Search | Korea Science

Building an RST-tagged Corpus and its Classification Scheme for Korean News Texts (한국어 수사구조 분류체계 수립 및 주석 코퍼스 구축)

Noh, Eunchung;Lee, Yeonsoo;Kim, YeonWoo;Lee, Do-Gil
- 한국어정보학회:학술대회논문집
- /
- 2016.10a
- /
- pp.33-38
- /
- 2016
수사구조는 텍스트의 각 구성 성분이 맺고 있는 관계를 의미하며, 필자의 의도는 논리적인 구조를 통해서 독자에게 더 잘 전달될 수 있다. 따라서 독자의 인지적 효과를 극대화할 수 있도록 수사구조를 고려하여 단락과 문장 구조를 구성하는 것이 필요하다. 그럼에도 불구하고 지금까지 수사구조에 기초한 한국어 분류체계를 만들거나 주석 코퍼스를 설계하려는 시도가 없었다. 본 연구에서는 기존 수사구조 이론을 기반으로, 한국어 보도문 형식에 적합한 30개 유형의 분류체계를 정제하고 최소 담화 단위별로 태깅한 코퍼스를 구축하였다. 또한 구축한 코퍼스를 토대로 중심문장을 비롯한 문장 구조의 특징과 분포 비율, 신문기사의 장르적 특성 등을 살펴봄으로써 텍스트에서 응집성의 실현 양상과 구문상의 특징을 확인하였다. 본 연구는 한국어 담화 구문에 적합한 수사구조 분류체계를 설계하고 이를 이용한 주석 코퍼스를 최초로 구축하였다는 점에서 의의를 갖는다.
PDF

Design of Multimedia Retrieval System based on XML (XML기반 멀티미디어 검색시스템의 설계)

Yoon, Mi-Hee;Cho, Dong-Uk
- Proceedings of the Korea Information Processing Society Conference
- /
- 2003.05a
- /
- pp.59-62
- /
- 2003
컴퓨팅 기술의 발달 밍 보편화로 인해 사용자들의 멀티미디어에 대한 요구가 증가하였고, 이러한 요구를 만족시키기 위해서는 단순한 텍스트 형식의 데이터가 아닌 멀티미디어 데이터, 특히 비디오 데이터에 대한 저장, 관리, 검색하는 기능이 필수적이다. 본 논문에서는 비디오데이터에 대한 효율적인 의미검색을 위해 주석기반 검색뿐만 아니라 특징기반 검색을 지원한다. 특히 사용자가 원하는 객체나 장면의 유사성 검색이 가능하며, 장면의 검색 결과로 제시된 장면을 선택한 후 선택된 장면을 기반으로 사용자가 원하는 좀 더 정확한 장면의 검색을 위한 SQBE(scene-query-by-example) 질의가 가능한 XML 기반 멀티미디어 검색시스템을 제안한다.
PDF

A Multimedia Database System using Method of annotation-based retrieval (주석 기반 검색 기법을 이용한 멀티미디어 데이터베이스 시스템)

Cho, Kyung-Mo
- Proceedings of the KAIS Fall Conference
- /
- 2010.05a
- /
- pp.319-322
- /
- 2010
본 논문에서는 특징기반 검색을 이용하여 대용량의 비디오 데이터에 대한 사용자의 다양한 의미검색을 지원하는 에이전트 기반에서의 자동화되고 통합된 비디오 의미기반 검색 시스템을 제안한다. 사용자의 기본적인 질의를 분석하고 질의에 의해 추출된 키 프레임의 이미지를 사용자가 선택함으로써 인덱싱 에이전트는 추출된 키 프레임의 주석에 대한 의미를 더욱 구체화시킨다. 또한, 사용자에 의해 선택된 키 프레임은 특징기반 검색의 질의 이미지가 되고 인덱싱 에이전트는 제안하는 다중 분할 칼라 히스토그램 기법을 통해 질의 이미지와 데이터베이스의 키 프레임들을 비교한 후 가장 유사한 키 프레임 이미지를 검색하여 사용자에게 디스플레이한다. 제안하여 구현된 시스템은 현저히 향상된 성능을 보였다.
PDF

Concept based Image Retrieval Using Similarity Measurement Between Concepts (개념간 유사성 측정을 이용한 개념 기반 이미지 검색)

조미영;최춘호;신주현;김판구
- Proceedings of the Korean Information Science Society Conference
- /
- 2003.04c
- /
- pp.253-255
- /
- 2003
기존의 개념 기반 이미지 검색에서는 이미지의 의미적 내용 인식을 위해 일반적으로 어휘적 정보나 텍스트 정보를 이용했다. 이러한 텍스트 정보 기반 이미지 검색은 전통적인 검색 방법인 키워드 검색 기술을 그대로 사용하여 쉽게 구현할 수 있으나 텍스트의 개념적 매칭이 아닌 스트링 매칭이므로 주석처리된 단어와 정확한 매칭이 없다면 찾을 수가 없었다. 이에 본 논문에서는 ontology의 일종인 WordNet을 이용하여 깊이 정보량 링크 타입, 밀도 등을 고려한 개념간 유사성 측정으로 패턴 매칭의 문제를 해결하고자 했다. 또한 키워드로 주석처리 되어 있는 Microsofts Design Gallery Live의 이미지를 이용하여 개념간 유사성 측정법을 실질적으로 개념 기반 이미지 검색에 적용해 보았다.
PDF

A Study on Intelligent Video Retrieval System based on query relaxation (질의완화를 기반으로 한 지능적인 비디오 검색 시스템)

Yoon, Mi-Hee;Cho, Dong-Uk
- Proceedings of the Korea Information Processing Society Conference
- /
- 2001.04b
- /
- pp.941-944
- /
- 2001
최근 하드웨어와 압축기술의 발달 및 보편화로 인해 사용자들의 비디오 데이터에 대한 요구가 증가하였다. 비디오 데이터는 비정형, 대용량의 특징을 가지고 있으므로 사용자의 다양한 요구를 만족시키기 위해서는 단순한 텍스트 형식의 데이터가 아닌 비디오 데이터에 대한 다양한 검색기법이 요구된다. 효율적인 비디오의 검색을 위해서는 사용자의 불완전한 질의에도 근사한 질의결과의 제시가 필요하다. 본 논문에서는 비디오데이터에 대한 효율적인 의미검색을 위해 주석기반과 특징기반을 혼합한 내용기반 검색을 지원하며 특히 사용자의 불완전한 질의에도 근접한 질의결과를 제시할 수 있는 지능적인 비디오 검색 시스템을 제안한다.
PDF

Applying Method WordNet for Concept based Image Retrieval system (개념 기반 이미지 검색 시스템을 위한 WordNet 적용 방안)

조미영;최준호;김판구
- Proceedings of the Korean Information Science Society Conference
- /
- 2002.10d
- /
- pp.487-489
- /
- 2002
기존의 키워드 기반 이미지 검색에서는 의미적 내용 인식을 위해 일반적으로 어휘적 정보나 텍스트 정보를 인간이 주석 형태로 달아주었다. 그러나 이런 텍스트 정보 기반 이미지 검색은 개념적 매칭이 아닌 스트링 매칭이므로 주석을 달아놓은 단어와 정확한 매칭이 없다면 찾을 수가 없다. 이러한 문제를 해결하기 위해 본 논문에서는 개념 기반 이미지 검색 시스템을 위한 WordNet의 적용 방안에 대해 연구했다. WordNet은 단언형이 아닌 단어의 의미 즉 synset이 구성 요소라는 특징을 이용해 각각의 이미지에 텍스트 정보 대신 적합한 개념의 Synset번호를 저장한다. 그리고 검색시 개념간의 유사성 측정을 이용해 검색어와 개념적으로 유사한 모든 이미지를 검색하도록 한다.
PDF

A WWW Images Automatic Annotation Based On Multi-cues Integration (멀티-큐 통합을 기반으로 WWW 영상의 자동 주석)

Shin, Seong-Yoon;Moon, Hyung-Yoon;Rhee, Yang-Won
- Journal of the Korea Society of Computer and Information
- /
- v.13 no.4
- /
- pp.79-86
- /
- 2008
As the rapid development of the Internet, the embedded images in HTML web pages nowadays become predominant. For its amazing function in describing the content and attracting attention, images become substantially important in web pages. All these images consist a considerable database. What's more, the semantic meanings of images are well presented by the surrounding text and links. But only a small minority of these images have precise assigned keyphrases. and manually assigning keyphrases to existing images is very laborious. Therefore it is highly desirable to automate the keyphrases extraction process. In this paper, we first introduce WWW image annotation methods, based on low level features, page tags, overall word frequency and local word frequency. Then we put forward our method of multi-cues integration image annotation. Also, show multi-cue image annotation method is more superior than other method through an experiment.
PDF

Development of Multimedia Annotation and Retrieval System using MPEG-7 based Semantic Metadata Model (MPEG-7 기반 의미적 메타데이터 모델을 이용한 멀티미디어 주석 및 검색 시스템의 개발)

An, Hyoung-Geun;Koh, Jae-Jin
- The KIPS Transactions:PartD
- /
- v.14D no.6
- /
- pp.573-584
- /
- 2007
As multimedia information recently increases fast, various types of retrieval of multimedia data are becoming issues of great importance. For the efficient multimedia data processing, semantics based retrieval techniques are required that can extract the meaning contents of multimedia data. Existing retrieval methods of multimedia data are annotation-based retrieval, feature-based retrieval and annotation and feature integration based retrieval. These systems take annotator a lot of efforts and time and we should perform complicated calculation for feature extraction. In addition. created data have shortcomings that we should go through static search that do not change. Also, user-friendly and semantic searching techniques are not supported. This paper proposes to develop S-MARS(Semantic Metadata-based Multimedia Annotation and Retrieval System) which can represent and extract multimedia data efficiently using MPEG-7. The system provides a graphical user interface for annotating, searching, and browsing multimedia data. It is implemented on the basis of the semantic metadata model to represent multimedia information. The semantic metadata about multimedia data is organized on the basis of multimedia description schema using XML schema that basically comply with the MPEG-7 standard. In conclusion. the proposed scheme can be easily implemented on any multimedia platforms supporting XML technology. It can be utilized to enable efficient semantic metadata sharing between systems, and it will contribute to improving the retrieval correctness and the user's satisfaction on embedding based multimedia retrieval algorithm method.
https://doi.org/10.3745/KIPSTD.2007.14-D.6.573 인용 PDF KSCI

A Named Entity Recognition Platform Based on Semi-Automatically Built NE-annotated Corpora and KoBERT (반자동구축된 개체명 주석코퍼스 DecoNAC과 KoBERT를 이용한 개체명인식 플랫폼 DecoNERO)

Kim, Shin-Woo;Hwang, Chang-Hoe;Yoon, Jeong-Woo;Lee, Seong-Hyeon;Choi, Soo-Won;Nam, Jee-Sun
- Annual Conference on Human and Language Technology
- /
- 2020.10a
- /
- pp.304-309
- /
- 2020
본 연구에서는 한국어 전자사전 DECO(Dictionnaire Electronique du COreen)와 다단어(Multi-Word Expressions: MWE) 개체명을 부분 패턴으로 기술하는 부분문법그래프(Local-Grammar Graph: LGG) 프레임에 기반하여 반자동으로 개체명주석 코퍼스 DecoNAC을 구축한 후, 이를 개체명 분석에 활용하고 또한 기계학습에 필요한 도메인별 학습 데이터로 활용하는 DecoNERO 개체명인식 플랫폼을 소개하는 데에 목적을 두었다. 최근 들어 좋은 성과를 보이는 것으로 보고되고 있는 기계학습 방법론들은 다양한 도메인을 기반으로한 대규모의 학습데이터를 필요로 한다. 본 연구에서는 정교하게 설계된 개체명 사전과 다단어 개체명 시퀀스에 대한 언어자원을 바탕으로 하는 반자동으로 학습데이터를 생성하는 방법론을 제안하였다. 본 연구에서 제안된 개체명주석 코퍼스 DecoNAC 기반 접근법의 성능을 실험하기 위해 온라인 뉴스 기사 텍스트를 바탕으로 실험을 진행하였다. 이 실험에서 DecoNAC을 적용한 경우, KoBERT 모델만으로 개체명을 인식한 결과에 비해 약 7.49%의 성능향상을 기대할 수 있음을 확인하였다.
PDF

Content based data search using semantic annotation (시맨틱 주석을 이용한 내용 기반 데이터 검색)

Kim, Byung-Gon;Oh, Sung-Kyun
- Journal of Digital Contents Society
- /
- v.12 no.4
- /
- pp.429-436
- /
- 2011
Various documents, images, videos and other materials on the web has been increasing rapidly. Efficient search of those things has become an important topic. From keyword-based search, internet search has been transformed to semantic search which finds the implications and the relations between data elements. Many annotation processing systems manipulating the metadata for semantic search have been proposed. However, annotation data generated by different methods and forms are difficult to process integrated search between those systems. In this study, in order to resolve this problem, we categorized levels of many annotation documents, and we proposed the method to measure the similarity between the annotation documents. Similarity measure between annotation documents can be used for searching similar or related documents, images, and videos regardless of the forms of the source data.
https://doi.org/10.9728/dcs.2011.12.4.429 인용 PDF KSCI

Search Result 331, Processing Time 0.025 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)