• Title/Summary/Keyword: Entity Resolution

Search Result 39, Processing Time 0.02 seconds

Author Entity Identification using Representative Properties in Linked Data (대표 속성을 이용한 저자 개체 식별)

  • Kim, Tae-Hong;Jung, Han-Min;Sung, Won-Kyung;Kim, Pyung
    • The Journal of the Korea Contents Association
    • /
    • v.12 no.1
    • /
    • pp.17-29
    • /
    • 2012
  • In recent years, Linked Data that is published under an open license shows increased growth rate and comes into the spotlight due to its interoperability and openness especially in government of developed countries. However there are relatively few out-links compared with its entire number of links and most of links refer a few hub dataset. These occur because of absence of technology that identifies entities in Linked data. In this paper, we present an improved author entity resolution method that using representative properties. To solve problems of previous methods that utilizes relation with other entities(owl:sameAs, owl:differentFrom and so on) or depends on Curation, we design and evaluate an automated realtime resolution process based on multi-ontologies that respects entity's type and its logical characteristics so as to verify entities consistency. The evaluation of author entity resolution shows positive results (The average of K measuring result is 0.8533.) with 29 author information that has obtained confirmation.

Simple and effective neural coreference resolution for Korean language

  • Park, Cheoneum;Lim, Joonho;Ryu, Jihee;Kim, Hyunki;Lee, Changki
    • ETRI Journal
    • /
    • v.43 no.6
    • /
    • pp.1038-1048
    • /
    • 2021
  • We propose an end-to-end neural coreference resolution for the Korean language that uses an attention mechanism to point to the same entity. Because Korean is a head-final language, we focused on a method that uses a pointer network based on the head. The key idea is to consider all nouns in the document as candidates based on the head-final characteristics of the Korean language and learn distributions over the referenced entity positions for each noun. Given the recent success of applications using bidirectional encoder representation from transformer (BERT) in natural language-processing tasks, we employed BERT in the proposed model to create word representations based on contextual information. The experimental results indicated that the proposed model achieved state-of-the-art performance in Korean language coreference resolution.

A Muti-Resolution Approach to Restaurant Named Entity Recognition in Korean Web

  • Kang, Bo-Yeong;Kim, Dae-Won
    • International Journal of Fuzzy Logic and Intelligent Systems
    • /
    • v.12 no.4
    • /
    • pp.277-284
    • /
    • 2012
  • Named entity recognition (NER) technique can play a crucial role in extracting information from the web. While NER systems with relatively high performances have been developed based on careful manipulation of terms with a statistical model, term mismatches often degrade the performance of such systems because the strings of all the candidate entities are not known a priori. Despite the importance of lexical-level term mismatches for NER systems, however, most NER approaches developed to date utilize only the term string itself and simple term-level features, and do not exploit the semantic features of terms which can handle the variations of terms effectively. As a solution to this problem, here we propose to match the semantic concepts of term units in restaurant named entities (NEs), where these units are automatically generated from multiple resolutions of a semantic tree. As a test experiment, we applied our restaurant NER scheme to 49,153 nouns in Korean restaurant web pages. Our scheme achieved an average accuracy of 87.89% when applied to test data, which was considerably better than the 78.70% accuracy obtained using the baseline system.

A Methodology for View Integration Using ERD Thesaurus (ERD시소러스를 이용한 뷰 통합 방법론)

  • Lee, Won-Jo;Koh, Jae-Jin;Jang, Gil-Sang
    • The KIPS Transactions:PartD
    • /
    • v.11D no.3
    • /
    • pp.553-562
    • /
    • 2004
  • This paper constructs ERD thesaurus that is storing information about Entity Relationship Diagram(ERD), and proposes an ERD thesaurus-based methodology for view integration in an important conceptual design step in designing databases. To show the usefulness of proposed methodology, the prototype for view integration support system is implemented for the applied case. As a result, ERD thesaurus-based methodology is more effective than the existing methodologies for view Integration in the aspects of affinity analysis, semantic conflicts resolution, and view Integration processes. Therefore, our methodology is expected to be utilized in integrating the existing fragmented schema or designing a large database integration.

Coreference Resolution for Korean Pronouns using Pointer Networks (포인터 네트워크를 이용한 한국어 대명사 상호참조해결)

  • Park, Cheoneum;Lee, Changki
    • Journal of KIISE
    • /
    • v.44 no.5
    • /
    • pp.496-502
    • /
    • 2017
  • Pointer Networks is a deep-learning model for the attention-mechanism outputting of a list of elements that corresponds to the input sequence and is based on a recurrent neural network (RNN). The coreference resolution for pronouns is the natural language processing (NLP) task that defines a single entity to find the antecedents that correspond to the pronouns in a document. In this paper, a pronoun coreference-resolution method that finds the relation between the antecedents and the pronouns using the Pointer Networks is proposed; furthermore, the input methods of the Pointer Networks-that is, the chaining order between the words in an entity-are proposed. From among the methods that are proposed in this paper, the chaining order Coref2 showed the best performance with an F1 of MUC 81.40 %. The method showed performances that are 31.00 % and 19.28 % better than the rule-based (50.40 %) and statistics-based (62.12 %) coreference resolution systems, respectively, for the Korean pronouns.

Korean Coreference Resolution using the Multi-pass Sieve (Multi-pass Sieve를 이용한 한국어 상호참조해결)

  • Park, Cheon-Eum;Choi, Kyoung-Ho;Lee, Changki
    • Journal of KIISE
    • /
    • v.41 no.11
    • /
    • pp.992-1005
    • /
    • 2014
  • Coreference resolution finds all expressions that refer to the same entity in a document. Coreference resolution is important for information extraction, document classification, document summary, and question answering system. In this paper, we adapt Stanford's Multi-pass sieve system, the one of the best model of rule based coreference resolution to Korean. In this paper, all noun phrases are considered to mentions. Also, unlike Stanford's Multi-pass sieve system, the dependency parse tree is used for mention extraction, a Korean acronym list is built 'dynamically'. In addition, we propose a method that calculates weights by applying transitive properties of centers of the centering theory when refer Korean pronoun. The experiments show that our system obtains MUC 59.0%, $B_3$ 59.5%, Ceafe 63.5%, and CoNLL(Mean) 60.7%.

Multi-task learning with contextual hierarchical attention for Korean coreference resolution

  • Cheoneum Park
    • ETRI Journal
    • /
    • v.45 no.1
    • /
    • pp.93-104
    • /
    • 2023
  • Coreference resolution is a task in discourse analysis that links several headwords used in any document object. We suggest pointer networks-based coreference resolution for Korean using multi-task learning (MTL) with an attention mechanism for a hierarchical structure. As Korean is a head-final language, the head can easily be found. Our model learns the distribution by referring to the same entity position and utilizes a pointer network to conduct coreference resolution depending on the input headword. As the input is a document, the input sequence is very long. Thus, the core idea is to learn the word- and sentence-level distributions in parallel with MTL, while using a shared representation to address the long sequence problem. The suggested technique is used to generate word representations for Korean based on contextual information using pre-trained language models for Korean. In the same experimental conditions, our model performed roughly 1.8% better on CoNLL F1 than previous research without hierarchical structure.

An Introduction to the Study of the Outlook on Highest Ruling Entity in Daesoonjinrohoe (I) - Focusing on Descriptions for Highest Ruling Entity and It's Meanings - (대순진리회 상제관 연구 서설 (I) - 최고신에 대한 표현들과 그 의미들을 중심으로 -)

  • Cha, Seon-keun
    • Journal of the Daesoon Academy of Sciences
    • /
    • v.21
    • /
    • pp.99-156
    • /
    • 2013
  • This paper is to indicate research tendencies of faith in Daesoonjinrihoe and controversial points of those, and to consider the outlook on Sangje after defining it as theological understanding and explanation for Gu-Cheon-Sang-Je (High-est ruling Entity that is the object of devotion in Daesoon-jinrihoe). As the first introduction to the work, various descriptions for Sangje are arranged and the meanings of those are analyzed. In brief, first, the name of Gu-Cheon-Eung-Won-Nweh-Seong-Bo-Hwa-Cheon-Jon, expresses the fact that the authority of Sangje (the Supreme Entity) is exposed by spatial concept Sangje dwells in Ninth Heaven. This fact can be compared with the doctrines Allah in Islam and Jehovah in Christianity each are dwelled in Seventh Heaven. And the name shows Sangje is the ruler who reigns over the universe by using yin and yang. Second, the name, Gu-Cheon-Eung-Won-Nweh-Seong-BoHwa-Cheon-Jon, is imported from China Taoism because it has been in Ok-Chu-Gyeong (the Gaoshang shenlei yushu). But in fact it's root is in Korea because Buyeo and Goguryeo, the ancient Korean nations, have the source of the name. While the name is not the Supreme Entity in China Taoism, it is the Supreme Entity in Daesoonjinrihoe. This fact is a important difference. Third, arbitrarily or not, the name, Gu-Cheon-Eung-Won-Nweh-Seong-Bo-Hwa-Cheon-Jon, is put on the image of 'resolution of grievances'. The reason is that many peoples in Korea and China has called the name for about 1,000 years ago to help their fortunes and escape predicaments. Forth, not only Gu-Cheon-Eung-Won-Nweh-Seong-Bo-Hwa-Cheon-Jon but also the name, Three Pure Ones and Ok-Cheon-Jin-Wang (Yuqingzhenwang) in China Taoism used as the Highest ruling Entity in Daesoonjinrihoe. But the relations between three Pure Ones and Ok-Cheon-Jin-Wang and Gu-Cheon-Eung-Won-Nweh-Seong-Bo-Hwa-Cheon-Jon in Dae-soonjinrihoe are different from that in China Taoism. Fifth, Sangje is associated with the Polaris divinity of Tae-Eul, view on God in Oriental Cosmology. The description Tae-Eul as well as Gu-Cheon-Eung-Won-Nweh-Seong-Bo-Hwa-Cheon-Jon is indicated Sangje is linked to the faith of Buyeo and Goguryeo. Sixth, Sangje is not only Mugeuk-Sin (The God of The Endless) who supervise the Endless but also Taegeuk-Ji-Cheon-Jon (The God of The Ultimate Reality) who supervise the Ultimate Reality. These descriptions directly display the fact Sangje is a creator. Seventh, in case explaining Sangje, the point of view is necessary that grasps the whole viewpoints Sangje 'was' Hidden God(deus otiosus) and 'is' Unhidden God after Incarnation. Eighth, Sangje is Cheon-Ju in Donghak, but different from that. Cheon-Ju in Donghak has both transcendence and immanence in tightrope tension, but Cheon-Ju in Daesoonjinrihoe emphasize transcendence than immanence. That difference is the result of the fact Cheon-Ju in Donghak was a being having revealed a man and Cheon-Ju in Daesoonjinrihoe was a being having incarnated after revealing a man. Ninth, Sangje is Gae-Byeok-Jang who is the manager of the transforming and ordering the Three Realms of the World by the Great Do which is the mutual beneficence of all life and Hae-Won-Sin who is the God of resolution of grievances.

Spontaneous Spinal Subarachnoid Hemorrhage with Spontaneous Resolution

  • Kim, Jin-Sung;Lee, Sang-Ho
    • Journal of Korean Neurosurgical Society
    • /
    • v.45 no.4
    • /
    • pp.253-255
    • /
    • 2009
  • Spontaneous spinal subarachnoid hematoma (SSH) is a rare entity to cause spinal cord or nerve root compression and is usually managed as surgical emergencies. We report a case of spontaneous SSH manifesting as severe lumbago, which demonstrated nearly complete clinical resolution with conservative treatment A 58-year-old female patient developed a large SSH, which was not related to blood dyscrasia, anticoagulation, lumbar puncture. or trauma. Patient had severe lumbago but no neurologic deficits. Because of absence of neurologic deficits, she was treated conservatively. Follow-up magnetic resonance (MR) image showed complete resolution. Conservative treatment of SSH may be considered if the patient with spontaneous SSH has no neurologic deficits.

CR-M-SpanBERT: Multiple embedding-based DNN coreference resolution using self-attention SpanBERT

  • Joon-young Jung
    • ETRI Journal
    • /
    • v.46 no.1
    • /
    • pp.35-47
    • /
    • 2024
  • This study introduces CR-M-SpanBERT, a coreference resolution (CR) model that utilizes multiple embedding-based span bidirectional encoder representations from transformers, for antecedent recognition in natural language (NL) text. Information extraction studies aimed to extract knowledge from NL text autonomously and cost-effectively. However, the extracted information may not represent knowledge accurately owing to the presence of ambiguous entities. Therefore, we propose a CR model that identifies mentions referring to the same entity in NL text. In the case of CR, it is necessary to understand both the syntax and semantics of the NL text simultaneously. Therefore, multiple embeddings are generated for CR, which can include syntactic and semantic information for each word. We evaluate the effectiveness of CR-M-SpanBERT by comparing it to a model that uses SpanBERT as the language model in CR studies. The results demonstrate that our proposed deep neural network model achieves high-recognition accuracy for extracting antecedents from NL text. Additionally, it requires fewer epochs to achieve an average F1 accuracy greater than 75% compared with the conventional SpanBERT approach.