통합 검색 | Korea Science

Using A Semantic Classification in Parsing Chinese: Some Preliminary Results

Gan, Kok-Wee
- 한국언어정보학회:학술대회논문집
- /
- 한국언어정보학회 1998년도 Language, Information and Computation = Selected Papers from the 12th Pacific Asia Conference on Language, Information and Computation, Singapore
- /
- pp.340-347
- /
- 1998
PDF

Using Syntax and Shallow Semantic Analysis for Vietnamese Question Generation

Phuoc Tran;Duy Khanh Nguyen;Tram Tran;Bay Vo
- KSII Transactions on Internet and Information Systems (TIIS)
- /
- 제17권10호
- /
- pp.2718-2731
- /
- 2023
This paper presents a method of using syntax and shallow semantic analysis for Vietnamese question generation (QG). Specifically, our proposed technique concentrates on investigating both the syntactic and shallow semantic structure of each sentence. The main goal of our method is to generate questions from a single sentence. These generated questions are known as factoid questions which require short, fact-based answers. In general, syntax-based analysis is one of the most popular approaches within the QG field, but it requires linguistic expert knowledge as well as a deep understanding of syntax rules in the Vietnamese language. It is thus considered a high-cost and inefficient solution due to the requirement of significant human effort to achieve qualified syntax rules. To deal with this problem, we collected the syntax rules in Vietnamese from a Vietnamese language textbook. Moreover, we also used different natural language processing (NLP) techniques to analyze Vietnamese shallow syntax and semantics for the QG task. These techniques include: sentence segmentation, word segmentation, part of speech, chunking, dependency parsing, and named entity recognition. We used human evaluation to assess the credibility of our model, which means we manually generated questions from the corpus, and then compared them with the generated questions. The empirical evidence demonstrates that our proposed technique has significant performance, in which the generated questions are very similar to those which are created by humans.
https://doi.org/10.3837/tiis.2023.10.007 인용 PDF HTML

뉴스 동영상 자동 의미 분석 알고리즘 (An Automatic News Video Semantic Parsing Algorithm)

전승철;박성한
- 대한전자공학회:학술대회논문집
- /
- 대한전자공학회 2001년도 하계종합학술대회 논문집(3)
- /
- pp.109-112
- /
- 2001
This paper proposes an efficient algorithm of extracting anchor blocks for a semantic structure of a news video. We define the FRFD to calculate the frame difference of anchor face position rather than simply uses the general frame difference. Since, The FRFD value is sensitive to existing face in frame, anchor block can be efficiently extracted. In this paper, an algorithm to extract a face position using partial decoded MPEG data is also proposed. In this way a news video can be structured semantically using the extracted anchor blocks.
PDF

분산 메모리 다중프로세서 환경에서의 병렬 음성인식 모델 (A Parallel Speech Recognition Model on Distributed Memory Multiprocessors)

정상화;김형순;박민욱;황병한
- 한국음향학회지
- /
- 제18권5호
- /
- pp.44-51
- /
- 1999
본 논문에서는 음성과 자연언어의 통합처리를 위한 효과적인 병렬계산모델을 제안한다. 음소모델은 연속 Hidden Markov Model(HMM)에 기반을 둔 문맥종속형 음소를 사용하며, 언어모델은 지식베이스를 기반으로 한다. 또한 지식베이스를 구성하기 위해 계층구조의 semantic network과 병렬 marker-passing을 추론 메카니즘으로 쓰는 memory-based parsing 기술을 사용한다. 본 연구의 병렬 음성인식 알고리즘은 분산메모리 MIMD(Multiple Instruction Multiple Data) 구조의 다중 Transputer 시스템을 이용하여 구현되었다. 실험결과, 본 연구의 지식베이스 기반 음성인식 시스템의 인식률이 word network 기반 음성인식 시스템보다 높게 나타났으며 code-phoneme 통계정보를 활용하여 인식성능의 향상도 얻을 수 있었다. 또한, 성능향상도(speedup) 관련 실험들을 통하여 병렬 음성인식 시스템의 실시간 구현 가능성을 확인하였다.
PDF

개방형 다중 데이터셋을 활용한 Combined Segmentation Network 기반 드론 영상의 의미론적 분할 (Semantic Segmentation of Drone Images Based on Combined Segmentation Network Using Multiple Open Datasets)

송아람
- 대한원격탐사학회지
- /
- 제39권5_3호
- /
- pp.967-978
- /
- 2023
본 연구에서는 다양한 드론 영상 데이터셋을 효과적으로 학습하여 의미론적 분할의 정확도를 향상시키기 위한 combined segmentation network (CSN)를 제안하고 검증하였다. CSN은 세 가지 드론 데이터셋의 다양성을 고려하기 위하여 인코딩 영역의 전체를 공유하며, 디코딩 영역은 독립적으로 학습된다. CSN의 경우, 학습 시 모든 데이터셋에 대한 손실값을 고려하기 때문에 U-Net 및 pyramid scene parsing network (PSPNet)으로 단일 데이터셋을 학습할 때보다 학습 효율이 떨어졌다. 그러나 국내 자율주행 드론 영상에 CSN을 적용한 결과, CSN이 PSPNet에 비해 초기 학습 없이도 영상 내 화소를 적절한 클래스로 분류할 수 있는 것을 확인하였다. 본 연구를 통하여 CSN이 다양한 드론 영상 데이터셋을 효과적으로 학습하고 새로운 지역에 대한 객체 인식 정확성을 향상시키는 데 중요한 도구로써 활용될 수 있을 것으로 기대할 수 있다.
https://doi.org/10.7780/kjrs.2023.39.5.3.7 인용 PDF HTML

Dependency Grammar and the Parsing of Chinese Sentences

Lai, Bong-Ycung-Tom;Huang, Changning
- 한국언어정보학회:학술대회논문집
- /
- 한국언어정보학회 1994년도 The Proceedings of the 1994 Kyoto Conference
- /
- pp.63-72
- /
- 1994
Dependency Grammar has been used by Iinguists as the basis of the syntactic components of their grammar formalisms. It has also been used in natural langauge parsing. In China, attempts have been made to use this grammar formalism to parse Chinese sentences using corpus based techniques. This paper reviews the properties of Dependency Grammar as embodied in four axioms for the well-formedness conditions for dependency structures. It is shown that allowing mul tiple governors as done by some followers of this formalism is unnecessary. The practice of augmenting Dependency Grammar with functional labels is discussed in the light of building functional structures when the sentence is parsed. This will also facilitate semantic interpretion.retion.
PDF

Parsing the Wh-Interrogative Construction in Korean

Yang, Jaehyung;Kim, Jong-Bok
- 한국언어정보학회지:언어와정보
- /
- 제17권2호
- /
- pp.51-66
- /
- 2013
Korean is a wh-in-situ language where the wh-expression stays in situ with an obligatory Q-particle marking its interrogative scope. This paper briefly reviews some basic properties of the wh-question construction in Korean and shows how a typed feature structure grammar, HPSG (Pollard and Sag 1994, Sag et al. 2003), together with the notions of 'type hierarchy' and 'constructions', can provide a robust basis for parsing the wh-construction in the language. We show that this system induces robust syntactic structures as well as enriched semantic representations for real-time applications such as machine translation, which require deep processing of the phenomena concerned.
PDF

요약파싱기법을 사용한 웹 접근성의 정적 분석 (Static Analysis of Web Accessibility Based on Abstract Parsing)

김현하;도경구
- 정보과학회 논문지
- /
- 제41권12호
- /
- pp.1099-1109
- /
- 2014
웹 접근성 평가 도구는 웹 사이트가 웹 접근성 지침을 잘 지키고 있는지 검사하는 도구이다. 국내외 법과 제도가 마련된 이후 지침 준수여부를 검사하는 도구가 많이 나왔지만, 대부분 동적으로 페이지를 수집해서 분석하는 방법을 사용한다. 특히 자동화된 도구들은 페이지를 수집한 후에 분석하는데, 실행환경이나 접근권한의 문제로 수집하지 못해서 분석결과에서 빠지는 경우가 발생할 수 있다. 본 연구는 기존 방법과 달리 정적으로 분석하여 웹 접근성을 평가하는 방법을 제안한다. 정적인 분석방법은 실행 가능한 모든 경로를 고려하기 때문에 놓치는 페이지 없이 분석할 수 있다. 요약해석기법에 파싱이론을 접목한 요약파싱 기술을 사용해서 동적으로 생성될 웹 페이지의 웹 접근성을 정적으로 분석하는 도구를 개발하였다. 실험 대상 PHP 프로그램을 제안하는 연구방법으로 개발한 도구와 비교 대상 도구에서 분석한 결과를 비교해서 비교 대상 도구에서는 접근권한이나 실행경로 등의 문제로 분석하지 못하고 놓치는 웹 페이지가 있음을 확인하였다.
https://doi.org/10.5626/JOK.2014.41.12.1099 인용

효율적인 한국어 파싱을 위한 최장일치 기반의 형태소 분석기 기능 확장 (Functional Expansion of Morphological Analyzer Based on Longest Phrase Matching For Efficient Korean Parsing)

이현영;이종석;강병도;양승원
- 디지털콘텐츠학회 논문지
- /
- 제17권3호
- /
- pp.203-210
- /
- 2016
한국어는 문장 구성소의 생략과 수식 범위가 자유롭기 때문에 파싱보다는 형태소 분석 단계에서 처리하면 좋은 경우가 있다. 본 논문에서는 파싱의 부담을 덜어 줄 수 있는 형태소 분석기의 기능 확장 방안을 제안한다. 이 방법은 미지어의 추정, 복합 명사 및 복합동사의 처리, 숫자 및 심볼의 처리에 의해 여러 형태소 열이 하나의 구문 범주를 가질 때 이것을 최장일치 방법으로 결합하고 의미 자질을 부여하여 하나의 구문 단위로 처리하는 것이다. 제안한 형태소 분석 방법은 불필요한 형태론적 모호성이 제거되고 형태소 분석 결과가 줄어들어 태거 및 파서의 정확률이 향상되었다. 또한, 실험을 통해 파싱트리는 평균 73.4%, 파싱 시간은 평균 52.9%로 줄었음을 보인다.
https://doi.org/10.9728/dcs.2016.17.3.203 인용 PDF KSCI

웹기반 언어 학습시스템을 위한 한국어 철자/문법 검사기의 성능 향상 (Improving a Korean Spell/Grammar Checker for the Web-Based Language Learning System)

남현숙;김광영;권혁철
- 인지과학
- /
- 제12권3호
- /
- pp.1-18
- /
- 2001
이 논문의 목적은 한국어 철자/문법 검사기를 교육적으로 활용한 웹 기반 국어 작문 학습 시스템의 구현이다. 웹 기반 학습시스템 \\`우리말 배움터\\`의 학습효과를 최대화하려면 한국어 철자/문법 검사기의 성능을 꾸준히 향상해야 한다 오늘날 자연어처리 시스템의 성능은 의미처리를 얼마나 정확하게 수행하는가에 달려있다 한국어 철자/문법 검사기에서 의미처리와 관련이 있는 부분은 철자 검사기에서 접사나 꼬리말과 파생하는 단어와 복합명사를 교정하는 처리기와 의미·문체 오류를 교정하는 문법 검사기이다. 본 시스템에서는 의미처리를 위하여 의존문법에 기반하여 부분문장분석과 연어관계정보를 이용한다. 여기에 더 세부적인 규칙을 추가하기 위해 단어를 개념적으로 분류하고 문장의 핵심요소인 동사를 하위범주화한 결과를 적용한다. 의미처리 기능을 강화한 철자/문법 검사기를 온라인으로 운영함으로써 웹에 기반한 한국어 학습시tm템과 통합된 환경에서 능동적이고 지능적인 학습 모형을 구현한다. 이 논문에서 다루는 의미처리의 대상은 주로 구문 단위이기 때문에 여러 개의 절이 모여 하나의 문장이 된 복문이나 중문은 다루지 못하고 있다. 또한 일률적인 체계 속에서 단어를 의미적으로 분류하는 데에도 많은 한계가 있다. 한편 이러한 자연어처리시스템을 웹 기반 학습시스템에 연결하여 효율적인 학습효과를 거두려면 학습내용 구성이나 인터페이스 설계 면에서도 고려해야 할 중요한 문제가 많다. 결론에서는 아직 완전하게 해결하지 못한 문제에 대해 고찰한다.
PDF

검색결과 62건 처리시간 0.022초

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

자세히 찾기

이미지 검색 (β)