Search | Korea Science

A Segmentation Method of Compound Nouns Using Syllable Preference (선호 음절 정보를 이용한 복합명사의 분해 방법)

Park Chan-Ee;Ryu Bang;Kim Sang-Bok
- Journal of Korea Multimedia Society
- /
- v.9 no.2
- /
- pp.151-159
- /
- 2006
The ratio of a segmentation algorithm of compound nouns causes an effect a lot in nouns which are not in the dictionary. The structure of Korean compound nouns are mostly derived from the Chinese characters and it includes some preference ratio. So it will be able to use segmentation rule of compound nouns. This paper suggests a segmentation algorithm using some preference ratio of Korean compound nouns which are not in the dictionary. The experiment resulted in getting 88.49% of correct segmentation and showed effective result from the comparative experimentation with other algorithm.
PDF

Korean Word Search App Using Meta-characters (메타문자를 사용한 한국어 사전 탐색 앱)

Kwon, Hong-Seok;Kim, Jae-Hoon
- Annual Conference on Human and Language Technology
- /
- 2011.10a
- /
- pp.110-113
- /
- 2011
스마트 폰의 보급이 대중화됨에 따라 다양한 앱들이 사용되고 있으나 효율적인 사전 탐색에 관한 앱은 그다지 많지 않다. 현재 공개된 한국어 사전 탐색 앱은 완전한 단어이거나 단어의 부분 문자열을 질의로 사용한다. 이 경우 완전한 단어를 기억하지 못하거나 한국어 정보처리를 위한 여러 형태의 음운 정보를 쉽게 탐색할 수 없다. 이러한 문제를 개선하기 위해 본 논문에서는 메타문자를 사용하여 효율적으로 단어를 탐색할 수 있는 앱을 개발한다. 본 논문에서 사용하는 메타문자는 임의의 음절을 표현하는 '*'와 '?'과 종성을 표현하는 ':'를 사용하며 사전구조는 자소 단위의 트라이를 사용한다. 또한 음절은 물론이고 자소(초성, 중성, 종성)로 구성된 질의를 탐색할 수 있다. 더구나 음절과 자소가 혼합된 질의도 사용할 수 있도록 하여 사용자의 편의를 크게 도모하였다.
PDF

A perceptual study on the correlation between the meaning of Korean polysemic ending and its boundary tone (동형다의 종결어미의 의미와 경계성조의 상관성에 대한 지각연구)

Youngsook Yune
- Phonetics and Speech Sciences
- /
- v.14 no.4
- /
- pp.1-10
- /
- 2022
The Korean polysemic ending '-(eu)lgeol' can has two different meanings, 'guess' and 'regret'. These are expressed by different boundary-tone types: a rising tone for guess, a falling one for regret. Therefore the sentence-final boundary-tone type is the most salient prosodic feature. However, besides tone type, the pitch difference between the final and penultimate syllables of '-(eu)lgeol' can also affect semantic discrimination. To investigate this aspect, we conducted a perception test using two sentences that were morphologically and syntactically identical. These two sentences were spoken using different boundary-tone types by a Korean native speaker. From these two sentences, the experimental stimuli were generated by artificially raising or lowering the pitch of the boundary syllable by 1Qt while fixing the pitch of the penultimate syllable and boundary-tone type. Thirty Korean native speakers participated in three levels of perceptual test, in which they were asked to mark whether the experimental sentences they listened to were perceived as guess or regret. The results revealed that regardless of boundary-tone types, the larger the pitch difference between the final and penultimate syllable in the positive direction, the more likely it is perceived as guess, and the smaller the pitch difference in the negative direction, the more likely it is perceived as regret.
https://doi.org/10.13064/KSSS.2022.14.4.001 인용 PDF KSCI

The Layered Structural Tagging Program for Seaching (언어자료 검색을 위한 계층구조형 형태소 분석 프로그램)

Kang, Yong-Hee
- Annual Conference on Human and Language Technology
- /
- 2001.10d
- /
- pp.89-96
- /
- 2001
1999년 제1회 형태소 분석기 및 품사태거 평가 워크숍 이후 표준안에 대한 새로운 대안이나 문제제기등을 제시한 논문은 전무하다. 본 연구에서는 평가대회 참가 이후 표준안을 수정한 새로운 유형의 형태소 분석 프로그램을 제작하여 그 실용성과 앞으로의 발전 가능성과 문제점을 밝혀, 계층구조형의 형태소분석 시스템을 채택하고 있는 일본의 JUMAN을 참조 새로운 유형의 형태소 분석형식을 제시한다. 본 연구는 일본방송협회 방송기술연구소(이하 NHK기술 연구소)의 의뢰에 인한 것이며 어절단위의 표준안과 다른 형태소 단위를 기본요소로 삼고 있으며 활용형을 갖고 있는 용언에 대해서는 활용형의 전개를 하고 있다. 어절단위로 탈피한 이유는 형태소 분석의 기본요소로써 어절단위 보다는 형태소 단위를 기준으로 삼는 것이 생산성이 높다고 생각된다. 어절정보와 문장정보는 XML(extensible makrup language)등의 별도의 정보를 주는 방법을 채택했다. 음절말음이 자음인지 모음인지의 음운 정보에 따라 활용형을 차별했으며 표준안과 달리 명사의 종류와 개념을 세분화했다. 아울러 조사와 어미등의 검색어와 함께 음절을 형성하고 있는 비검색어 대상은 배제하는 프로그램과 표준안의 어절방식으로 출력하는 3가지 프로그램을 작성했다. 본 연구에서는 계층구조의 형태소분석 프로그램의 가능성과 한국어의 특성을 고려한 출력항목등을 고찰하는 것을 목적으로 한다.
PDF

A Reverse Segmentation Algorithm of Compound Nouns (복합명사의 역방향 분해 알고리즘)

Lee, Hyeon-Min;Park, Hyeok-Ro
- The KIPS Transactions:PartB
- /
- v.8B no.4
- /
- pp.357-364
- /
- 2001
본 논문에서는 단위명사 사전과 접사 사전을 이용하여 한국어 복합명사를 분해하는 새로운 알고리즘을 제안한다. 한국어 복합명사는 그 구조에 있어서 중심어가 뒤에 나타난다는 점에 착안하여 본 논문에서 제안한 분해 알고리즘은 복합명사를 끝음절에서 첫음절 방향 즉 역방향으로 분해를 시도한다. ETRI의 태깅된 코퍼스로부터 추출한 복합명사 3,230개에 대해 실험한 결과 약 96.6%의 분해 정확도를 얻었다. 미등록어를 포함한 복합명사의 경우는 77.5%의 분해 정확도를 나타냈다. 실험에 사용된 데이터중의 미등록어는 대부분 접사를 포함한 파행어로서, 제안한 복합명사 분해 알고리즘은 접사가 부착된 미등록어 분석에 있어서 보다 높은 분석 정확도를 나타냄을 알 수 있었다.
PDF

Segmental and prosodic environments and vowel devoicing in Korean (분절음적, 운율적 환경과 무성모음의 실현)

Shin Ji-Young;Chae Eun-Ae
- Proceedings of the Acoustical Society of Korea Conference
- /
- spring
- /
- pp.309-312
- /
- 2002
무성모음화 현상이 어떠한 분절음적, 운율적 환경에서 주로 실현되는가를 알아보기 위하여 선행자음의 분절음적 환경, 후행자음의 분절음적 환경, 해당 강세구의 음절수, 운율 구조상의 위치 등 모두 네 가지를 변수로 실험을 진행하였다. 모두 10명의 화자(남5, 여5)가 발화한 1140개의 자료에 나타난 행당 모음의 길이를 측정하는 방법으로 분석을 실시하였다. 그 결과 선행자음은 [+기식성]과 [+지속성]을 가진 환경이, 후행 자음은 [-지속정]과 [기식성]을 가진 환경이 무성모음화가 잘 일어나는 환경인 것으로 밝혀졌다. 음절수의 증가는 큰 영향을 주지 않는 것으로 보였고, 대체로 두 번째 강세구의 단어초에 위치하는 경우에 모음의 길이가 짧거나 무성모음화되는 경향이 관찰되었다.
PDF

Multi-head Attention and Pointer Network Based Syllables Dependency Parser (멀티헤드 어텐션과 포인터 네트워크 기반의 음절 단위 의존 구문 분석)

Kim, Hong-jin;Oh, Shin-hyeok;Kim, Dam-rin;Kim, Bo-eun;Kim, Hark-soo
- Annual Conference on Human and Language Technology
- /
- 2019.10a
- /
- pp.546-548
- /
- 2019
구문 분석은 문장을 구성하는 어절들 사이의 관계를 파악하여 문장의 구조를 이해하는 기술이다. 구문 분석은 구구조 분석과 의존 구문 분석으로 나누어진다. 한국어처럼 어순이 자유로운 언어에는 의존 구문 분석이 더 적합하다. 의존 구문 분석은 문장을 구성하고 있는 어절 간의 의존 관계를 분석하는 작업으로, 각 어절의 지배소를 찾아내어 의존 관계를 분석한다. 본 논문에서는 멀티헤드 어텐션과 포인터 네트워크를 이용한 음절 단위 의존 구문 분석기를 제안하며 UAS 92.16%, LAS 89.71%의 성능을 보였다.
PDF

Metrical Structure Change Phenomenon of K-Pop Songs : Focusing on Dance Music (K-Pop 노랫말의 운율구조 변화 현상 : 댄스음악을 중심으로)

Seo, Keun-Young
- Journal of Korea Entertainment Industry Association
- /
- v.14 no.7
- /
- pp.343-362
- /
- 2020
English is a stress-timed language that has a phonetic system in which the speech is restructured by stress changes. On the other hand, Korean is a syllable-timed language in which each syllable is pronounced at almost the same length and intensity, and Korean and English have distinctly different metrical systems in general speech. However, as the language of the lyrics in K-Pop music is mixed in both languages, Korean and English, the Korean lyrics in K-Pop music have a metrical system by stress changes as in English. The writer's view is that the change in the metrical structure of Korean lyrics is inevitable in order to sustain the new Korean Wave. Therefore, in this study, dance music - a major genre of K-Pop music that focuses on rhythm expression - is classified into 1998, 2003, and 2009 according to the changes in the Korean Wave, and the metrical structure of each period is compared and analyzed. Based on this, the current K-Pop metrical structure features are derived and the K-Pop Korean writing method is proposed that deviates from the existing limited writing method which allocates one syllable per note. The author hopes this research will be used as a methodology for writing lyrics in Korean songs in K-Pop, as well as a way to encourage the use of Korean lyrics.
https://doi.org/10.21184/jkeia.2017.10.11.7.343 인용

A Study on the Spoken Korean Citynames Using Multi-Layered Perceptron of Back-Propagation Algorithm (오차 역전파 알고리즘을 갖는 MLP를 이용한 한국 지명 인식에 대한 연구)

Song, Do-Sun;Lee, Jae-Gheon;Kim, Seok-Dong;Lee, Haing-Sei
- The Journal of the Acoustical Society of Korea
- /
- v.13 no.6
- /
- pp.5-14
- /
- 1994
This paper is about an experiment of speaker-independent automatic Korean spoken words recognition using Multi-Layered Perceptron and Error Back-propagation algorithm. The object words are 50 citynames of D.D.D local numbers. 43 of those are 2 syllables and the rest 7 are 3 syllables. The words were not segmented into syllables or phonemes, and some feature components extracted from the words in equal gap were applied to the neural network. That led independent result on the speech duration, and the PARCOR coefficients calculated from the frames using linear predictive analysis were employed as feature components. This paper tried to find out the optimum conditions through 4 differerent experiments which are comparison between total and pre-classified training, dependency of recognition rate on the number of frames and PAROCR order, recognition change due to the number of neurons in the hidden layer, and the comparison of the output pattern composition method of output neurons. As a result, the recognition rate of $89.6\%$ is obtaimed through the research.
PDF

A Pragmatically-oriented Study of Focus and Intonation (억양과 초점에 관한 화용론적 연구)

Lee Yeong-kil
- Proceedings of the Acoustical Society of Korea Conference
- /
- autumn
- /
- pp.379-382
- /
- 1999
모든 문장에는 '새로운' 정보를 전달하기 위한 초점이 있고 높낮돋들림을 포함하는 초점범위는 다시 정보 초점을 필수 요소로 갖는 정보 구조 경계를 갖는다. 모호성이 없는 적절한 초점 구조를 결정하기 위해 '국어 초점 원리'를 도입함으로써 초점 성분의 영역이 확인되고 화맥에 의한 초점 해석이 가능해진다. 초점 성분을 설명하고 높낮돋들림과 초점 돋들림의 관계를 기술하는 '기본초점규칙'이 필요하며 '정보 구조 원리'에 의해 '새로운' 정보가 선택되어 초점 범위는 화맥에 의해 구체화된다. 정보 구조가 문법 체계의 모든 의미 계층과 관계를 가지며 정보 구조의 경계 안에 정보 초점으로 실현되는 초점 돋들림이 있게 되므로 기본 초점 규칙은 '초점 돋들림 원리'로 수정되어 초점 범위 내의 음절에 초점 돋들림이 할당된다.
PDF

Search Result 75, Processing Time 0.019 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)