Search | Korea Science

Real Time Speaker Close-Up and Tracking System Using the Lip Varying Informations (입술 움직임 변화량을 이용한 실시간 화자의 클로즈업 및 트레킹 시스템 구현)

양운모;장언동;윤태승;곽내정;안재형
- Proceedings of the Korea Multimedia Society Conference
- /
- 2002.05d
- /
- pp.547-552
- /
- 2002
본 논문에서는 다수의 사람이 존재하는 입력영상에서 입술 움직임 정보를 이용한 실시간 화자의 클로즈업(close-up) 시스템을 구현한다. 칼라 CCD 카메라를 통해 입력되는 동영상에서 화자를 검출한 후 입술 움직임 정보를 이용하여 다른 한 대의 카메라로 화자를 클로즈업한다. 구현된 시스템은 얼굴색 정보와 형태 정보를 이용하여 각 사람의 얼굴 및 입술 영역을 검출한 후, 입술 영역 변화량을 이용하여 화자를 검출한다. 검출된 화자를 클로즈업하기 위하여 PTZ(Pan/Tilt/Zoom) 카메라를 사용하였으며, RS-232C 시리얼 포트를 이용하여 카메라를 제어한다. 실험결과 3인 이상의 입력 동영상에서 정확하게 화자를 검출할 수 있으며, 움직이는 화자의 얼굴 트레킹이 가능하다.
PDF

Real Time Speaker Close-Up System using The Lip Motion Informations (입술 움직임 정보를 이용한 실시간 화자 클로즈업 시스템 구현)

권혁봉;장언동;윤태승;안재형
- Journal of Korea Multimedia Society
- /
- v.4 no.6
- /
- pp.510-517
- /
- 2001
In this paper, we implement a real time speaker close-up system using lip motion information from input images having some people. After detecting a speaker from input moving pictures through one color CCD camera, the other camera closes up the speaker by using lip motion information. The implemented system detects a face and lip area of each person by means of a facial color and a morphological information, and then finds out a speaker by using lip area variation. A PTZ(Pan/Tilt/Zoom) camera is used in order to close up the detected speaker and it is controlled by RS-232C serial port. Consequently, we can exactly detect a speaker in input moving pictures including more than three people.
PDF

Coarticulation Model of Hangul Visual speedh for Lip Animation (입술 애니메이션을 위한 한글 발음의 동시조음 모델)

Gong, Gwang-Sik;Kim, Chang-Heon
- Journal of KIISE:Computer Systems and Theory
- /
- v.26 no.9
- /
- pp.1031-1041
- /
- 1999
기존의 한글에 대한 입술 애니메이션 방법은 음소의 입모양을 몇 개의 입모양으로 정의하고 이들을 보간하여 입술을 애니메이션하였다. 하지만 발음하는 동안의 실제 입술 움직임은 선형함수나 단순한 비선형함수가 아니기 때문에 보간방법에 의해 중간 움직임을 생성하는 방법으로는 음소의 입술 움직임을 효과적으로 생성할 수 없다. 또 이 방법은 동시조음도 고려하지 않아 음소들간에 변화하는 입술 움직임도 표현할 수 없었다. 본 논문에서는 동시조음을 고려하여 한글을 자연스럽게 발음하는 입술 애니메이션 방법을 제안한다. 비디오 카메라로 발음하는 동안의 음소의 움직임들을 측정하고 입술 움직임 제어 파라미터들을 추출한다. 각각의 제어 파라미터들은 L fqvist의 스피치 생성 제스처 이론(speech production gesture theory)을 이용하여 실제 음소의 입술 움직임에 근사한 움직임인 지배함수(dominance function)들로 정의되고 입술 움직임을 애니메이션할 때 사용된다. 또, 각 지배함수들은 혼합함수(blending function)와 반음절에 의한 한글 합성 규칙을 사용하여 결합하고 동시조음이 적용된 한글을 발음하게 된다. 따라서 스피치 생성 제스처 이론을 이용하여 입술 움직임 모델을 구현한 방법은 기존의 보간에 의해 중간 움직임을 생성한 방법보다 실제 움직임에 근사한 움직임을 생성하고 동시조음도 고려한 움직임을 보여준다.Abstract The existing lip animation method of Hangul classifies the shape of lips with a few shapes and implements the lip animation with interpolating them. However it doesn't represent natural lip animation because the function of the real motion of lips, during articulation, isn't linear or simple non-linear function. It doesn't also represent the motion of lips varying among phonemes because it doesn't consider coarticulation. In this paper we present a new coarticulation model for the natural lip animation of Hangul. Using two video cameras, we film the speaker's lips and extract the lip control parameters. Each lip control parameter is defined as dominance function by using L fqvist's speech production gesture theory. This dominance function approximates to the real lip animation of a phoneme during articulation of one and is used when lip animation is implemented. Each dominance function combines into blending function by using Hangul composition rule based on demi-syllable. Then the lip animation of our coarticulation model represents natural motion of lips. Therefore our coarticulation model approximates to real lip motion rather than the existing model and represents the natural lip motion considered coarticulation.

A Tracking Method of Robust Lip Movement Image Regions for Blocking the External Acoustic Noise (외부응향잡음 차단을 위한 강인한 입술움직임 영상영역 추적방법)

Kim, Eung-Kyeu
- Proceedings of the KIEE Conference
- /
- 2009.07a
- /
- pp.1913_1914
- /
- 2009
본 논문에서 조명환경하에서 음성/영상 연동시스템을 통해서 외부음향잡음의 차단을 위한 강인한 입술움직임 영상영역을 추적하는 한 가지 방법을 제안한다. 조명환경하에서 강인한 입술움직임 영상영역을 추적하기 위해 온라인상에서 입술움직임 표준영상을 수집하였고 다양한 조명환경에 적응하는 입술 움직임 영상의 특징들을 추출하였다. 동시에 온라인 템플릿 영상을 획득하였고, 이 영상들을 템플릿 정합을 위해 사용했다. 음성/영상처리시스템의 연동결과, 다양한 조명환경하에서 그 연동률을 99.3%까지 높일 수 있었고 음향잡음에 의한 음성인식 실행을 원천적으로 차단할 수 있었다.
PDF

Change in lip movement during speech by aging: Based on a double vowel (노화에 따른 발화 시 입술움직임의 변화: 이중모음을 중심으로)

Park, Hee-June
- Phonetics and Speech Sciences
- /
- v.13 no.1
- /
- pp.73-79
- /
- 2021
This study investigated the change in lip movement during speech according to aging. For the study, 15 elderly women with an average of 69 years and 15 young women with an average of 22 years were selected. To measure the movement of the lips, the ratio between the minimum point and the maximum point of movement when pronouncing a double vowel was analyzed in pixel units using image analysis software. For clinical utility, the software was produced by applying an automated algorithm and compared with the results of handwork. This study found that the range of the width and length of lips in double vowel tasks was smaller for the elderly than that of the young. A strong positive correlation was found between manual and automated methods, indicating that both methods are useful for extracting lip contours. Based on the above results, it was found that the range of the lips decreased when ignited as aging progressed. Therefore, monitoring the condition of lip performance by simply measuring the movement of lips before aging progresses, and performing exercises to maintain lip range, will prevent pronunciation problems caused by aging.
https://doi.org/10.13064/KSSS.2021.13.1.073 인용 PDF KSCI

Vowels(a,e,i,o,u) Analysis Using Optical Flow (Optical Flow를 이용한 단모음(아,에,이,오,우) 분석)

이미애;박기수
- Proceedings of the Korea Multimedia Society Conference
- /
- 2002.05c
- /
- pp.299-302
- /
- 2002
컴퓨터를 이용한 독순 연구는 Man Machine Interface, 지적부호화에 있어서의 송신측 기술, 청각 장애인의 독순 훈련 시스템 등 다방면에서 그 응용이 기대된다. 본 논문은, 움직임 정보는 입술의 에지영역에 집중하고 있음에 주목하여, 입술 에지영역의 Optical Flow 추정값을 독순정보로 이용하는 방법을 제안한다. 휘도값을 갖지 않는 에지에, 선형 가상 휘도값를 정해주어 Optical Flow를 추정하는 VGM을 도입해 특징 파라미터를 계산하고, 마할라노비스 평방거리(Mahalanobis's square distance)에 기초한 최대우도판별함수를 이용하여 단모음을 분석하는 알고리즘을 제안한다.
PDF

A Study on Spatio-temporal Features for Korean Vowel Lipreading (한국어 모음 입술독해를 위한 시공간적 특징에 관한 연구)

오현화;김인철;김동수;진성일
- The Journal of the Acoustical Society of Korea
- /
- v.21 no.1
- /
- pp.19-26
- /
- 2002
This paper defines the visual basic speech units, visemes and investigates various visual features of a lip for the effective Korean lipreading. First, we analyzed the visual characteristics of the Korean vowels from the database of the lip image sequences obtained from the multi-speakers, thereby giving a definition of seven Korean vowel visemes. Various spatio-temporal features of a lip are extracted from the feature points located on both inner and outer lip contours of image sequences and their classification performances are evaluated by using a hidden Markov model based classifier for effective lipreading. The experimental results for recognizing the Korean visemes have demonstrated that the feature victor containing the information of inner and outer lip contours can be effectively applied to lipreading and also the direction and magnitude of the movement of a lip feature point over time is quite useful for Korean lipreading.
PDF KSCI

Improvement of Lipreading Performance Using Gabor Filter for Ship Environment (선박 환경에서 Gabor 여파기를 적용한 입술 읽기 성능향상)

Shin, Do-Sung;Lee, Seong-Ro;Kwon, Jang-Woo
- The Journal of Korean Institute of Communications and Information Sciences
- /
- v.35 no.7C
- /
- pp.598-603
- /
- 2010
In this paper, we work for Lipreading using visual information for ship environment. Lipreading is studied for using image information including lips of a speaker at the existing speech recognition system. This technique is a compensation method to increase recognition rate decreasing remarkably in noisy circumstances. Proposed way improved the rate of recognition improving methode of preprocessing using the Gabor Filter for Ship Environment. The experiment were carried out under changing of light with time in the ship environment with lip image. For Comparing with recognition, make a compare with between method of lip region of interest (ROI) before Gabor filtering and after Gabor filtering. In the case of using method of lip ROI before Gabor filtering, the result of the experiments applying to the proposed ways recognition resulting in 44% of recognition.
PDF KSCI

기능성 음성 질환(Functional Voice Disorders)과 성대의 움직임

안철민
- Proceedings of the KSLP Conference
- /
- 2003.11a
- /
- pp.190-192
- /
- 2003
음성은 단순히 성대에서 만들어지는 것이 아니다. 호흡을 시작으로 성대의 접촉과 점막 진동에 의해 만들어진 소리가 공명강을 거쳐 입술, 혀의 움직임을 거쳐 최종적으로 의미를 전달하는 소리로 완성된다. 기능성 음성 질환은 이러한 과정 중에서 발성 방법과 같은 기능적 문제에 의하여 발생하게 된다. 따라서 기능성 음성 질환이 있을 때 이러한 과정의 움직임에 대한 조사가 필요하다. (중략)
PDF

Supporting the Korean Lip Synchronization and Facial Expression (한글 입술 움직임과 얼굴 표정이 동기화된 3차원 개인 아바타 대화방 시스템)

Lee, Jung;Oh, Beom-Soo;Jeong, Won-Ki;Kim, Chang-Hun
- Proceedings of the Korean Information Science Society Conference
- /
- 2000.04b
- /
- pp.640-642
- /
- 2000
대화방 시스템은 텍스트화 화상을 이용한 대화방 또는 메시지 전달시스템이 널리 사용되고 있다. 본 논문은 3차원 아바타가 등장하는 대화방 시스템을 생성 및 관리하는 기술을 제안한다. 본 아바타 대화방의 특징은 사진을 가지고 간단히 3차원 개인 아바타로 변환 생성하는 기술, 3차원 개인 아바타의 한글 발음에 적합한 입술 움직임, 메시지에 따른 적절한 표정변화 등이다. 특히, 3차원 개인 아바타는 사진만으로 생성이 가능하며, 텍스쳐 매핑된 3차원 아바타는 실시간으로 사실감있는 대화방 서비스가 가능하도록 제어된다.
PDF

Search Result 33, Processing Time 0.032 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)