Search | Korea Science

On an Adaptation of Announcement Sound Level in White Noise Environment (백색소음 환경에서 음성안내레벨 적응에 관한 연구)

Yun, Jong-Jin;Bae, Myung-Jin
- Journal of the Institute of Electronics Engineers of Korea SP
- /
- v.49 no.1
- /
- pp.112-118
- /
- 2012
In daily life, there are many information broadcasting by using voice information systems. If surrounding noises are mixed with the information signals, the clarity of the signal become down graded too much to understand. Surrounding noises are not uniformed, but very irregular signals always changing. Therefore, it is very hard to control the output signals along with the irregular signals. This paper suggests a method to change the level of the voice information adapting to the surround noise in the white noise environment. The surround noise level is measured by subtracting the stored output voice signal from the voice signal degraded by the noise. The noise is used to estimation of SNR. And, the method to change the output level of voice signal adapting to the noise level. The suggested adaptive voice information system has the advantage to improve listeners' speech perception and to use amplifier's energy effectively.
PDF KSCI

Voice Analysis and Treatment Result According to Configuration of Sulcus Vocalis (성대구증의 형태에 따른 음향학적 분석 및 치료 결과)

Yang, Ho Cherl;Jeong, Byoung Seo;Kim, Dong Young;Woo, Joo Hyun
- Journal of the Korean Society of Laryngology, Phoniatrics and Logopedics
- /
- v.23 no.2
- /
- pp.119-123
- /
- 2012
Background and Objectives : Sulcus vocalis could be classified into type I, type IIa, and type IIb. There have been a little reports about voice quality and treatment results related with types of sulcus vocalis. The authors conducted an analysis of voice and treatment according to different types of sulcus vocalis. Materials and Methods : This study was based on a retrospective chart review. The sulcus types were classified into type I and type II. Objective and subjective voice assessments were analyzed. Patients were treated individually with voice therapy, percutaneous steroid injection, and injection laryngoplasty. Comparison was performed on the voice difference between type I group and type II group, and between pre-treatment and post-treatment of each types. Results : One hundred and one patients were enrolled into this study, and 49 patients were type I and 52 patients were type II. Type I group showed longer mean maximal phonation time (MPT) than type II group, although other voice parameters didn't show any difference between two groups. Even after the management, almost all of the voice parameters didn't show improvement except MPT of type II group. Conclusion：Although the type I sulcus has been known as a non-pathologic lesion, it can result in some degree of voice change and discomfort, and thus need an active management. In this study, voice therapy, percutaneous steroid injection, and injection laryngoplasty showed limited effect to the both types of sulcus vocalis. Further studies for management of sulcus vocalis were needed.
PDF

Validity of Voice Handicap Index and Voice Analysis following Laryngeal Microsurgery for Benign Vocal Cord Lesions (양성 성대 질환 환자의 후두 미세 수술 전후 음성 장애 지수 및 음성 분석의 유용성)

Park, Young-Hak;Lee, Jeong-Hak;Joo, Young-Hoon;Park, Sung-Sin;Bang, Choong-Il;Kim, Min-Sik;Cho, Seung-Ho
- Journal of the Korean Society of Laryngology, Phoniatrics and Logopedics
- /
- v.16 no.1
- /
- pp.23-27
- /
- 2005
Background and Objectives : Voice disorders can cause problems in patients with benign vocal cord lesions emotionally, physically, economically and functionally. Neither subjective nor objective voice examinations can evaluate such factors adequately. The Voice Handicap Index (VHI) subjectively evaluates voice disorders in terms of physical, functional, emotional factors and measures the patient's perception of the impact of voice disorder. The purpose of this study is to evaluate the usefulness of VHI in the patients with benign vocal cord lesions. Materials and Method : The authors evaluated 37 patients who experienced laryngeal microsurgery for benign vocal cord lesions from september 2003 to August 2004. The VHI was used to measure the postoperative changes of the patient's perception and acoustic analysis and aerodynamic tests were also done. Statistical analysis was done using paired t-test and Pearson's correlation. Results : The VHI scores showed statistically significant reductions postoperatively. In acoustic analysis, jitter and shimmer had statistically significant reductions after surgery but noise-to-harmonics ratio did not. A statistically significant change in the average MFR and MPT perioperatively was found. The relationship between VHI and acoustic, aerodynamic analysis attained statistical significance. Conclusion : The VHI is a useful assessment tool to monitor the patient's self-perception of voice change after the surgery of benign vocal cord lesions. The VHI measurement, when combined with acoustic and aerodynamic analyses, will be helpful in comparing functional outcomes after voice surgery.
PDF

Clinical Study of Aged Patients with Hoarseness (노인애성환자에 대한 임상적연구)

안철민;권기환
- Journal of the Korean Society of Laryngology, Phoniatrics and Logopedics
- /
- v.7 no.1
- /
- pp.27-31
- /
- 1996
The voice of aged persons is known generally to be somewhat different from that of other adults, suggesting that laryngeal change occurs with advancing age. However, because knowledge of the voice characteristics of aged persons is limited, it is difficult to judge whether their voices arc normal. Chart review and laryngoscopic examination from ninety-one patients with hoarseness over the age of 60(1st group) and one hundred sixteen patients with hoarseness below the age of 50(2nd group) were done to define aging related voice disorders. The following results were obtained. 1) Associated diseases related to laryngeal disease were hypertension(12%), pulmonary disease(4.4%), thyroid disease(1.1%) in 1st group and hypertension(9.5%), thyroid disease(1.7%) in 2nd group. 2) The underlying diseases causing hoarseness in order of frequency were benign vocal fold lesion(37.7%), inflammatory disease(36.8%), functional dysphonia(17%) in 1st group and benign vocal fold lesion(43.6%), functional dysphonia(26.3%), inflammatory disease(16.5%) in 2nd group. 3) In stroboscopic findings, atrophy and sulcus of vocal cords are more prevalent in males than in females and edema of vocal cords is more common in females. Generally the voice characteristics of aged persons depend on the mass of the vocal folds which may be decreased through atrophy or be increased by edema. However, other factors such as systemic diseases, drug side effects and compensatory mechanism to presbylaryngis must be taken into account in diagnosing and treating voice disorders in aged persons.
PDF

Artificial Intelligence for Clinical Research in Voice Disease (후두음성 질환에 대한 인공지능 연구)

Jungirl, Seok;Tack-Kyun, Kwon
- Journal of the Korean Society of Laryngology, Phoniatrics and Logopedics
- /
- v.33 no.3
- /
- pp.142-155
- /
- 2022
Diagnosis using voice is non-invasive and can be implemented through various voice recording devices; therefore, it can be used as a screening or diagnostic assistant tool for laryngeal voice disease to help clinicians. The development of artificial intelligence algorithms, such as machine learning, led by the latest deep learning technology, began with a binary classification that distinguishes normal and pathological voices; consequently, it has contributed in improving the accuracy of multi-classification to classify various types of pathological voices. However, no conclusions that can be applied in the clinical field have yet been achieved. Most studies on pathological speech classification using speech have used the continuous short vowel /ah/, which is relatively easier than using continuous or running speech. However, continuous speech has the potential to derive more accurate results as additional information can be obtained from the change in the voice signal over time. In this review, explanations of terms related to artificial intelligence research, and the latest trends in machine learning and deep learning algorithms are reviewed; furthermore, the latest research results and limitations are introduced to provide future directions for researchers.
https://doi.org/10.22469/jkslp.2022.33.3.142 인용 PDF KSCI

Production and perception of Korean word-initial stops from a sound change perspective (음 변화 관점에서 바라본 한국어 어두 폐쇄음의 발화 및 지각)

Kim, Jin-Woo
- Phonetics and Speech Sciences
- /
- v.13 no.3
- /
- pp.39-51
- /
- 2021
Based on spontaneous speech data collected in 2020, this study examined the production and perception of Korean lenis, aspirated, and fortis stops. Unlike the controlled experiments of previous studies, lenis and aspirated stops of males in their 30s were not distinguished by voice onset time (VOT) in spontaneous speech. Perceptual experiments were conducted on young females, the leaders of language change. F0 was found to serve as the primary cue for the perception of lenis stops, and then VOT distinguished the aspirated and fortis stops. The fact that the sounds were always perceived as lenis stops when F0 was low, irrespective of whether VOT was short or long, showed that F0 plays an absolute role in the perception of lenis stops. However, in some cases the aspirated and lenis stops were distinguished only by VOT, which does not happen in production. In terms of sound change, disagreement between production and perception systems occurs when sound change is in progress. In particular, when production change precedes perception change, it indicates that the sound change is in its latter stages. Young females still maintain the previous system in perception because the distinction of lenis and aspirated stops by VOT was valid in their parents' generation. In other words, VOT is still used for perception to communicate with other groups.
https://doi.org/10.13064/KSSS.2021.13.3.039 인용 PDF KSCI

Voice Color Conversion Based on the Formants and Spectrum Tilt Modification (포먼트 이동과 스펙트럼 기울기의 변환을 이용한 음색 변환)

Son Song-Young;Hahn Min-Soo
- MALSORI
- /
- no.45
- /
- pp.63-77
- /
- 2003
The purpose of voice color conversion is to change the speaker identity perceived from the speech signal. In this paper, we propose a new voice color conversion algorithm through the formant shifting and the spectrum-tilt modification in the frequency domain. The basic idea of this technique is to convert the positions of source formants into those of target speaker's formants through interpolation and decimation and to modify the spectrum-tilt by utilizing the information of both speakers' spectrum envelops. The LPC spectrum is adopted to evaluate the position of formant and the information of spectrum-tilt. Our algorithm enables us to convert the speaker identity rather successfully while maintaining good speech quality, since it modifies speech waveforms directly in the frequency domain.
PDF

Significance of Acoustic Parameter - RAP, PPQ, APQ- in Hoarseness (애성환자에서 음향지표인 RAP, PPQ 및 APQ의 유용성)

안철민;이종혁;강현국;이용배
- Journal of the Korean Society of Laryngology, Phoniatrics and Logopedics
- /
- v.6 no.1
- /
- pp.22-26
- /
- 1995
Change of voice, espicially hoarseness show irregular vibration of vocal cord. So, computerized acoustic analysis has presented many acoustic parameters for objective evaluation of voice. We objectively investigated the vocal vibration of normal persons and hoarseness patients in Korea. The RAP(relative average perturbation), PPQ(pitch period perturbation quotient) and APQ(amplitude perturbation quotient) of normal persons were compared with that of hoarseness patients with multidimensional voice program for the possibility of distinguishing the pathologic vocal vibration from normal. Authors agree that RAP, PPQ and APQ showed interesting differences between the normal and the hoarseness patients by the multivariate statistical analysis. In conculusion, relative average perturbation, pitch period perturbation and amplitude perturbation quotient might be meangingful screening parameters distinguishing hoarseness patients from normal.
PDF

An Implementation of Travel Information Service Using VoiceXML and GPS (VoiceXML과 GPS를 이용한 여행정보 서비스의 구현)

Oh, Jae-Gyu;Kim, Sun-Hyung
- Journal of the Korea Academia-Industrial cooperation Society
- /
- v.8 no.6
- /
- pp.1443-1448
- /
- 2007
In this paper, we implement a distributed computing environment-based travel information service that can use web(internet) and speech interface at the same time and can apply location information, using voice and web browser-based VoiceXML and GPS, to escape the limitations of traditional web(internet)-based travel information services. Because of IVR(Interactive Voice Response) of traditional call center has operated to a pre-installation scenario, it takes much a service time and has the inconveniences that must repeat speech recording according to the revised scenarios in case change response contents. However, suggested VoiceXML and GPS-based travel information service system has advantages that reorganization of system setups is easy, because it consists of the method to update server after make individual conversation scenarios by file format(document), and can provide usefully various travel information in environmental restriction conditions such as the back regions environment, according as our prototype find user's present location using GPS information and then provide various travel information service by this information.
PDF

Voice Activity Detection Method Using Psycho-Acoustic Model Based on Speech Energy Maximization in Noisy Environments (잡음 환경에서 심리음향모델 기반 음성 에너지 최대화를 이용한 음성 검출 방법)

Choi, Gab-Keun;Kim, Soon-Hyob
- The Journal of the Acoustical Society of Korea
- /
- v.28 no.5
- /
- pp.447-453
- /
- 2009
This paper introduces the method for detect voices and exact end point at low SNR by maximizing voice energy. Conventional VAD (Voice Activity Detection) algorithm estimates noise level so it tends to detect the end point inaccurately. Moreover, because it uses relatively long analysis range for reflecting temporal change of noise, computing load too high for application. In this paper, the SEM-VAD (Speech Energy Maximization-Voice Activity Detection) method which uses psycho-acoustical bark scale filter banks to maximize voice energy within frames is introduced. Stable threshold values are obtained at various noise environments (SNR 15 dB, 10 dB, 5 dB, 0 dB). At the test for voice detection in car noisy environment, PHR (Pause Hit Rate) was 100%accurate at every noise environment, and FAR (False Alarm Rate) shows 0% at SNR15 dB and 10 dB, 5.6% at SNR5 dB and 9.5% at SNR0 dB.
https://doi.org/10.7776/ASK.2009.28.5.447 인용 PDF KSCI

Search Result 360, Processing Time 0.025 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)