Search | Korea Science

Frequency-Weighting linear predictive analysis of speech (Frequency-Weighting을 이용한 음성의 선형상측)

김상준;윤종관;조동활
- The Journal of the Acoustical Society of Korea
- /
- v.4 no.1
- /
- pp.43-54
- /
- 1985
이 논문에서는 Frequency weighting을 이용하여 선형예측 부호화기의 명료성을 개선하는 방법 을 연구한다. 잡음이 섞이지 않은 음성에 대해서는 음성을 분석하기전에 frequency weighting을 행한다. 또한 잡음이 섞인 음성인 경우에는 잡음성분을 spectral subtraction 방법에 의해서 제거한 다음에 frequency weighting을 준다. 이 때 frequency weighting을 주기 위해서 귀의 특성과 연관되어 잘 알려 진 C- message weighting 함수, flanagan weighting 함수 및 articulation index를 약간 수정한 weighting 함수를 사용했다. 여러 객관적인 distance measure를 사용하여 frequency weighting 방법의 성능을 측정하고 귀로 들어 본 결과, frequency weighting 방법을 사용하여 선형예측 방법에 의한 합성 음의 명료도를 효율적으로 개선할 수 있었다.
PDF

Comparison of Speech Intelligibility depending on the Sound Source Location in the Classrooms of Middle and High Schools (음원의 위치에 따른 중${\cdot}$고등학교 교실의 음성명료도 비교)

Lee Hwan-Hee;Haan Chan-Hoon
- Proceedings of the Acoustical Society of Korea Conference
- /
- spring
- /
- pp.487-490
- /
- 2002
학교 교육의 특성상 많은 부분이 교실에서의 음성정보 전달에 의해 이루어지고 있는 점을 감안하면 바람직한 청취환경의 개선이 검토되어야 한다. 또한 중${\cdot}$ 고등학교의 수학능력시험의 국어, 영어 듣기평가 및 다양한 어학 시험이 시청각 시설을 통해 이루어지고 있는 실정이므로 교실의 음환경은 매우 중요한 요소라하겠다. 본 논문에서는 음환경을 좌우하는 음원의 위치에 따라 명료 도가 어떻게 달라지는지를 실험을 통하여 검증하고, 명료도가 높고, 교실 전체에 균등한 분포를 보이는 음원의 위치를 찾아내고자 하였다. 교실 내의 음원의 위치로는 일반적으로 많이 쓰이고 있는 column(벽면 노출형)과 ceiling(천정 매입형) 위치와 임의의 음원 cluster(전면 중앙)를 선정하여 음장 파라메터를 측정한 결과 RASTI 는 세 타입 모두 $0.54\~0.55$로 값으로 근소한 차이를 보이고 있으며, 잔향시간은 ceiling>cluster>column의 순서로 나타났다. 일반적으로 잔향과 명료도와의 관계는 반비례하는 것으로 알려져 있으나, 실험 결과 잔향시간이 1.33초로 가장 긴 column 스피커의 경우 D50 값이 약 $47\%$로 가장 높은 값으로 나타났다. 이것은 column형 스피커의 경우 음원과 각 학생의 위치에 대한 평균 직접음선거리가 가장 짧기 때문인 것으로 나타났다.
PDF

Robust Speech Reinforcement Based on Gain-Modification incorporating Speech Absence Probability (음성 부재 확률을 이용한 음성 강화 이득 수정 기법)

Choi, Jae-Hun;Chang, Joon-Hyuk
- Journal of the Institute of Electronics Engineers of Korea SP
- /
- v.47 no.1
- /
- pp.175-182
- /
- 2010
In this paper, we propose a robust speech reinforcement technique to enhance the intelligibility of the degraded speech signal under the ambient noise environments based on soft decision scheme incorporating a speech absence probability (SAP) with speech reinforcement gains. Since the ambient noise significantly decreases the intelligibility of the speech signal, the speech reinforcement approach to amplify the estimated clean speech signal from the background noise environments for improving the intelligibility and clarity of the corrupted speech signal was proposed. In order to estimate the robust reinforcement gain rather than the conventional speech reinforcement method between speech active periods and nonspeech periods or transient intervals, we propose the speech reinforcement algorithm based on soft decision applying the SAP to the estimation of speech reinforcement gains. The performances of the proposed algorithm are evaluated by the Comparison Category Rating (CCR) of the measurement for subjective determination of transmission quality in ITU-T P.800 under various ambient noise environments and show better performances compared with the conventional method.
PDF KSCI

From Clarity To Human Voice (명료도에서 사람 목소리로 - TTS에 관하여)

권철홍
- Proceedings of the Acoustical Society of Korea Conference
- /
- 1998.06c
- /
- pp.139-142
- /
- 1998
그 동안 TTS 음성합성의 평가 척도로 명료도(Clarity)와 자연성(Naturalness)을 기준으로 삼았다. 이제는 합성음의 평가 기준이 사람 목소리와 이해도가 되는 것이 좋겠다고 생각한다. 본 논문은 사람 목소리와 이해도라는 척도 중에서 사람 목소리에 관한 주제를 다루고자 한다. 이를 위하여 음성 DB의 합성 단위로 CVC type을 기본으로 하고, CV, VC type으로 보강한 단위를 선정하여 음성 DB를 구축하였다. 그리고 합성 알고리즘은 음색을 살리며 피치 변경이 용이한 PS-RELP 알고리즘을 제안하였다.
PDF

A method of wall absorption treatment for enhancing the speech intelligibility at a directional microphone array in a room (실내 공간 내 지향성 마이크 어레이에서의 음성 명료도 개선을 위한 벽면 흡음 처리 방법)

Ko, Byeong-Yun;Ih, Jeong-Guon;Cho, Wan-Ho
- The Journal of the Acoustical Society of Korea
- /
- v.40 no.6
- /
- pp.649-659
- /
- 2021
Wall absorption treatment effectively reduces reverberation, but requires a large area for a live room and each wall absorption affects speech intelligibility differently. In this study, we try to find the most effective wall for the absorption treatment using the beamforming array microphone in terms of speech intelligibility. The absorption importance factor is defined by using the collision number of reflected sounds on each wall. It allows estimating how much the speech signal will be enhanced by the absorption treatment. A cuboid room with a size of 107 m³ and a reverberation time of 1.1 s is selected for the simulation. When a Helmholtz-type absorption is treated on the wall with the most significant importance factor, the modified clarity for 500 and 1k Hz is improved by 5.1 dB and 4.8 dB respectively, and the speech transmission index is enhanced by 0.06. The difference in results between the proposed method and commercial simulation code is less than a Just-Noticeable Difference (JND). The absorption treatment on the wall with the most significant importance factor shows improvement greater than the wall with the largest area, and its difference is larger than a JND value.
https://doi.org/10.7776/ASK.2021.40.6.649 인용 PDF KSCI

Adaptation of Classification Model for Improving Speech Intelligibility in Noise (음성 명료도 향상을 위한 분류 모델의 잡음 환경 적응)

Jung, Junyoung;Kim, Gibak
- Journal of Broadcast Engineering
- /
- v.23 no.4
- /
- pp.511-518
- /
- 2018
This paper deals with improving speech intelligibility by applying binary mask to time-frequency units of speech in noise. The binary mask is set to "0" or "1" according to whether speech is dominant or noise is dominant by comparing signal-to-noise ratio with pre-defined threshold. Bayesian classifier trained with Gaussian mixture model is used to estimate the binary mask of each time-frequency signal. The binary mask based noise suppressor improves speech intelligibility only in noise condition which is included in the training data. In this paper, speaker adaptation techniques for speech recognition are applied to adapt the Gaussian mixture model to a new noise environment. Experiments with noise-corrupted speech are conducted to demonstrate the improvement of speech intelligibility by employing adaption techniques in a new noise environment.
https://doi.org/10.5909/JBE.2018.23.4.511 인용 PDF KSCI KPUBS

Comparison of Sound Pressure Level and Speech Intelligibility of Emergency Broadcasting System at T-junction Corridor Space (T자형 복도 공간의 비상 방송용 확성기 배치별 음압 레벨과 음성 명료도 비교)

Jeong, Jeong-Ho;Lee, Sung-Chan
- Fire Science and Engineering
- /
- v.33 no.1
- /
- pp.105-112
- /
- 2019
In this study, an architectural acoustics simulation was conducted to examine the clear and uniform transmission of emergency broadcasting sound in a T junction corridor space. The sound absorption performance of the corridor space and the location and spacing of the loudspeaker for emergency broadcasting were varied. The distribution of the sound pressure level and the distribution of sound transmission indices (STI, RASTI) were compared. The simulation showed that the loudspeaker for emergency broadcasting should be installed approximately 10 m from the center of the T junction corridor connection for clear voice transmission. Narrowing the 25 m installation interval of the NFSC shows that an even clearer and sufficient volume of emergency broadcast sound can be delivered evenly.
https://doi.org/10.7731/KIFSE.2019.33.1.105 인용 PDF KSCI HTML

Comparison of Sound Pressure Level and Speech Intelligibility of Emergency Broadcasting System at Longitudinal Corridor (장방향 복도 공간의 비상방송설비에 대한 음압 레벨과 음성 명료도 비교)

Jeong, Jeong-Ho;Lee, Sung-Chan
- Fire Science and Engineering
- /
- v.32 no.4
- /
- pp.42-49
- /
- 2018
In this study, in order to investigate whether or not the emergency broadcasting sound generated from an emergency broadcasting speaker is clearly transmitted to the occupant through architectural sound simulation, when the loudspeaker for emergency broadcasting is installed at intervals of 25 m according to NFSC 202 for a rectangular hallway. The sound pressure level and speech intelligibility index were analyzed according to changes in building finishing materials. With a reflective material finishing, sound pressure level satisfied the standard while speech intelligibility index was low. As a result of applying the sound absorbing material finishing, clarity and speech transmission index was improved to a level that could be understood by the occupant, whereas the sound pressure level delivered to the occupant decreased in the same space.
https://doi.org/10.7731/KIFSE.2018.32.4.042 인용 PDF KSCI HTML

Speech Intelligibility Analysis on the Laser Detected Sound of the Glass Windows (유리창의 레이저 탐지음에 대한 음성명료도 분석)

Kim, Seock-Hyun;Lee, Hyun-Woo;Kim, Hee-Dong
- The Journal of the Acoustical Society of Korea
- /
- v.28 no.2
- /
- pp.127-134
- /
- 2009
In this study, possibility of the laser eavesdropping is investigated on the window glasses with various thicknesses, Glass windows are excited by maximum length sequency (MLS) signal and the vibration sound is detected by a laser doppler vibrometer. From the detected sound, speech intelligibility is objectively estimated. Speech transmission index (STI), which is based on the modulation transfer function (MTF). is calculated for the estimation. Finally, disturbing wave effect on the speech intelligibility is analysed by using an outside speaker and a window shaker attached on the glass window. The purpose of the study is to estimate the possibility of remote eavesdropping by the laser sensor and to evaluate the performance of the homemade window shaker to protect from the remote eavesdropping.
https://doi.org/10.7776/ASK.2009.28.2.127 인용 PDF KSCI

Investigation of the listening environment of classrooms for elderly people using speech intelligibility tests (음성명료도 시험에 의한 노인 교육시설의 청취환경 조사)

Park, Chan-Jae;Kim, Bo-Gyeong;Haan, Chan-Hoon
- The Journal of the Acoustical Society of Korea
- /
- v.40 no.1
- /
- pp.18-30
- /
- 2021
The ultimate goal of the present study is to establish the acoustical performance standards of classroom for the elderly who are incomplete hearing people. As a pilot survey, the present study was conducted to investigate the listening environment and the actual condition of speech perception performance of education facilities for elderly, Acoustic performances of two education facilities for elderly in Cheongju were measured and questionnaire survey was done to elderly people. Also, speech intelligibility tests were undertaken by Consonant Vowel Consonant (CVC) and Phonetically Balanced Words (PBW) methods. The questionnaire survey showed that the elderly were satisfied with the listening environment of the educational facilities in general. Also, it was found that acoustical performances satisfy with the acoustic criteria of general classrooms in Korea. However, the results of the speech intelligibility test showed that the scores of elderly were significantly lower than twenties with normal hearing. It was also revealed that the scores are reduced as the age increases. Thus, it was concluded that the acoustical performance standards of educational facilities for the normal hearing were not suitable for educational facilities for the elderly.
https://doi.org/10.7776/ASK.2021.40.1.018 인용 PDF KSCI

Search Result 189, Processing Time 0.029 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)