통합 검색 | Korea Science

음성 강화를 위한 a priori SNR 추정기반 적응 바람소리 저감 방법 (An Adaptive Wind Noise Reduction Method Based on a priori SNR Estimation for Speech Eenhancement)

서지훈;이석필
- 전기학회논문지
- /
- 제64권12호
- /
- pp.1756-1760
- /
- 2015
This paper focuses on a priori signal to noise ratio (SNR) estimation method for the speech enhancement. There are many researches for speech enhancement with several ambient noise cancellation methods. The method based on spectral subtraction (SS) which is widely used in noise reduction has a trade-off between the performance and the distortion of the signals. So the need of adaptive method like an estimated a priori SNR being able to making a high performance and low distortion is increasing. The decision directed (DD) approach is used to determine a priori SNR in noisy speech signals. A priori SNR is estimated by using only the magnitude components and consequently follows a posteriori SNR with one frame delay. We propose a modified a priori SNR estimator and the weighted rational transfer function for speech enhancement with wind noises. The experimental result shows the performance of our proposed estimator is better Perceptual Evaluation of Speech Quality scores (PESQ, ITU-T P.862) compare to the conventional DD approach-based systems and different noise reduction methods.
https://doi.org/10.5370/KIEE.2015.64.12.1756 인용 PDF KSCI

SNR Enhancement Algorithm Using Multiple Chirp Symbols with Clock Drift for Accurate Ranging

Jang, Seong-Hyun;Kim, Yeong-Sam;Yoon, Sang-Hun;Chong, Jong-Wha
- ETRI Journal
- /
- 제33권6호
- /
- pp.841-848
- /
- 2011
A signal-to-noise ratio (SNR) enhancement algorithm using multiple chirp symbols with clock drift is proposed for accurate ranging. Improvement of the ranging performance can be achieved by using the multiple chirp symbols according to Cramer-Rao lower bound; however, distortion caused by clock drift is inevitable practically. The distortion induced by the clock drift is approximated as a linear phase term, caused by carrier frequency offset, sampling time offset, and symbol time offset. SNR of the averaged chirp symbol obtained from the proposed algorithm based on the phase derotation and the symbol averaging is enhanced. Hence, the ranging performance is improved. The mathematical analysis of the SNR enhancement agrees with the simulations.
https://doi.org/10.4218/etrij.11.0111.0013 인용 PDF KSCI

Speech Enhancement Using Phase-Dependent A Priori SNR Estimator in Log-Mel Spectral Domain

Lee, Yun-Kyung;Park, Jeon Gue;Lee, Yun Keun;Kwon, Oh-Wook
- ETRI Journal
- /
- 제36권5호
- /
- pp.721-729
- /
- 2014
We propose a novel phase-based method for single-channel speech enhancement to extract and enhance the desired signals in noisy environments by utilizing the phase information. In the method, a phase-dependent a priori signal-to-noise ratio (SNR) is estimated in the log-mel spectral domain to utilize both the magnitude and phase information of input speech signals. The phase-dependent estimator is incorporated into the conventional magnitude-based decision-directed approach that recursively computes the a priori SNR from noisy speech. Additionally, we reduce the performance degradation owing to the one-frame delay of the estimated phase-dependent a priori SNR by using a minimum mean square error (MMSE)-based and maximum a posteriori (MAP)-based estimator. In our speech enhancement experiments, the proposed phase-dependent a priori SNR estimator is shown to improve the output SNR by 2.6 dB for both the MMSE-based and MAP-based estimator cases as compared to a conventional magnitude-based estimator.
https://doi.org/10.4218/etrij.14.2214.0039 인용 PDF KSCI KPUBS

위상 정보를 고려한 로그멜 영역에서의 2단계 선험 SNR 추정 (Two-step a priori SNR Estimation in the Log-mel Domain Considering Phase Information)

이윤경;권오욱
- 말소리와 음성과학
- /
- 제3권1호
- /
- pp.87-94
- /
- 2011
The decision directed (DD) approach is widely used to determine a priori SNR from noisy speech signals. In conventional speech enhancement systems with a DD approach, a priori SNR is estimated by using only the magnitude components and consequently follows a posteriori SNR with one frame delay. We propose a phase-dependent two-step a priori SNR estimator based on the minimum mean square error (MMSE) in the log-mel spectral domain so that we can consider both magnitude and phase information, and it can overcome the performance degradation caused by one frame delay. From the experimental results, the proposed estimator is shown to improve the output SNR of enhanced speech signals by 2.3 dB compared to the conventional DD approach-based system.
PDF

음성 향상에서 강인한 새로운 선행 SNR 추정 기법에 관한 연구 (A Novel Approach to a Robust A Priori SNR Estimator in Speech Enhancement)

박윤식;장준혁
- 한국음향학회지
- /
- 제25권8호
- /
- pp.383-388
- /
- 2006
본 논문에서는 잡음 환경에서 단일 마이크로폰의 음성 향상에 대한 새로운 기법을 제시했다. 일반적으로 널리 알려진 스펙트럼 차감법에 근거한 음성 향상 기술은 신호 대 잡음비에 따른 스펙트럼 이득으로 표현된다. 대표적인 Ephraim과 Malah의 decision-directed (DD) 추정치는 잡음 구간에서 효율적으로 뮤지컬 잡음을 제거하지만 음성 구간에서는 이전 프레임의 음성 스펙트럼 성분에 더 큰 비중을 두기 때문에 a priori SNR의 프레임 지연이 발생한다. 따라서 DD에 의해 추정된 a priori SNR이 적용된 잡음 제거 이득은 현재 프레임보다 이전 프레임에 영향을 받으므로 음성 전이 구간에서 잡음 제거 성능을 저하시킨다. 본 논문은 DD의 가중치 파라미터에 Sigmoid Type의 함수를 적용하여 계산적으로는 간단하지만 효과적인 음성 향상 알고리즘을 제안한다. 제안된 접근 방식은 DD의 주요 파라미터인 a priori SNR 지연의 문제점을 해결하면서 뮤지컬 잡음 제거에 우수한 DD의 이점은 유지한다. 제안된 알고리즘의 성능은 다양한 잡음 환경에서 ITU-T P.862 Perceptual Evaluation of Speech Quality (PESQ) 와 Mean Opinion Score (MOS). 그리고 음성 스펙트로그램 (Spectrogram)에 의해 평가했고 기존의 DD의 고정된 가중치 파라미터를 사용했을 때 보다 향상된 결과를 나타내었다.
https://doi.org/10.7776/ASK.2006.25.8.383 인용 PDF KSCI

사이코어쿠스틱스 모델을 이용한 음성 향상 (Speech enhancement using psychoacoustics model)

권철현;신대규;박상희
- 대한전기학회:학술대회논문집
- /
- 대한전기학회 1999년도 추계학술대회 논문집 학회본부 B
- /
- pp.748-750
- /
- 1999
In this study, a speech enhancement is presented based on the utilization of well-known auditory mechanism, noise masking. The speech enhancement approach adopted here is to derive an modifier that achieves audible noise suppression. This modification selectively affects the perceptually significant spectral values, and is therefore less prone to introduction of unwanted distortions than methods that affect the complete STSA and produces more enhanced results at low SNR as well as at high SNR. The speech enhancement method adopted here needs exact estimation of the minimum specteal value per critical band because it uses only the minimum spectral value per critical band. For this, the method adopted here uses the modified spectral subtraction that is more flexible than power spectral subtraction. So, the result in experiment represented better SNR than before.
PDF

3.0T 자기공명영상을 이용한 유방 검사시 IDEAL기법의 유용성 평가 (Evaluation of Usefulness of IDEAL(Iterative decomposition of water and fat with echo asymmetry and least squares estimation) Technique in 3.0T Breast MRI)

조재환
- 디지털콘텐츠학회 논문지
- /
- 제11권2호
- /
- pp.217-224
- /
- 2010
유방암중 관상피내암으로 진단 받은 환자를 대상으로 기존의 지방 억제 기법인 CHESS와 새로운 기법인 IDEAL을 정량적으로 비교 분석하여 IDEAL기법의 효과와 유용성을 고찰 해보고자 한다. 조직학적으로 관상피 내암으로 진단 받은 환자 20명을 대상으로 3.0T MR scanner를 이용하여 CHESS 기법과 IDEAL 기법을 이용하여 지방 억제한 횡이완 강조 영상과 조영 증강 전후의 종이완 강조 영상을 획득하였다. 분석 결과 횡이완 강조 영상과 조영 증강 전, 후의 종이완 강조 영상에서 신호대 잡음비는 병변 부위에서는 차이를 보이지 않는 반면 유관조직과 지방조직에서는 IDEAL 기법을 이용한 그룹에서 높은 신호대 잡음비를 보였으며 두 그룹에서의 대조도대 잡음비는 IDEAL 기법을 이용한 그룹에서 높은 대조도대 잡음비를 보였다.
PDF KSCI

결정지향 SNR 추정방식에서의 추정오차 보정기법을 통한 SNR 추정성능개선 (Performance Enhancement of Decision Directed SNR Estimation by Correction Scheme of SNR Estimation Error)

곽재민
- 한국항행학회논문지
- /
- 제16권6호
- /
- pp.982-987
- /
- 2012
본 논문에서는 AWGN 채널에서 특정 판정영역 내에 수신된 샘플들을 이용해 DD(Decision Directed) 방식으로 SNR을 추정하는 경우에 발생하는 SNR 추정 오차에 대해서 분석하였다. 수신기에서 특정 성좌점에 해당하는 기준 판정영역에 수신된 샘플들로 이상적인 수신점과의 에러벡터를 이용하여 SNR을 추정하는 경우, 다른 판정영역에 대응되는 송신심볼이 잡음의 영향으로 기준 판정영역으로 넘어온 샘플들까지 포함되므로, 추정된 신호 성좌점의 평균치가 이동함으로써 DD방식의 SNR 추정이 부정확하게 이루어진다. 이러한 현상을 변형된 확률밀도 함수를 기반으로 설명하고 실제 SNR과 추정 SNR과의 오차를 유도하여 정량적으로 분석하였다. 또한 컴퓨터 시뮬레이션을 통한 SNR 추정오차가 이론적으로 유도된 SNR 추정오차와 일치하고, 제안한 보정기법을 통해 SNR 추정 성능을 크게 개선할 수 있음을 확인하였다.
https://doi.org/10.12673/jkoni.2012.16.6.982 인용 PDF KSCI

효과적인 복소 스펙트럼 기반 음성 향상을 위한 시간과 주파수 영역 손실함수 조합에 관한 연구 (A study on loss combination in time and frequency for effective speech enhancement based on complex-valued spectrum)

정재희;김우일
- 한국음향학회지
- /
- 제41권1호
- /
- pp.38-44
- /
- 2022
잡음에 오염된 음성의 명료도와 음질을 향상시키고자 음성 향상을 수행한다. 본 연구에서는 복소값 스펙트럼을 이용한 마스크기반 음성 향상에서 시간 영역 손실함수와 주파수 영역 손실함수에 따른 학습 결과를 비교하였다. 시간 영역의 음성 파형과 주파수 영역의 스펙트럼의 세부정보를 고려해 두 영역의 장점을 활용할 수 있도록 손실함수 조합에 관해 연구를 진행하였다. 시간 영역 손실함수는 Scale Invariant-Source to Noise Ratio(SI-SNR)을 이용해 계산하고, 주파수 영역 손실함수는 복소값 스펙트럼과 크기 스펙트럼을 Mean Squared Error(MSE)로 계산하여 사용하였고, sin 함수를 이용해 위상에 대한 손실함수를 계산하였다. 손실함수 조합은 시간 영역 손실함수인 SI-SNR과 각 주파수 영역 손실함수를 조합하였다. 또한 크기 값과 위상 값을 모두 고려할 수 있도록 SI-SNR과 크기 스펙트럼, 위상에 관련된 손실함수들도 조합하여 실험을 진행하였다. 음성 향상 결과는 Source-to-Distortion Ratio(SDR), Perceptual Evaluation of Speech Quality(PESQ), Short-Time Objective Intelligibility(STOI)를이용해 성능 비교 평가를 진행하였다. 음성 향상 결과를 확인해보기 위해 스펙트럼 상에서 비교를 진행하였다. TIMIT 데이터베이스를 이용한 실험 결과, 시간 영역 또는 주파수 영역 손실함수보다 SI-SNR과 크기 스펙트럼을 조합한 손실함수를 사용하여 음성 향상을 학습했을 때 가장 높은 성능을 보였다.
https://doi.org/10.7776/ASK.2022.41.1.038 인용 PDF KSCI

비정상 잡음환경에서 음질향상을 위한 적응 임계 치 알고리즘 (Adaptive Threshold for Speech Enhancement in Nonstationary Noisy Environments)

이수정;김순협
- 한국음향학회지
- /
- 제27권7호
- /
- pp.386-393
- /
- 2008
본 논문에서는 비정상 잡음환경에서 음질향상을 위한 새로운 방법을 제안한다. 정상 잡음환경에서 음질향상을 위한 잡음제거 방법으로 주파수 차감법이 잘 알려져 있다. 그러나 실제 잡음환경은 대 부분 비정상적인 특성을 나타낸다. 제안한 방법은 다양한 잡음 과 비정상 환경에서 잘 동작 할 수 있도록 적응 임계 치를 위한 자동제어 파라미터를 사용한다. 특히, 자동제어 파라미터는 a posteriori SNR을 이용한 선형함수를 적용하여 잡음레벨의 증감에 따라 적응 임계 치를 제어한다. 제안한 알고리즘은 음질향상을 위해 Hangover (HO)을 이용한 주파수 차감법과 결합한다. 알고리즘의 성능은 다양한 잡음환경에서 ITU-T P.835 signal distortion (SIG)와 segment signal to-noise ratio (SNR)로 평가하여 (HO)을 이용한 음성검출과 minimum statistics (MS) 방법에 비해 우수한 결과를 나타냈다
https://doi.org/10.7776/ASK.2008.27.7.386 인용 PDF KSCI

검색결과 190건 처리시간 0.023초

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

자세히 찾기

이미지 검색 (β)