통합 검색 | Korea Science

LPC 계수의 최적 양자화에 기초한 음성 코더 구현 (Implementation of a CELP coder based on optimum quantization of the LPC coefficients)

이우종;박지태;장태규
- 대한전기학회:학술대회논문집
- /
- 대한전기학회 2001년도 하계학술대회 논문집 D
- /
- pp.2516-2518
- /
- 2001
The quantization of the LPC parameters is a very important aspect of the speech compression algorithm. This paper analyzes the quantization effect of the LPC coefficients and presents the implementation of a fixed-point CELP coder based on the LPC analysis.
PDF

Multi-frame AR model을 이용한 LPC 계수 양자화 (Quantization of LPC Coefficients Using a Multi-frame AR-model)

정원진;김무영
- 한국음향학회지
- /
- 제31권2호
- /
- pp.93-99
- /
- 2012
음성코딩 시 성도는 Linear Predictive Coding (LPC) 계수를 이용해서 모델링 한다. 일반적으로 LPC 계수는 양자화와 선형보간 관점에서 유리한 Line Spectral Frequency (LSF) 파라미터로 변경하여 사용한다. 10차 이상의 다차원 LSF 데이터를 벡터 양자화를 이용하여 직접 코딩하게 되면 벡터 내 상관관계 (intra-frame correlation)를 모두 이용할 수 있으므로 rate-distortion 관점에서는 높은 효율을 기대할 수 있다. 하지만, 계산량과 메모리 요구량이 높아져서 실제 코딩 시스템에서는 사용할 수 없게 되므로, 차원을 나누어 압축하는 Split Vector Quantization (SVQ)이 이용된다. 또한, LSF 데이터는 과거 벡터와의 벡터 간 상관관계 (inter-frame correlation)가 높으므로, 이를 이용한 Predictive Split Vector Quantization (PSVQ)이 사용되고 있다. PSVQ는 SVQ 보다 높은 rate-distortion 성능을 보인다. 본 논문에서는 음성 저장 장치를 위한 최적의 PSVQ를 구현하기 위해서 다수의 과거 프레임 정보와의 벡터 간상관관계 (inter-frame correlation)를 고려한 Multi-Frame AR-model 기반 SVQ (MF-AR-SVQ)를 제안하였다. 기존 PSVQ와 비교해 보았을 때, MF-AR-SVQ는 계산량과 메모리 요구량의 큰 증가 없이, 평균 spectral distortion 관점에서 약 1비트의 성능 향상을 보였다.
https://doi.org/10.7776/ASK.2012.31.2.093 인용 PDF KSCI

LPC 벡터 양자화를 이용한 가변률 CELP 음성코딩에 관한 연구 (Variable Rate CELP Coding with Phonetic Segmentation using LPC Vector Quantization)

정영호
- 한국음향학회:학술대회논문집
- /
- 한국음향학회 1994년도 제11회 음성통신 및 신호처리 워크샵 논문집 (SCAS 11권 1호)
- /
- pp.205-209
- /
- 1994
This paper presents a variable rate speech coding method with phonetic segmentation, called for PSVXC. Multiple access techniques that require efficient encoding of speech to achieve capacity improvements are currently emerging in the cellular telephone system. The variable rate speech coder have the reduced average data rate required to transmit conversational speech. Each frame of active speech is classified into one of four phonetic classes. A distinct coding configuration and bit-rate is applied to each category. And also a split vector quantization is used to accurately quantize the LPC information using LSP parameters.
PDF

이동통신 음성 부화화기를 위한 선형 예측 계수(LPC)의 효율적 양자화 방법 (Efficient quantization of LPC parameters for vocoder of mobile communications)

이인성;우홍채
- 전자공학회논문지S
- /
- 제34S권4호
- /
- pp.50-56
- /
- 1997
In this paper, efficient quantization methods of line spectrum pairs (LSP) which has good performances and low complexity and memory are proosed for vocoder of mobile communication system. The adaptive quantization method utilizing the ordering property of LSP parameters is used in a scalar quantizer and a vector-scalar hybrid quantizer. The proposed scalar quantization algorithm needs 31 bits/frame to maintain the transparent quality of speech. The improved vector-scalar quantizer achieves an average spectral distortion of 1dB using 26 bits/frame. The proposed methods are evaluated in the channel errors and changed the predictor structure to maintain the robustness to channel errors.
PDF

성도 면적 함수를 이용한 음성 인식에 관한 연구 (A Study on Speech Recognition using Vocal Tract Area Function)

송제혁;김동준
- 대한의용생체공학회:의공학회지
- /
- 제16권3호
- /
- pp.345-352
- /
- 1995
The LPC cepstrum coefficients, which are an acoustic features of speech signal, have been widely used as the feature parameter for various speech recognition systems and showed good performance. The vocal tract area function is a kind of articulatory feature, which is related with the physiological mechanism of speech production. This paper proposes the vocal tract area function as an alternative feature parameter for speech recognition. The linear predictive analysis using Burg algorithm and the vector quantization are performed. Then, recognition experiments for 5 Korean vowels and 10 digits are executed using the conventional LPC cepstrum coefficients and the vocal tract area function. The recognitions using the area function showed the slightly better results than those using the conventional LPC cepstrum coefficients.
PDF

DMS 모델과 이중 스펙트럼 특징을 이용한 HMM에 의한 음성 인식 (HMM-based Speech Recognition using DMS Model and Double Spectral Feature)

안태옥
- 한국산학기술학회논문지
- /
- 제7권4호
- /
- pp.649-655
- /
- 2006
본 논문은 화자 독립의 음성인식을 위한 연구로써, DMS 모델에 의한 DMSVQ(Dynamic Multi-Section Vector Quantization) 코드북과 이중 스펙트럼 특징을 이용한 HMM(Hidden Markov Model) 음성인식 방법을 제안한다. 정적 스펙트럼 특징으로서는 LPC ?S스트럼 계수를 이용하였고, 동적 스펙트럼 특징으로는 LPC ?S스트럼의 회귀계수를 사용하였다. 이들 두개의 스펙트럼 특징들을 각각 VQ 코드북으로 양자화되고, DMS 모델을 이용한 HMM은 입력으로써 정적 스펙트럼 특징과 동적 스펙트럼 특징을 받아드림으로써 모델링된다. 제안된 방법에 의한 인식 실험은 기존의 다양한 인식 방법에 의한 인식 실험들과 비교를 위해 동일한 데이터와 조건 하에서 수행하였다. 실험 결과, 본 연구에서 제안한 방법이 기존의 방법들보다 우수한 방법임을 입증하였다.
PDF

신경 회로망을 이용한 음성 신호의 벡터 양자화 (Speech Signal Vector Quantization Using Neural Network)

백승복;김상희
- 대한전자공학회:학술대회논문집
- /
- 대한전자공학회 1999년도 추계종합학술대회 논문집
- /
- pp.1015-1018
- /
- 1999
This paper describes a vector quantization for speech signal coding using neural networks. We processed speech signal using LPC method that extracts speech signal feature, and speech signal feature is quantized using competitive neural network kohonen self-organization feature map.
PDF

LPC Cepstral 벡터 양자화에 의한 저 전송율 CELP 음성부호기의 스펙트럼 표기 (Spectrum Representation Based on LPC Cepstral VQ for Low Bit Rate CELP Coder)

정재호
- 한국통신학회논문지
- /
- 제19권4호
- /
- pp.761-771
- /
- 1994
본 논문에서는, 매우 낮은 전송율이 요구되는 음성통신의 환경하에서 CELP 음성 부호기를 사용할 경우, 스펙트럼에 대한 정보를 어떻게 효과적으로 나타낼 것인가에 대하여 고찰하였다. 구체적으로, 스펙트럼에 대한 정보를 나타내는 LPC 파라메타를 cepstrum으로 변형시키고, 변형된 LPC cepstrum계수들을 효과적으로 벡터 양자화하는 방법을 제시하였다. 벡터 양자화에 사용되는 코드-북의 설계를 위하여, 주파수 대역에서 서로 다른 의미를 갖는 세계의 cepstral distance measure들을 시도하였으며, 각각에 대한 성능이 분석되어졌다. 시뮬레이션을 통하여, 본 논문에서 제시한 LPC cepstral 벡터 양자화 방식이 스펙트럼에 대한 정보를 매우 효과적으로 나타낼 수 있음을 보였다.
PDF

블록 제한 트렐리스 부호화 양자화 기법을 이용한 협대역 음성 부호화기용 LPC 계수 양자화기 설계 (Designing a Quantizer of LPC Parameters for the Narrowband Speech Coder using Block-Constrained Trellis Coded Quantization)

전자경;박상국;강상원
- 한국통신학회논문지
- /
- 제32권3C호
- /
- pp.234-240
- /
- 2007
본 논문에서는 기존의 트렐리스 부호화 양자화 기법을 이용, 변형하여 저 복잡도 블록 제한 격자 부호화 양자화 기법 (Block-Constrained Trellis Coded Quantization, 이하 BC-TCQ)을 제안하곤 이를 이용한 협대역 음성 부호화기용 예측 BC-TCQ를 설계하였다. 트렐리스 부호화 양자화 기법은 일종의 벡터 양자화 방식으로 부호화에 요구되는 벡터 코드북을 트렐리스 구조에 기반한 스칼라 코드북으로 구성함으로써 VQ와 비교 할 만한 성능을 보일 뿐 아니라 복잡도가 훨씬 작은 특성을 보인다. 본 논문에서 제안한 예측 BC-TCQ는 프레임당 26비트에서 IS-641 음성 부호화기보다 평균 SD가 0.4107dB 향상되었으며, 더하기 연산이 64.54%, 곱하기 연산이 76.93%, 비교 연산이 2.35% 감소하였다.
PDF KSCI

A Line Spectrum Frequency Pairs Representation for Spectral Envelop Quantization

Park, Youngho;Lee, Won-Cheol;Bae, Myung-Jin
- 대한전자공학회:학술대회논문집
- /
- 대한전자공학회 2000년도 제13회 신호처리 합동 학술대회 논문집
- /
- pp.787-790
- /
- 2000
This paper introduces a new type of representation of the LSPs as a promising alternative used for transmitting the LPC parameters. Major contribution in this paper is that the vocal track information embedded on the spectral envelope can be represented in terms of the reduced number of LSF compared tn the conventional. Hence, it provides a possibility that LPC parameters could be quantized at a reduced bit rate without causing any major spectral distortion. The simulation result illustrates the capability of the proposed LSPs representation as an efficient quantization method via a proper rejection of the redundant pairs of pole and zero along the unit circle.
PDF

검색결과 28건 처리시간 0.023초

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

자세히 찾기

이미지 검색 (β)