Search | Korea Science

An Audio Coding Technique Employing the Inter-channel Phase Difference Skip (채널 간 위상차 파라미터 생략 기법을 이용한 오디오 부호화)

Kim, Hyun-Hwi;Kim, Rin-Chul
- Journal of Broadcast Engineering
- /
- v.21 no.3
- /
- pp.369-379
- /
- 2016
This paper deals with an efficient method for skipping inter-channel phase differences (IPD) in the MPEG surround of the unified speech and audio coding (USAC). Based on the psycho-acoustic sensitivity on the IPD, we estimate a threshold on IPD, below which we can not notice degradation in spatial cue. We propose an IPD skip method, in which any IPDs within the threshold are set to zero and are not transmitted. The proposed IPD skip method gives about 38% savings in terms of bit amount for IPD. Nevertheless, in the MUSHRA test, the proposed method does not show any noticeable degradation in the decoded audio quality.
https://doi.org/10.5909/JBE.2016.21.3.369 인용 PDF KSCI KPUBS HTML

Noise-Biased Compensation of Minimum Statistics Method using a Nonlinear Function and A Priori Speech Absence Probability for Speech Enhancement (음질향상을 위해 비선형 함수와 사전 음성부재확률을 이용한 최소통계법의 잡음전력편의 보상방법)

Lee, Soo-Jeong;Lee, Gang-Seong;Kim, Sun-Hyob
- The Journal of the Acoustical Society of Korea
- /
- v.28 no.1
- /
- pp.77-83
- /
- 2009
This paper proposes a new noise-biased compensation of minimum statistics(MS) method using a nonlinear function and a priori speech absence probability(SAP) for speech enhancement in non-stationary noisy environments. The minimum statistics(MS) method is well known technique for noise power estimation in non-stationary noisy environments. It tends to bias the noise estimate below that of true noise level. The proposed method is combined with an adaptive parameter based on a sigmoid function and a priori speech absence probability (SAP) for biased compensation. Specifically. we apply the adaptive parameter according to the a posteriori SNR. In addition, when the a priori SAP equals unity, the adaptive biased compensation factor separately increases ${\delta}_{max}$ each frequency bin, and vice versa. We evaluate the estimation of noise power capability in highly non-stationary and various noise environments, the improvement in the segmental signal-to-noise ratio (SNR), and the Itakura-Saito Distortion Measure (ISDM) integrated into a spectral subtraction (SS). The results shows that our proposed method is superior to the conventional MS approach.
https://doi.org/10.7776/ASK.2009.28.1.077 인용 PDF KSCI

Development of a Diphone-Based Audiote System (다이폰단위의 합성방법을 이용한 오디오텍스 시스템의 구현에 관한 연구)

이승훈
- Proceedings of the Acoustical Society of Korea Conference
- /
- 1994.06c
- /
- pp.99-102
- /
- 1994
당 연구실에서 개발했던 초기의 오디오텍스 시스템은 LSP 파라미터를 이용한 무제한 한국어 음성합성 장치로서 합성데이타베이스는 640개의 반음절로 구성되어 있었다. 그러나 이 시스템은 일반 사용자들에게 음성합성 서비스를 제공하기에는 damwlf이 너무 미흡하였으므로 음원모델의 수정, 에너지 contour의 조절등을 사용하여 어느 정도 음질개선을 꾀하였으나 만족할 만한 수준에는 도달하지 못했다. 그래서 합성단위를 다이폰단위로 수정한 새로운 오디오텍스 시스템을 ngus하였다. 다이폰단위의 오디오텍스시스템은 한국어의여러가지 음운환경을 고려하여 1228개의 합성단위로 구성되어 있으며 LSP 파라미터를 이용한 합성방식을 채택하고 있다. 또한 음원생성시 수정된 LF 모델에 자음의 명료도 및 자연성을 높이기 위해 TMS320C30 DSP chip, MC68020 CPU, 고속 메모리소자, 및 VRTOS를 사용하여 시스템을 구현하였으며, 청취실험결과 기존의 합성방법보다 자연성 및 명료도에서 개선된 음질을 얻을 수 있었다.
PDF

A Study on a comparison and analysis of Speaking rate estimation for adaptive bit rate on CELP vocoder (가변전송률 CELP 부호화기 설계를 위한 발성률 비교 분석에 관한 연구)

Jang KyungA;Min SoYeon;Bae MyungJin
- Proceedings of the Acoustical Society of Korea Conference
- /
- spring
- /
- pp.105-108
- /
- 2004
음성 부호화 기술은 전송률과 복잡도를 줄이고 음질을 향상시키는 방향으로 진행되고 있다. 현재 상용화되고 있는 CELP형 보코더는 낮은 전송률에 비해 우수한 음질을 제공한다. 본 논문에서는 기존의 방식과 다르게 보코더 단에 입력 음성이 들어가기 앞서 전처리 기법을 수행하는 전처리단을 부가하여 전송률을 낮추는 방법을 소개하고, 소개된 방법들을 각기 비교하고 분석하고자 한다. 전처리기법들을 음성 인식이나 합성에서 사용되는 파라미터들을 적용시켰으며, 처리시간이나 계산시간에 있어 기존의 방식에서 많은 영향을 미치지 않은 간단한 알고리즘으로 구현하였다. 소개하는 전처리단에서는 기존의 코딩방식에서 사용하지 않은 파라미터들, 발성율, 지속시간, PSOLA 방식들을 이용하였다.
PDF

Implementation of Demisyllable database for formant synthesizer (포만트 합성기용 반음절 세트의 구축에 관한 연구)

이정석
- Proceedings of the Acoustical Society of Korea Conference
- /
- 1992.06a
- /
- pp.81-84
- /
- 1992
포만트형 합성기에 사용될 반음절 데이터 베이스의 구성과 필요한 파라미터의 추출 과정에 대하여 논한다. 포만트 합성기는 많은 구동 파라미터를 필요로 하기 때문에 저장 장소를 절약하기 위해서 적절한 합성단위의 선택과 합성단위의 효율적인 표현이 필요하다. 본 연구에서는 포만트 합성기에 있어서 합성음의 음질에 큰 영향을 미치는 포만트궤적의 추출과 데이터베이스의 구성에 대하여 기술한다.
PDF

Speech Enhancement using the Neural Network Filter (신경망필터를 이용한 음질향상)

김종우;공성곤
- Proceedings of the Korean Institute of Intelligent Systems Conference
- /
- 2000.05a
- /
- pp.102-105
- /
- 2000
본 논문에서는 잡음환경에서의 음성신호복원(Speech Enhancement) 시스템 구현을 목적으로 한다 이를 위한 적응필터로서 LMS(Least Mean Square)알고리즘 FIR필터를 제안한다. 또 정밀 필터로서 신경망 필터를 제안한다. 잡음환경에서의 음성신호 복원 시스템은 잡음에 의해 왜곡된 음성신호에서 잡음성분만을 제거함으로써 음성신호를 복원하는 시스템이다. 일반적으로 잡음은 시변특성과, 비선형적인 전달특성을 갖는다. 그러므로 파라미터가 고정된 필터로는 제어하기가 힘들다. 이러한 이유로 본 논문에서는 LMS알고리즘 적응필터를 적용한다. 신경망 필터는 오차 역전파 학습 알고리즘에 의해 오차를 최소화하는 방향으로 필터의 파라미터를 수정한다. 제안한 필터로 잡음환경에서의 음성신호복원 시스템을 구성하고, 실험을 통해 필터의 성능을 확인한다.
PDF

Speech Quality Estimation Algorithm using a Harmonic Modeling of Reverberant Signals (반향 음성 신호의 하모닉 모델링을 이용한 음질 예측 알고리즘)

Yang, Jae-Mo;Kang, Hong-Goo
- Journal of Broadcast Engineering
- /
- v.18 no.6
- /
- pp.919-926
- /
- 2013
The acoustic signal from a distance sound source in an enclosed space often produces reverberant sound that varies depending on room impulse response. The estimation of the level of reverberation or the quality of the observed signal is important because it provides valuable information on the condition of system operating environment. It is also useful for designing a dereverberation system. This paper proposes a speech quality estimation method based on the harmonicity of received signal, a unique characteristic of voiced speech. At first, we show that the harmonic signal modeling to a reverberant signal is reasonable. Then, the ratio between the harmonically modeled signal and the estimated non-harmonic signal is used as a measure of standard room acoustical parameter, which is related to speech clarity. Experimental results show that the proposed method successfully estimates speech quality when the reverberation time varies from 0.2s to 1.0s. Finally, we confirm the superiority of the proposed method in both background noise and reverberant environments.
https://doi.org/10.5909/JBE.2013.18.6.919 인용 PDF KSCI KPUBS HTML

Adaptive Threshold for Speech Enhancement in Nonstationary Noisy Environments (비정상 잡음환경에서 음질향상을 위한 적응 임계 치 알고리즘)

Lee, Soo-Jeong;Kim, Sun-Hyob
- The Journal of the Acoustical Society of Korea
- /
- v.27 no.7
- /
- pp.386-393
- /
- 2008
This paper proposes a new approach for speech enhancement in highly nonstationary noisy environments. The spectral subtraction (SS) is a well known technique for speech enhancement in stationary noisy environments. However, in real world, noise is mostly nonstationary. The proposed method uses an auto control parameter for an adaptive threshold to work well in highly nonstationary noisy environments. Especially, the auto control parameter is affected by a linear function associated with an a posteriori signal to noise ratio (SNR) according to the increase or the decrease of the noise level. The proposed algorithm is combined with spectral subtraction (SS) using a hangover scheme (HO) for speech enhancement. The performances of the proposed method are evaluated ITU-T P.835 signal distortion (SIG) and the segment signal to-noise ratio (SNR) in various and highly nonstationary noisy environments and is superior to that of conventional spectral subtraction (SS) using a hangover (HO) and SS using a minimum statistics (MS) methods.
https://doi.org/10.7776/ASK.2008.27.7.386 인용 PDF KSCI

Transcoding Algorithm for SMV and G.729A Vocoders via Direct Parameter Transformation (G.729A와 SMV 음성부호화기를 위한 파라미터 직접 변환 방식의 상호부호화 알고리듬)

장달원;서성호;이선일;유창동
- Journal of the Institute of Electronics Engineers of Korea SP
- /
- v.40 no.6
- /
- pp.71-83
- /
- 2003
In this paper, a novel transcoding algorithm for the G.729A and the Selectable Mode Vocoder(SMV) vocoders via direct parameter transformation is proposed. In contrast to the conventional tandem transcoding algorithm, the proposed algorithm converts the parameters of one coder to the other without going through the decoding and encoding processes. In transcoder from SMV to G.729A, LSP conversion algorithm, pitch delay conversion algorithm and transcoding algorithm in lower rate are proposed, and in transcoder from G.729A to SMV, LSP conversion algorithm, pitch delay conversion algorithm and rate selection algorithm are proposed. Evaluation results show that while exhibiting better computational and delay characteristics, the proposed algorithm produces equivalent or Improved speech quality to that produced by the tandem transcoding algorithm.
PDF KSCI

A Study on Tape Transport Characteristics of Belt Driven System (벨트구동계의 동특성 해석을 통한 주행특성 분석)

유진형;김남응;주관정
- Proceedings of the Korean Society for Noise and Vibration Engineering Conference
- /
- 1994.10a
- /
- pp.205-209
- /
- 1994
본 연구에서는 오디오용 데크의 음질특성과 벨트구동계의 회전진동 특성과의 연관성을 규명하고자 한다. 이를 위해 벨트구동계를 다자유도 진동계로 모델링하여 회전진동 특성을 규명하고 실험으로 확인하였다. 이의 결과를 실제 데크에 응용하기 위해 각 파라미터의 변화에 따른 특성변화의 경향을 살펴보았으며, 당사에서 개발중인 신모델에서 회전비의 적절한 설계를 제안하여 우수한 음질특성을 확인하였다.
PDF

Search Result 70, Processing Time 0.03 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)