Papers

Scalable Multiband Binaural Renderer for MPEG-H 3D Audio

International Journal
2011~2015
작성자
이진영
작성일
2015-08-01 22:05
조회
1746
Authors : Taegyu Lee, Hyun Oh Oh, Jeongil Seo, Young-Cheol Park, Dae Hee Youn

Year : 2015

Publisher / Conference : IEEE Journal of Selected Topics in Signal Processing

Volume : 9, issue 5

Page : 907-920

To provide immersive 3D multimedia service, MPEG has launched MPEG-H, ISO/IEC 23008, “High Efficiency Coding and Media Delivery in Heterogeneous Environments.” As part of the audio, MPEG-H 3D Audio has been standardized based on a multichannel loudspeaker configuration (e.g., 22.2). Binaural rendering is a key application of 3D audio; however, previous studies focus on binaural rendering with low complexity such as IIR filter design for HRTF or pre-/post-processing to solve in-head localization or front-back confusion. In this paper, a new binaural rendering algorithm is proposed to support the large number of input channel signals and provide high-quality in terms of timbre, parts of this algorithm were adopted into the MPEG-H 3D Audio. The proposed algorithm truncates binaural room impulse response at mixing time, the transition point from the early-reflections to the late reverberation part. Each part is processed independently by variable order filtering in frequency domain (VOFF) and parametric late reverberation filtering (PLF), respectively. Further, a QMF domain tapped delay line (QTDL) is proposed to reduce complexity in the high-frequency band, based on human auditory perception and codec characteristics. In the proposed algorithm, a scalability scheme is adopted to cover a wide range of applications by adjusting the threshold of mixing time. Experimental results show that the proposed algorithm is able to provide the audio quality of a binaural rendered signal using full-length binaural room impulse responses. A scalability test also shows that the proposed scalability scheme smoothly compromises between audio quality and computational complexity.
전체 364
254 International Conference Eunwoo Song, Frank K. Soong, Hong-Goo Kang "Improved Time-Frequency Trajectory Excitation Vocoder for DNN-Based Speech Synthesis" in INTERSPEECH, 2016
253 Domestic Conference Hyeonjoo Kang, Young-sun Joo, Wonsuk Jun, Hong-goo Kang "다층신경망 기반 다중 화자 음성변환 시스템" in 한국음향학회 제 33회 음성통신 및 신호처리 학술대회, 2016
252 International Conference Eunwoo Song, Hong-Goo Kang "Multi-class learning algorithm for deep neural network-based statistical parametric speech synthesis" in EUSIPCO, 2016
251 Domestic Conference Min-jae Hwang, JeeSok Lee, Misuk Lee, and Hong-Goo Kang "사전 분석법을 통한 스프레드 스펙트럼 기반 오디오 워터마킹 알고리즘의 성능 향상" in 한국음향학회 제 33회 음성통신 및 신호처리 학술대회, 2016
250 Domestic Conference 문현기, 박영철, 윤대희 "위상 일치와 가변 지수 감쇄 가중치 방법이 적용된 가상 저음 시스템" in 2016 년 한국방송·미디어공학회 하계학술대회, 2016
249 Domestic Conference 문현기, 박영철, 윤대희 "주파수 종속 믹싱 타임 추정 기법 분석" in 한국음향학회 춘계학술대회, 2016
248 Domestic Conference 서지호, 박영철, 윤대희 "헤드폰/이어폰 환경에서의 제약 최적화 기반의 효율적인 피드백 능동 소음 제어 필터 설계 알고리즘" in 한국음향학회 춘계학술대회, 2016
247 Domestic Conference 박규태, 박영철, 윤대희 "가변 가중치 곡선을 적용한 가상저음시스템" in 한국음향학회 춘계학술대회, 2016
246 Domestic Conference 양해민, 변경근, 강홍구 "RTCP를 이용한 심층 신경망 기반 음질평가 점수 대역 분별 알고리즘" in 한국음향학회 춘계학술대회, 2016
245 Domestic Conference 김글빛, 이진규, 강홍구 "문장종속 화자검증 시스템을 위한 비음수 행렬 분해 기반 잡음 제거" in 한국음향학회 춘계학술대회, 2016