标题: TDOA INFORMATION BASED VAD FOR ROBUST SPEECH RECOGNITION IN DIRECTIONAL AND DIFFUSE NOISE FIELD
作者: Huang, Kuan-Lang
Chi, Tai-Shih
电机资讯学士班
Undergraduate Honors Program of Electrical Engineering and Computer Science
关键字: Diffuse noise;directional interference;phase difference;time difference of arrival;voice activity detector
公开日期: 2012
摘要: A two-microphone algorithm is proposed to improve automatic speech recognition (ASR) rates when target speech is corrupted by directional interferences and diffuse noise simultaneously. The algorithm adopts the time difference of arrival (TDOA) to suppress directional interferences and a TDOA-information based voice activity detector (VAD) to suppress diffuse noise. Simulation results show the proposed algorithm is effective in improving ASR rates in a sound field mixed with a directional interference and diffuse noise. Compared with the phase difference (PD) algorithm, the proposed method gives comparable recognition rates when facing a directional interference and much higher and more robust recognition rates when diffuse noise emerges.
URI: http://hdl.handle.net/11536/21519
ISBN: 978-1-4673-2507-3
期刊: 2012 8TH INTERNATIONAL SYMPOSIUM ON CHINESE SPOKEN LANGUAGE PROCESSING
起始页: 126
结束页: 130
显示于类别:Conferences Paper