Tracking a moving speaker using excitation source information

TitleTracking a moving speaker using excitation source information
Publication TypeConference Papers
Year of Publication2003
AuthorsRaykar VC, Duraiswami R, Yegnanarayana B, Prasanna SRM
Conference NameEighth European Conference on Speech Communication and Technology
Date Published2003///

Microphone arrays are widely used to detect, locate, and track a stationary or moving speaker. The first step is to estimate the time delay, between the speech signals received by a pair of microphones. Conventional methods like generalized cross-correlation are based on the spectral content of the vocal tract system in the speech signal. The spectral content of the speech signal is affected due to degradations in the speech signal caused by noise and reverberation. However, features corresponding to the excitation source of speech are less affected by such degradations. This paper proposes a novel method to estimate the time delays using the excitation source information in speech. The estimated delays are used to get the position of the moving speaker. The proposed method is compared with the spectrum-based approach using real data from a microphone array setup.