Learning a Precedence Effect-Like Weighting Function for the Generalized Cross-Correlation Framework

Authors:
K. W. Wilson;T. Darrell
Affiliations:
Lab. of Comput. Sci. & Artificial Intelligence, Massachusetts Inst. of Technol., Cambridge, MA;-
Venue:
IEEE Transactions on Audio, Speech, and Language Processing
Year:
2006

Citing 0
Cited 2

3D-audio matting, postediting, and rerendering from field recordings

EURASIP Journal on Applied Signal Processing
Sequential organization of speech in reverberant environments by integrating monaural grouping and binaural localization

IEEE Transactions on Audio, Speech, and Language Processing - Special issue on processing reverberant speech: methodologies and applications

Quantified Score

Hi-index	0.00

Visualization

Abstract

Speech source localization in reverberant environments has proved difficult for automated microphone array systems. Because of its nonstationary nature, certain features observable in the reverberant speech signal, such as sudden increases in audio energy, provide cues to indicate time-frequency regions that are particularly useful for audio localization. We exploit these cues by learning a mapping from reverberated signal spectrograms to localization precision using ridge regression. Using the learned mappings in the generalized cross-correlation framework, we demonstrate improved localization performance. Additionally, the resulting mappings exhibit behavior consistent with the well-known precedence effect from psychoacoustic studies