HMM-based separation of acoustic transfer function for single-channel sound source localization

15 years 8 months ago

Download www.me.cs.scitec.kobe-u.ac.jp

This paper presents a sound source (talker) localization method using only a single microphone, where a HMM (Hidden Markov Model) of clean speech is introduced to estimate the acoustic transfer function from a user’s position. The new method is able to carry out this estimation without measuring impulse responses. The frame sequence of the acoustic transfer function is estimated by maximizing the likelihood of training data uttered from a given position, where the cepstral parameters are used to effectively represent useful clean speech. Using the estimated frame sequence data, the GMM (Gaussian Mixture Model) of the acoustic transfer function is created to deal with the inﬂuence of a room impulse response. Then, for each test data set, we ﬁnd a maximum-likelihood GMM from among the estimated GMMs corresponding to each position. The effectiveness of this method has been conﬁrmed by talker localization experiments performed in a room environment.

Ryoichi Takashima, Tetsuya Takiguchi, Yasuo Ariki

Real-time Traffic

Acoustic Transfer Function | Clean Speech | ICASSP 2010 | Impulse Response | Signal Processing |

claim paper

Post Info
More Details (n/a)

Added	06 Dec 2010
Updated	06 Dec 2010
Type	Conference
Year	2010
Where	ICASSP
Authors	Ryoichi Takashima, Tetsuya Takiguchi, Yasuo Ariki

Comments (0)

Sciweavers

HMM-based separation of acoustic transfer function for single-channel sound source localization

Acoustic Transfer Function | Clean Speech | ICASSP 2010 | Impulse Response | Signal Processing |

Explore & Download

Productivity Tools

Sciweavers