Single-Channel Talker Localization Based on Discrimination of Acoustic Transfer Functions
This paper presents a sound source (talker) localization method using only a single microphone, where a Gaussian Mixture Model (GMM) of clean speech is introduced to estimate the acoustic transfer function from a user’s position. The new method is able to carry out this estimation without...
Main Authors: | , , , |
---|---|
Format: | Article |
Language: | English |
Published: |
SpringerOpen
2009-01-01
|
Series: | EURASIP Journal on Advances in Signal Processing |
Online Access: | http://dx.doi.org/10.1155/2009/918404 |
_version_ | 1811284359799373824 |
---|---|
author | Tetsuya Takiguchi Yuji Sumida Ryoichi Takashima Yasuo Ariki |
author_facet | Tetsuya Takiguchi Yuji Sumida Ryoichi Takashima Yasuo Ariki |
author_sort | Tetsuya Takiguchi |
collection | DOAJ |
description | This paper presents a sound source (talker) localization method using only a single microphone, where a Gaussian Mixture Model (GMM) of clean speech is introduced to estimate the acoustic transfer function from a user’s position. The new method is able to carry out this estimation without measuring impulse responses. The frame sequence of the acoustic transfer function is estimated by maximizing the likelihood of training data uttered from a given position, where the cepstral parameters are used to effectively represent useful clean speech. Using the estimated frame sequence data, the GMM of the acoustic transfer function is created to deal with the influence of a room impulse response. Then, for each test dataset, we find a maximum-likelihood (ML) GMM from among the estimated GMMs corresponding to each position. The effectiveness of this method has been confirmed by talker localization experiments performed in a room environment. |
first_indexed | 2024-04-13T02:27:35Z |
format | Article |
id | doaj.art-5919de5a265246ee87a956515f52ed6f |
institution | Directory Open Access Journal |
issn | 1687-6172 1687-6180 |
language | English |
last_indexed | 2024-04-13T02:27:35Z |
publishDate | 2009-01-01 |
publisher | SpringerOpen |
record_format | Article |
series | EURASIP Journal on Advances in Signal Processing |
spelling | doaj.art-5919de5a265246ee87a956515f52ed6f2022-12-22T03:06:43ZengSpringerOpenEURASIP Journal on Advances in Signal Processing1687-61721687-61802009-01-01200910.1155/2009/918404Single-Channel Talker Localization Based on Discrimination of Acoustic Transfer FunctionsTetsuya TakiguchiYuji SumidaRyoichi TakashimaYasuo ArikiThis paper presents a sound source (talker) localization method using only a single microphone, where a Gaussian Mixture Model (GMM) of clean speech is introduced to estimate the acoustic transfer function from a user’s position. The new method is able to carry out this estimation without measuring impulse responses. The frame sequence of the acoustic transfer function is estimated by maximizing the likelihood of training data uttered from a given position, where the cepstral parameters are used to effectively represent useful clean speech. Using the estimated frame sequence data, the GMM of the acoustic transfer function is created to deal with the influence of a room impulse response. Then, for each test dataset, we find a maximum-likelihood (ML) GMM from among the estimated GMMs corresponding to each position. The effectiveness of this method has been confirmed by talker localization experiments performed in a room environment.http://dx.doi.org/10.1155/2009/918404 |
spellingShingle | Tetsuya Takiguchi Yuji Sumida Ryoichi Takashima Yasuo Ariki Single-Channel Talker Localization Based on Discrimination of Acoustic Transfer Functions EURASIP Journal on Advances in Signal Processing |
title | Single-Channel Talker Localization Based on Discrimination of Acoustic Transfer Functions |
title_full | Single-Channel Talker Localization Based on Discrimination of Acoustic Transfer Functions |
title_fullStr | Single-Channel Talker Localization Based on Discrimination of Acoustic Transfer Functions |
title_full_unstemmed | Single-Channel Talker Localization Based on Discrimination of Acoustic Transfer Functions |
title_short | Single-Channel Talker Localization Based on Discrimination of Acoustic Transfer Functions |
title_sort | single channel talker localization based on discrimination of acoustic transfer functions |
url | http://dx.doi.org/10.1155/2009/918404 |
work_keys_str_mv | AT tetsuyatakiguchi singlechanneltalkerlocalizationbasedondiscriminationofacoustictransferfunctions AT yujisumida singlechanneltalkerlocalizationbasedondiscriminationofacoustictransferfunctions AT ryoichitakashima singlechanneltalkerlocalizationbasedondiscriminationofacoustictransferfunctions AT yasuoariki singlechanneltalkerlocalizationbasedondiscriminationofacoustictransferfunctions |