logo Idiap Research Institute        
 [BibTeX] [Marc21]
A Sector-Based Approach for Localization of Multiple Speakers with Microphone Arrays
Type of publication: Conference paper
Citation: lathoud04b
Booktitle: Proceedings of the 2004 SAPA Workshop
Year: 2004
Month: 10
Address: Jeju Island, Korea
Note: IDIAP-RR 04-15
Crossref: lathoud-rr-04-15:
Abstract: Microphone arrays are useful in meeting rooms, where speech needs to be acquired and segmented. For example, automatic speech segmentation allows enhanced browsing experience, and facilitates automatic analysis of large amounts of data. Spontaneous multi-party speech includes many overlaps between speakers; moreover other audio sources such as laptops and projectors can be active. For these reasons, locating multiple wideband sources in a reasonable amount of time is highly desirable. In existing multisource localization approaches, search initialization is very often an issue left open. We propose here a methodology for estimating speech activity in a given sector of the space rather than at a particular point. In experiments on more than one hour of speech from real meeting room multisource recordings, we show that the sector-based greatly reduces the search space. At the same time, it achieves effective localization of multiple concurrent speakers.
Userfields: ipdinar={2004}, ipdmembership={speech},
Keywords:
Projects Idiap
Authors Lathoud, Guillaume
McCowan, Iain A.
Added by: [UNK]
Total mark: 0
Attachments
  • lathoud04b.pdf
  • lathoud04b.ps.gz
Notes