Publication

Segmenting Multiple Concurrent Speakers Using Microphone Arrays

Related publications (41)

Graph Chatbot

Chat with Graph Search

Ask any question about EPFL courses, lectures, exercises, research, news, etc. or try the example questions below.

DISCLAIMER: The Graph Chatbot is not programmed to provide explicit or categorical answers to your questions. Rather, it transforms your questions into API requests that are distributed across the various IT services officially administered by EPFL. Its purpose is solely to collect and recommend relevant references to content that you can explore to help you answer your questions.

Segmenting Multiple Concurrent Speakers Using Microphone Arrays

Guillaume Lathoud, Darren Moore

Speaker turn detection is an important task for many speech processing applications. However, accurate segmentation can be hard to achieve if there are multiple concurrent speakers (overlap), as is typically the case in multi-party conversations. In such c ...

2003

Location Based Speaker Segmentation

Guillaume Lathoud

This paper proposes a technique that segments into speaker turns based on their location, essentially implementing a discrete source tracking system. In many multi-party conversations, such as meetings or teleconferences, the location of participants is re ...

2003

Microphone Array Speech Recognition : Experiments on Overlapping Speech in Meetings

Darren Moore

This paper investigates the use of microphone arrays to acquire and recognise speech in meetings. Meetings pose several interesting problems for speech processing, as they consist of multiple competing speakers within a small space, typically around a tabl ...

2003

Audio-Visual Speaker Tracking with Importance Particle Filters

Daniel Gatica-Perez, Jean-Marc Odobez, Guillaume Lathoud, Darren Moore

We present a probabilistic methodology for audio-visual (AV) speaker tracking, using an uncalibrated wide-angle camera and a microphone array. The algorithm fuses 2-D object shape and audio information via importance particle filters (I-PFs), allowing for ...

2003

An Online Audio Indexing System

Hervé Bourlard, Jitendra Ajmera

This paper presents overview of an online audio indexing system, which creates a searchable index of speech content embedded in digitized audio files. This system is based on our recently proposed offline audio segmentation techniques. As the data arrives ...

IDIAP2003

Small Microphone Array: Algorithms and Hardware

Darren Moore

This report describes the processing algorithms and gives an overview of the hardware for the small microphone array unit in the IM2.RTMAP (Real-time Microphone Array Processing) project. The algorithms include techniques for speech enhancement, speaker lo ...

IDIAP2003

Location Based Speaker Segmentation

Guillaume Lathoud

IDIAP2002

Microphone Array Speech Recognition : Experiments on Overlapping Speech in Meetings

Darren Moore

IDIAP2002

Audio-Visual Speaker Tracking with Importance Particle Filters

Daniel Gatica-Perez, Jean-Marc Odobez, Guillaume Lathoud, Darren Moore

IDIAP2002

R/D optimal linear prediction

Martin Vetterli, Paolo Prandoni

A common technique to extend linear prediction to nonstationary signals is time segmentation: the signal is split into small portions and the modelization is carried out locally. The accuracy of the analysis is, however, dependent on the window size and on ...

2000