Skip to main navigation Skip to search Skip to main content

Audio-Visual Source Separation with Alternating Diffusion Maps

Research output: Chapter in Book/Report/Conference proceedingChapterpeer-review

Abstract

In this chapter we consider the separation of multiple sound sources of different types including multiple speakers and transients, which are measured by a single microphone and by a video camera. We address the problem of separating a particular sound source from all other sources focusing specifically on obtaining an underlying representation of it while attenuating all other sources. By pointing the video camera merely to the desired sound source, the problem becomes equivalent to extracting the common source to the audio and the video modalities while ignoring the other sources. We use a kernel-based method, which is particularly designed for this task, providing an underlying representation of the common source. We demonstrate the usefulness of the obtained representation for the activity detection of the common source and discuss how it may be further used for source separation.

Original languageEnglish GB
Title of host publicationAUDIO SOURCE SEPARATION
Pages365-382
Number of pages18
DOIs
StatePublished - 2018

Publication series

NameSignals and Communication Technology

ASJC Scopus subject areas

  • Control and Systems Engineering
  • Signal Processing
  • Computer Networks and Communications
  • Electrical and Electronic Engineering

Fingerprint

Dive into the research topics of 'Audio-Visual Source Separation with Alternating Diffusion Maps'. Together they form a unique fingerprint.

Cite this