Skip to main navigation Skip to search Skip to main content

Novel Two-Stage Audiovisual Speech Filtering in Noisy Environments

  • Andrew Abel*
  • , Amir Hussain
  • *Corresponding author for this work

Research output: Contribution to journalArticlepeer-review

14 Scopus citations

Abstract

In recent years, the established link between the various human communication production domains has become more widely utilised in the field of speech processing. In this work, we build on previous work by the authors and present a novel two-stage audiovisual speech enhancement system, making use of audio-only beamforming, automatic lip tracking, and pre-processing with visually derived Wiener speech filtering. Initial results have demonstrated that this two-stage multimodal speech enhancement approach can produce positive results with noisy speech mixtures that conventional audio-only beamforming would struggle to cope with, such as in very noisy environments with a very low signal to noise ratio, and when the type of noise is difficult for audio-only beamforming to process.

Original languageEnglish
Pages (from-to)200-217
Number of pages18
JournalCognitive Computation
Volume6
Issue number2
DOIs
StatePublished - Jun 2014
Externally publishedYes

Keywords

  • Audiovisual speech processing
  • Multimodal speech filtering
  • Speech enhancement

ASJC Scopus subject areas

  • Computer Vision and Pattern Recognition
  • Computer Science Applications
  • Cognitive Neuroscience

Fingerprint

Dive into the research topics of 'Novel Two-Stage Audiovisual Speech Filtering in Noisy Environments'. Together they form a unique fingerprint.

Cite this