Search CORE

393 research outputs found

An adaptive stereo basis method for convolutive blind audio source separation

Author: Abdallah
Abdallah
Aharon
Amari
Amari
Araki
Bell
Cardoso
Cardoso
Cardoso
Davies
Douglas
Emmanuel Vincent
Hyvärinen
Ikeda
Ikram
Jafari
Jourjine
Knapp
Kurita
Lewicki
Makino
Maria G. Jafari
Mark D. Plumbley
Matsuda
Mike E. Davies
Mitianoudis
Mitianoudis
O’Grady
Parra
Samer A. Abdallah
Saruwatari
Sawada
Schmidt
Smaragdis
Torkkola
Vincent
Vincent
Viste
Yilmaz
Zhang
Publication venue: 'Elsevier BV'
Publication date: 01/01/2008
Field of study

NOTICE: this is the author’s version of a work that was accepted for publication in Neurocomputing. Changes resulting from the publishing process, such as peer review, editing, corrections, structural formatting, and other quality control mechanisms may not be reflected in this document. Changes may have been made to this work since it was submitted for publication. A definitive version was subsequently published in PUBLICATION, [71, 10-12, June 2008] DOI:neucom.2007.08.02

Crossref

UCL Discovery

Edinburgh Research Explorer

Queen Mary Research Online

Multimodal methods for blind source separation of audio sources

Author: Syed M.R. Naqvi (7200659)
Publication venue
Publication date: 01/01/2009
Field of study

The enhancement of the performance of frequency domain convolutive blind source separation (FDCBSS) techniques when applied to the problem of separating audio sources recorded in a room environment is the focus of this thesis. This challenging application is termed the cocktail party problem and the ultimate aim would be to build a machine which matches the ability of a human being to solve this task. Human beings exploit both their eyes and their ears in solving this task and hence they adopt a multimodal approach, i.e. they exploit both audio and video modalities. New multimodal methods for blind source separation of audio sources are therefore proposed in this work as a step towards realizing such a machine. The geometry of the room environment is initially exploited to improve the separation performance of a FDCBSS algorithm. The positions of the human speakers are monitored by video cameras and this information is incorporated within the FDCBSS algorithm in the form of constraints added to the underlying cross-power spectral density matrix-based cost function which measures separation performance. [Continues.

Loughborough University Institutional Repository

Convolutive Blind Source Separation Methods

Author: Kjems Ulrik
Larsen Jan
Parra Lucas C.
Pedersen Michael Syskind
Publication venue: 'Springer Fachmedien Wiesbaden GmbH'
Publication date: 01/01/2008
Field of study

In this chapter, we provide an overview of existing algorithms for blind source separation of convolutive audio mixtures. We provide a taxonomy, wherein many of the existing algorithms can be organized, and we present published results from those algorithms that have been applied to real-world audio separation tasks

CiteSeerX

Online Research Database In Technology

Source Separation for Hearing Aid Applications

Author: Pedersen Michael Syskind
Publication venue: Technical University of Denmark
Publication date: 01/11/2006
Field of study

Online Research Database In Technology

Blind Source Separation for Speech Application Under Real Acoustic Environment

Author: Hiroshi Saruwatari
Yu Takahashi
Publication venue: 'IntechOpen'
Publication date: 10/10/2012
Field of study

IntechOpen

Multichannel Speech Enhancement

Author: Lino Garcia
Soledad Torres-Guijarro
Publication venue: 'IntechOpen'
Publication date: 01/10/2008
Field of study

IntechOpen

Developing A System For Blind Acoustic Source Localization And Separation

Author: Kulkarni Raghavendra
Publication venue: DigitalCommons@WayneState
Publication date: 01/01/2013
Field of study

This dissertation presents innovate methodologies for locating, extracting, and separating multiple incoherent sound sources in three-dimensional (3D) space; and applications of the time reversal (TR) algorithm to pinpoint the hyper active neural activities inside the brain auditory structure that are correlated to the tinnitus pathology. Specifically, an acoustic modeling based method is developed for locating arbitrary and incoherent sound sources in 3D space in real time by using a minimal number of microphones, and the Point Source Separation (PSS) method is developed for extracting target signals from directly measured mixed signals. Combining these two approaches leads to a novel technology known as Blind Sources Localization and Separation (BSLS) that enables one to locate multiple incoherent sound signals in 3D space and separate original individual sources simultaneously, based on the directly measured mixed signals. These technologies have been validated through numerical simulations and experiments conducted in various non-ideal environments where there are non-negligible, unspecified sound reflections and reverberation as well as interferences from random background noise. Another innovation presented in this dissertation is concerned with applications of the TR algorithm to pinpoint the exact locations of hyper-active neurons in the brain auditory structure that are directly correlated to the tinnitus perception. Benchmark tests conducted on normal rats have confirmed the localization results provided by the TR algorithm. Results demonstrate that the spatial resolution of this source localization can be as high as the micrometer level. This high precision localization may lead to a paradigm shift in tinnitus diagnosis, which may in turn produce a more cost-effective treatment for tinnitus than any of the existing ones

Digital Commons@Wayne State University