Search CORE

150 research outputs found

Making Faces - State-Space Models Applied to Multi-Modal Signal Processing

Author: Lehn-Schiøler Tue
Publication venue: Technical University of Denmark
Publication date: 01/01/2005
Field of study

CASA 2009:International Conference on Computer Animation and Social Agents

Author: Egges Arjan
Publication venue: Centre for Telematics and Information Technology (CTIT)
Publication date: 05/06/2009
Field of study

University of Twente Research Information

Biorthogonality in lapped transforms : a study in high-quality audio compression

Author: Cheung Shiufun
Publication venue: Massachusetts Institute of Technology
Publication date: 01/01/1996
Field of study

Thesis (Ph. D.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 1996.Includes bibliographical references (leaves 76-82).by Shiufun Cheung.Ph.D

Sparse and structured decomposition of audio signals on hybrid dictionaries using musical priors

Author: Kowalski Matthieu
Papadopoulos Hélène
Publication venue: 'Acoustical Society of America (ASA)'
Publication date: 02/04/2013
Field of study

International audienceThis paper investigates the use of musical priors for sparse expansion of audio signals of music, on an overcomplete dual-resolution dictionary taken from the union of two orthonormal bases that can describe both transient and tonal components of a music audio signal. More specifically, chord and metrical structure information are used to build a structured model that takes into account dependencies between coefficients of the decomposition, both for the tonal and for the transient layer. The denoising task application is used to provide a proof of concept of the proposed musical priors. Several configurations of the model are analyzed. Evaluation on monophonic and complex polyphonic excerpts of real music signals shows that the proposed approach provides results whose quality measured by the signal-to-noise ratio is competitive with state-of-the-art approaches, and more coherent with the semantic content of the signal. A detailed analysis of the model in terms of sparsity and in terms of interpretability of the representation is also provided, and shows that the model is capable of giving a relevant and legible representation of Western tonal music audio signals

HAL-CentraleSupelec