Analysis of Multimodal Signals Using Redundant Representations

Monaci, G.; Divorra Escoda, O.; Vandergheynst, P.

doi:10.1109/ICIP.2005.1530349

Monaci, G.; Divorra Escoda, O.; Vandergheynst, P.

2005

Download

Formats

Format
BibTeX
MARCXML
TextMARC
MARC
DublinCore
EndNote
NLM
RefWorks
RIS

Files

Abstract

In this work we explore the potentialities of a framework for the representation of audio-visual signals using decompositions on overcomplete dictionaries. Redundant decompositions may describe audio-visual sequences in a concise fashion, preserving good representation properties thanks to the use of redundant, well designed, dictionaries. We expect that this will help us overcome two typical problems of multimodal fusion algorithms. On one hand, classical representation techniques, like pixel-based measures (for the video) or Fourier-like transforms (for the audio), take into account only marginally the physics of the problem. On the other hand, the input signals have large dimensionality. The results we obtain by making use of sparse decompositions of audio-visual signals over redundant codebooks are encouraging and show the potentialities of the proposed approach to multimodal signal representation.

Details

Title Analysis of Multimodal Signals Using Redundant Representations

Author(s) Monaci, G. ; Divorra Escoda, O. ; Vandergheynst, P.

Published in IEEE International Conference on Image Processing 2005

Pages 46-49

Conference IEEE International Conference on Image Processing (ICIP '05), Genova

Date 2005

Keywords

LTS2

Note Winner of IBM Student Paper Award

DOI https://doi.org/10.1109/ICIP.2005.1530349

Laboratories LTS2

Record Appears in Scientific production and competences > STI - School of Engineering > IEM - Institut d'Electricité et de Microtechnique > LTS2 - Signal Processing Laboratory 2
Conference Papers
Work produced at EPFL
Published

Record creation date 2006-06-14

Files

Abstract

Details

PDF