Serveur d'exploration sur la recherche en informatique en Lorraine

Attention, ce site est en cours de développement !
Attention, site généré par des moyens informatiques à partir de corpus bruts.
Les informations ne sont donc pas validées.

Multiple-order non-negative matrix factorization for speech enhancement

Identifieur interne : 003468 ( Hal/Curation ); précédent : 003467; suivant : 003469

Multiple-order non-negative matrix factorization for speech enhancement

Auteurs : Xabier Jaureguiberry [France] ; Emmanuel Vincent [France] ; Gaël Richard [France]

Source :

RBID : Hal:hal-01023399

English descriptors

Abstract

Amongst the speech enhancement techniques, statistical models based on Non-negative Matrix Factorization (NMF) have received great attention. In a single channel configuration, NMF is used to describe the spectral content of both the speech and noise sources. As the number of components can have a crucial influence on separation quality, we here propose to investigate model order selection based on the variational Bayesian approximation to the marginal likelihood of models of different orders. To go further, we propose to use model averaging to combine several single-order NMFs and we show that a straightforward application of model averaging principles is inefficient as it turned out to be equivalent to model selection. We thus introduce a parameter to control the entropy of the model order distribution which makes the averaging effective. We also show that our probabilistic model nicely extends to a multiple-order NMF model where several NMFs are jointly estimated and averaged. Experiments are conducted on real data from the CHiME challenge and give an interesting insight on the entropic parameter and model order priors. Separation results are also promising as model averaging outperforms single-order model selection. Finally, our multiple-order NMF shows an interesting gain in computation time.

Url:

Links toward previous steps (curation, corpus...)


Links to Exploration step

Hal:hal-01023399

Le document en format XML

<record>
<TEI>
<teiHeader>
<fileDesc>
<titleStmt>
<title xml:lang="en">Multiple-order non-negative matrix factorization for speech enhancement</title>
<author>
<name sortKey="Jaureguiberry, Xabier" sort="Jaureguiberry, Xabier" uniqKey="Jaureguiberry X" first="Xabier" last="Jaureguiberry">Xabier Jaureguiberry</name>
<affiliation wicri:level="1">
<hal:affiliation type="laboratory" xml:id="struct-162010" status="VALID">
<orgName>Laboratoire Traitement et Communication de l'Information</orgName>
<orgName type="acronym">LTCI</orgName>
<desc>
<address>
<addrLine>46 rue Barrault F-75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.ltci.telecom-paristech.fr/</ref>
</desc>
<listRelation>
<relation active="#struct-300362" type="direct"></relation>
<relation active="#struct-302102" type="direct"></relation>
<relation name="UMR5141" active="#struct-441569" type="direct"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-300362" type="direct">
<org type="institution" xml:id="struct-300362" status="VALID">
<orgName>Télécom ParisTech</orgName>
<desc>
<address>
<addrLine>46 rue Barrault 75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.telecom-paristech.fr</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-302102" type="direct">
<org type="institution" xml:id="struct-302102" status="VALID">
<orgName>Institut Mines-Télécom</orgName>
<desc>
<address>
<addrLine>46 rue Barrault -75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.mines-telecom.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR5141" active="#struct-441569" type="direct">
<org type="institution" xml:id="struct-441569" status="VALID">
<idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
<author>
<name sortKey="Vincent, Emmanuel" sort="Vincent, Emmanuel" uniqKey="Vincent E" first="Emmanuel" last="Vincent">Emmanuel Vincent</name>
<affiliation wicri:level="1">
<hal:affiliation type="researchteam" xml:id="struct-205127" status="OLD">
<idno type="RNSR">200118295L</idno>
<orgName>Analysis, perception and recognition of speech</orgName>
<orgName type="acronym">PAROLE</orgName>
<date type="end">2014-06-30</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/equipes/parole</ref>
</desc>
<listRelation>
<relation active="#struct-129671" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-423086" type="direct"></relation>
<relation active="#struct-206040" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-129671" type="direct">
<org type="laboratory" xml:id="struct-129671" status="VALID">
<idno type="RNSR">198618246Y</idno>
<orgName>INRIA Nancy - Grand Est</orgName>
<desc>
<address>
<addrLine>615 rue du Jardin Botanique 54600 Villers-lès-Nancy</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/nancy</ref>
</desc>
<listRelation>
<relation active="#struct-300009" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-300009" type="indirect">
<org type="institution" xml:id="struct-300009" status="VALID">
<orgName>Institut National de Recherche en Informatique et en Automatique</orgName>
<orgName type="acronym">Inria</orgName>
<desc>
<address>
<addrLine>Domaine de VoluceauRocquencourt - BP 10578153 Le Chesnay Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/en/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-423086" type="direct">
<org type="department" xml:id="struct-423086" status="VALID">
<orgName>Department of Natural Language Processing & Knowledge Discovery</orgName>
<orgName type="acronym">LORIA - NLPKD</orgName>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr/la-recherche-en/departements/Knowledge-and-Language-Management</ref>
</desc>
<listRelation>
<relation active="#struct-206040" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-206040" type="indirect">
<org type="laboratory" xml:id="struct-206040" status="VALID">
<idno type="IdRef">067077927</idno>
<idno type="RNSR">198912571S</idno>
<idno type="IdUnivLorraine">[UL]RSI--</idno>
<orgName>Laboratoire Lorrain de Recherche en Informatique et ses Applications</orgName>
<orgName type="acronym">LORIA</orgName>
<date type="start">2012-01-01</date>
<desc>
<address>
<addrLine>Campus Scientifique BP 239 54506 Vandoeuvre-lès-Nancy Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr</ref>
</desc>
<listRelation>
<relation active="#struct-300009" type="direct"></relation>
<relation active="#struct-413289" type="direct"></relation>
<relation name="UMR7503" active="#struct-441569" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-413289" type="indirect">
<org type="institution" xml:id="struct-413289" status="VALID">
<idno type="IdRef">157040569</idno>
<idno type="IdUnivLorraine">[UL]100--</idno>
<orgName>Université de Lorraine</orgName>
<orgName type="acronym">UL</orgName>
<date type="start">2012-01-01</date>
<desc>
<address>
<addrLine>34 cours Léopold - CS 25233 - 54052 Nancy cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.univ-lorraine.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR7503" active="#struct-441569" type="indirect">
<org type="institution" xml:id="struct-441569" status="VALID">
<idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
<placeName>
<settlement type="city">Nancy</settlement>
<settlement type="city">Metz</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Université de Lorraine</orgName>
</affiliation>
</author>
<author>
<name sortKey="Richard, Gael" sort="Richard, Gael" uniqKey="Richard G" first="Gaël" last="Richard">Gaël Richard</name>
<affiliation wicri:level="1">
<hal:affiliation type="laboratory" xml:id="struct-162010" status="VALID">
<orgName>Laboratoire Traitement et Communication de l'Information</orgName>
<orgName type="acronym">LTCI</orgName>
<desc>
<address>
<addrLine>46 rue Barrault F-75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.ltci.telecom-paristech.fr/</ref>
</desc>
<listRelation>
<relation active="#struct-300362" type="direct"></relation>
<relation active="#struct-302102" type="direct"></relation>
<relation name="UMR5141" active="#struct-441569" type="direct"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-300362" type="direct">
<org type="institution" xml:id="struct-300362" status="VALID">
<orgName>Télécom ParisTech</orgName>
<desc>
<address>
<addrLine>46 rue Barrault 75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.telecom-paristech.fr</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-302102" type="direct">
<org type="institution" xml:id="struct-302102" status="VALID">
<orgName>Institut Mines-Télécom</orgName>
<desc>
<address>
<addrLine>46 rue Barrault -75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.mines-telecom.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR5141" active="#struct-441569" type="direct">
<org type="institution" xml:id="struct-441569" status="VALID">
<idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
</titleStmt>
<publicationStmt>
<idno type="wicri:source">HAL</idno>
<idno type="RBID">Hal:hal-01023399</idno>
<idno type="halId">hal-01023399</idno>
<idno type="halUri">https://hal.archives-ouvertes.fr/hal-01023399</idno>
<idno type="url">https://hal.archives-ouvertes.fr/hal-01023399</idno>
<date when="2014-06-29">2014-06-29</date>
<idno type="wicri:Area/Hal/Corpus">003468</idno>
<idno type="wicri:Area/Hal/Curation">003468</idno>
</publicationStmt>
<sourceDesc>
<biblStruct>
<analytic>
<title xml:lang="en">Multiple-order non-negative matrix factorization for speech enhancement</title>
<author>
<name sortKey="Jaureguiberry, Xabier" sort="Jaureguiberry, Xabier" uniqKey="Jaureguiberry X" first="Xabier" last="Jaureguiberry">Xabier Jaureguiberry</name>
<affiliation wicri:level="1">
<hal:affiliation type="laboratory" xml:id="struct-162010" status="VALID">
<orgName>Laboratoire Traitement et Communication de l'Information</orgName>
<orgName type="acronym">LTCI</orgName>
<desc>
<address>
<addrLine>46 rue Barrault F-75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.ltci.telecom-paristech.fr/</ref>
</desc>
<listRelation>
<relation active="#struct-300362" type="direct"></relation>
<relation active="#struct-302102" type="direct"></relation>
<relation name="UMR5141" active="#struct-441569" type="direct"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-300362" type="direct">
<org type="institution" xml:id="struct-300362" status="VALID">
<orgName>Télécom ParisTech</orgName>
<desc>
<address>
<addrLine>46 rue Barrault 75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.telecom-paristech.fr</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-302102" type="direct">
<org type="institution" xml:id="struct-302102" status="VALID">
<orgName>Institut Mines-Télécom</orgName>
<desc>
<address>
<addrLine>46 rue Barrault -75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.mines-telecom.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR5141" active="#struct-441569" type="direct">
<org type="institution" xml:id="struct-441569" status="VALID">
<idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
<author>
<name sortKey="Vincent, Emmanuel" sort="Vincent, Emmanuel" uniqKey="Vincent E" first="Emmanuel" last="Vincent">Emmanuel Vincent</name>
<affiliation wicri:level="1">
<hal:affiliation type="researchteam" xml:id="struct-205127" status="OLD">
<idno type="RNSR">200118295L</idno>
<orgName>Analysis, perception and recognition of speech</orgName>
<orgName type="acronym">PAROLE</orgName>
<date type="end">2014-06-30</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/equipes/parole</ref>
</desc>
<listRelation>
<relation active="#struct-129671" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-423086" type="direct"></relation>
<relation active="#struct-206040" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-129671" type="direct">
<org type="laboratory" xml:id="struct-129671" status="VALID">
<idno type="RNSR">198618246Y</idno>
<orgName>INRIA Nancy - Grand Est</orgName>
<desc>
<address>
<addrLine>615 rue du Jardin Botanique 54600 Villers-lès-Nancy</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/nancy</ref>
</desc>
<listRelation>
<relation active="#struct-300009" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-300009" type="indirect">
<org type="institution" xml:id="struct-300009" status="VALID">
<orgName>Institut National de Recherche en Informatique et en Automatique</orgName>
<orgName type="acronym">Inria</orgName>
<desc>
<address>
<addrLine>Domaine de VoluceauRocquencourt - BP 10578153 Le Chesnay Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/en/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-423086" type="direct">
<org type="department" xml:id="struct-423086" status="VALID">
<orgName>Department of Natural Language Processing & Knowledge Discovery</orgName>
<orgName type="acronym">LORIA - NLPKD</orgName>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr/la-recherche-en/departements/Knowledge-and-Language-Management</ref>
</desc>
<listRelation>
<relation active="#struct-206040" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-206040" type="indirect">
<org type="laboratory" xml:id="struct-206040" status="VALID">
<idno type="IdRef">067077927</idno>
<idno type="RNSR">198912571S</idno>
<idno type="IdUnivLorraine">[UL]RSI--</idno>
<orgName>Laboratoire Lorrain de Recherche en Informatique et ses Applications</orgName>
<orgName type="acronym">LORIA</orgName>
<date type="start">2012-01-01</date>
<desc>
<address>
<addrLine>Campus Scientifique BP 239 54506 Vandoeuvre-lès-Nancy Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr</ref>
</desc>
<listRelation>
<relation active="#struct-300009" type="direct"></relation>
<relation active="#struct-413289" type="direct"></relation>
<relation name="UMR7503" active="#struct-441569" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-413289" type="indirect">
<org type="institution" xml:id="struct-413289" status="VALID">
<idno type="IdRef">157040569</idno>
<idno type="IdUnivLorraine">[UL]100--</idno>
<orgName>Université de Lorraine</orgName>
<orgName type="acronym">UL</orgName>
<date type="start">2012-01-01</date>
<desc>
<address>
<addrLine>34 cours Léopold - CS 25233 - 54052 Nancy cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.univ-lorraine.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR7503" active="#struct-441569" type="indirect">
<org type="institution" xml:id="struct-441569" status="VALID">
<idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
<placeName>
<settlement type="city">Nancy</settlement>
<settlement type="city">Metz</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Université de Lorraine</orgName>
</affiliation>
</author>
<author>
<name sortKey="Richard, Gael" sort="Richard, Gael" uniqKey="Richard G" first="Gaël" last="Richard">Gaël Richard</name>
<affiliation wicri:level="1">
<hal:affiliation type="laboratory" xml:id="struct-162010" status="VALID">
<orgName>Laboratoire Traitement et Communication de l'Information</orgName>
<orgName type="acronym">LTCI</orgName>
<desc>
<address>
<addrLine>46 rue Barrault F-75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.ltci.telecom-paristech.fr/</ref>
</desc>
<listRelation>
<relation active="#struct-300362" type="direct"></relation>
<relation active="#struct-302102" type="direct"></relation>
<relation name="UMR5141" active="#struct-441569" type="direct"></relation>
</listRelation>
<tutelles>
<tutelle active="#struct-300362" type="direct">
<org type="institution" xml:id="struct-300362" status="VALID">
<orgName>Télécom ParisTech</orgName>
<desc>
<address>
<addrLine>46 rue Barrault 75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.telecom-paristech.fr</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-302102" type="direct">
<org type="institution" xml:id="struct-302102" status="VALID">
<orgName>Institut Mines-Télécom</orgName>
<desc>
<address>
<addrLine>46 rue Barrault -75634 Paris Cedex 13</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.mines-telecom.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR5141" active="#struct-441569" type="direct">
<org type="institution" xml:id="struct-441569" status="VALID">
<idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc>
<address>
<country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
</analytic>
</biblStruct>
</sourceDesc>
</fileDesc>
<profileDesc>
<textClass>
<keywords scheme="mix" xml:lang="en">
<term>Model Averaging</term>
<term>Non-negative Matrix Factorization</term>
<term>Speech Enhancement</term>
<term>Variational Bayes</term>
</keywords>
</textClass>
</profileDesc>
</teiHeader>
<front>
<div type="abstract" xml:lang="en">Amongst the speech enhancement techniques, statistical models based on Non-negative Matrix Factorization (NMF) have received great attention. In a single channel configuration, NMF is used to describe the spectral content of both the speech and noise sources. As the number of components can have a crucial influence on separation quality, we here propose to investigate model order selection based on the variational Bayesian approximation to the marginal likelihood of models of different orders. To go further, we propose to use model averaging to combine several single-order NMFs and we show that a straightforward application of model averaging principles is inefficient as it turned out to be equivalent to model selection. We thus introduce a parameter to control the entropy of the model order distribution which makes the averaging effective. We also show that our probabilistic model nicely extends to a multiple-order NMF model where several NMFs are jointly estimated and averaged. Experiments are conducted on real data from the CHiME challenge and give an interesting insight on the entropic parameter and model order priors. Separation results are also promising as model averaging outperforms single-order model selection. Finally, our multiple-order NMF shows an interesting gain in computation time.</div>
</front>
</TEI>
<hal api="V3">
<titleStmt>
<title xml:lang="en">Multiple-order non-negative matrix factorization for speech enhancement</title>
<author role="aut">
<persName>
<forename type="first">Xabier</forename>
<surname>Jaureguiberry</surname>
</persName>
<email>xabier.jaureguiberry@telecom-paristech.fr</email>
<idno type="idhal">xabierj</idno>
<idno type="halauthor">872426</idno>
<affiliation ref="#struct-162010"></affiliation>
<affiliation ref="#struct-27016"></affiliation>
</author>
<author role="aut">
<persName>
<forename type="first">Emmanuel</forename>
<surname>Vincent</surname>
</persName>
<email>emmanuel.vincent@inria.fr</email>
<idno type="idhal">emmanuelv</idno>
<idno type="halauthor">571022</idno>
<affiliation ref="#struct-205127"></affiliation>
</author>
<author role="aut">
<persName>
<forename type="first">Gaël</forename>
<surname>Richard</surname>
</persName>
<email>gael.richard@telecom.paristech.fr</email>
<idno type="halauthor">476013</idno>
<affiliation ref="#struct-162010"></affiliation>
<affiliation ref="#struct-27016"></affiliation>
</author>
<editor role="depositor">
<persName>
<forename>Xabier</forename>
<surname>Jaureguiberry</surname>
</persName>
<email>xabierj@gmail.com</email>
</editor>
</titleStmt>
<editionStmt>
<edition n="v1" type="current">
<date type="whenSubmitted">2014-07-12 16:55:38</date>
<date type="whenModified">2016-01-19 01:06:38</date>
<date type="whenReleased">2014-07-16 15:21:05</date>
<date type="whenProduced">2014-06-29</date>
<date type="whenEndEmbargoed">2014-07-12</date>
<ref type="file" target="https://hal.archives-ouvertes.fr/hal-01023399/document">
<date notBefore="2014-07-12"></date>
</ref>
<ref type="file" subtype="author" n="1" target="https://hal.archives-ouvertes.fr/hal-01023399/file/interspeech14_vf.pdf">
<date notBefore="2014-07-12"></date>
</ref>
</edition>
<respStmt>
<resp>contributor</resp>
<name key="185586">
<persName>
<forename>Xabier</forename>
<surname>Jaureguiberry</surname>
</persName>
<email>xabierj@gmail.com</email>
</name>
</respStmt>
</editionStmt>
<publicationStmt>
<distributor>CCSD</distributor>
<idno type="halId">hal-01023399</idno>
<idno type="halUri">https://hal.archives-ouvertes.fr/hal-01023399</idno>
<idno type="halBibtex">jaureguiberry:hal-01023399</idno>
<idno type="halRefHtml">Interspeech, Jun 2014, Singapour, Singapore. pp.4, 2014</idno>
<idno type="halRef">Interspeech, Jun 2014, Singapour, Singapore. pp.4, 2014</idno>
</publicationStmt>
<seriesStmt>
<idno type="stamp" n="CNRS">CNRS - Centre national de la recherche scientifique</idno>
<idno type="stamp" n="INRIA">INRIA - Institut National de Recherche en Informatique et en Automatique</idno>
<idno type="stamp" n="ENST">Ecole Nationale Supérieure des Télécommunications</idno>
<idno type="stamp" n="INRIA-LORRAINE">INRIA Nancy - Grand Est</idno>
<idno type="stamp" n="LORIA2">Publications du LORIA</idno>
<idno type="stamp" n="INRIA-NANCY-GRAND-EST">INRIA Nancy - Grand Est</idno>
<idno type="stamp" n="LORIA-TALC" p="LORIA">Traitement automatique des langues et des connaissances</idno>
<idno type="stamp" n="LORIA">LORIA - Laboratoire Lorrain de Recherche en Informatique et ses Applications</idno>
<idno type="stamp" n="UNIV-LORRAINE">Université de Lorraine</idno>
<idno type="stamp" n="INSTITUT-TELECOM">Institut Télécom</idno>
<idno type="stamp" n="TELECOM-PARISTECH" p="INSTITUT-TELECOM">Télécom ParisTech</idno>
<idno type="stamp" n="PARISTECH">ParisTech</idno>
<idno type="stamp" n="INRIA_TEST">INRIA - Institut National de Recherche en Informatique et en Automatique</idno>
<idno type="stamp" n="GRID5000">Grid'5000</idno>
<idno type="stamp" n="INRIA2">INRIA 2</idno>
</seriesStmt>
<notesStmt>
<note type="audience" n="2">International</note>
<note type="invited" n="0">No</note>
<note type="popular" n="0">No</note>
<note type="peer" n="1">Yes</note>
<note type="proceedings" n="1">Yes</note>
</notesStmt>
<sourceDesc>
<biblStruct>
<analytic>
<title xml:lang="en">Multiple-order non-negative matrix factorization for speech enhancement</title>
<author role="aut">
<persName>
<forename type="first">Xabier</forename>
<surname>Jaureguiberry</surname>
</persName>
<email>xabier.jaureguiberry@telecom-paristech.fr</email>
<idno type="idHal">xabierj</idno>
<idno type="halAuthorId">872426</idno>
<affiliation ref="#struct-162010"></affiliation>
<affiliation ref="#struct-27016"></affiliation>
</author>
<author role="aut">
<persName>
<forename type="first">Emmanuel</forename>
<surname>Vincent</surname>
</persName>
<email>emmanuel.vincent@inria.fr</email>
<idno type="idHal">emmanuelv</idno>
<idno type="halAuthorId">571022</idno>
<affiliation ref="#struct-205127"></affiliation>
</author>
<author role="aut">
<persName>
<forename type="first">Gaël</forename>
<surname>Richard</surname>
</persName>
<email>gael.richard@telecom.paristech.fr</email>
<idno type="halAuthorId">476013</idno>
<affiliation ref="#struct-162010"></affiliation>
<affiliation ref="#struct-27016"></affiliation>
</author>
</analytic>
<monogr>
<meeting>
<title>Interspeech</title>
<date type="start">2014-06-29</date>
<settlement>Singapour</settlement>
<country key="SG">Singapore</country>
</meeting>
<imprint>
<biblScope unit="pp">4</biblScope>
<date type="datePub">2014-09-14</date>
</imprint>
</monogr>
</biblStruct>
</sourceDesc>
<profileDesc>
<langUsage>
<language ident="en">English</language>
</langUsage>
<textClass>
<keywords scheme="author">
<term xml:lang="en">Speech Enhancement</term>
<term xml:lang="en">Variational Bayes</term>
<term xml:lang="en">Non-negative Matrix Factorization</term>
<term xml:lang="en">Model Averaging</term>
</keywords>
<classCode scheme="halDomain" n="info.info-ts">Computer Science [cs]/Signal and Image Processing</classCode>
<classCode scheme="halDomain" n="spi.signal">Engineering Sciences [physics]/Signal and Image processing</classCode>
<classCode scheme="halTypology" n="COMM">Conference papers</classCode>
</textClass>
<abstract xml:lang="en">Amongst the speech enhancement techniques, statistical models based on Non-negative Matrix Factorization (NMF) have received great attention. In a single channel configuration, NMF is used to describe the spectral content of both the speech and noise sources. As the number of components can have a crucial influence on separation quality, we here propose to investigate model order selection based on the variational Bayesian approximation to the marginal likelihood of models of different orders. To go further, we propose to use model averaging to combine several single-order NMFs and we show that a straightforward application of model averaging principles is inefficient as it turned out to be equivalent to model selection. We thus introduce a parameter to control the entropy of the model order distribution which makes the averaging effective. We also show that our probabilistic model nicely extends to a multiple-order NMF model where several NMFs are jointly estimated and averaged. Experiments are conducted on real data from the CHiME challenge and give an interesting insight on the entropic parameter and model order priors. Separation results are also promising as model averaging outperforms single-order model selection. Finally, our multiple-order NMF shows an interesting gain in computation time.</abstract>
<particDesc>
<org type="consortium">Grid'5000</org>
</particDesc>
</profileDesc>
</hal>
</record>

Pour manipuler ce document sous Unix (Dilib)

EXPLOR_STEP=$WICRI_ROOT/Wicri/Lorraine/explor/InforLorV4/Data/Hal/Curation
HfdSelect -h $EXPLOR_STEP/biblio.hfd -nk 003468 | SxmlIndent | more

Ou

HfdSelect -h $EXPLOR_AREA/Data/Hal/Curation/biblio.hfd -nk 003468 | SxmlIndent | more

Pour mettre un lien sur cette page dans le réseau Wicri

{{Explor lien
   |wiki=    Wicri/Lorraine
   |area=    InforLorV4
   |flux=    Hal
   |étape=   Curation
   |type=    RBID
   |clé=     Hal:hal-01023399
   |texte=   Multiple-order non-negative matrix factorization for speech enhancement
}}

Wicri

This area was generated with Dilib version V0.6.33.
Data generation: Mon Jun 10 21:56:28 2019. Site generation: Fri Feb 25 15:29:27 2022