Improving the recognition of pathological voice using the discriminant HLDA transformation
Identifieur interne : 000903 ( Main/Exploration ); précédent : 000902; suivant : 000904Improving the recognition of pathological voice using the discriminant HLDA transformation
Auteurs : Othman Lachhab [France] ; Joseph Di Martino [France] ; El Hassane Ibn Elhaj [Maroc] ; Ahmed Hammouch [France]Source :
English descriptors
- mix :
Abstract
In this paper, we propose a simple and fast method for evaluating the pathological voice (esophageal) by applying the continuous speech recognition in a speaker dependent mode, on our own database of the pathological voice, we call FPSD (French Pathological Speech Database). The recognition system used is implemented using the HTK platform, based on HMM/GMM monophone models. The acoustic vectors are linearly transformed by the HLDA (Heteroscedastic Linear Discriminant Analysis) method to reduce their size in a smaller space with good discriminative properties. The obtained phone recognition rate (63.59 %) is very promising when we know that esophageal voice contains unnatural sounds, difficult to understand.
Url:
Affiliations:
Links toward previous steps (curation, corpus...)
- to stream Hal, to step Corpus: 002A37
- to stream Hal, to step Curation: 002A37
- to stream Hal, to step Checkpoint: 000837
- to stream Main, to step Merge: 000904
- to stream Main, to step Curation: 000903
Le document en format XML
<record><TEI><teiHeader><fileDesc><titleStmt><title xml:lang="en">Improving the recognition of pathological voice using the discriminant HLDA transformation</title>
<author><name sortKey="Lachhab, Othman" sort="Lachhab, Othman" uniqKey="Lachhab O" first="Othman" last="Lachhab">Othman Lachhab</name>
<affiliation wicri:level="1"><hal:affiliation type="laboratory" xml:id="struct-208694" status="INCOMING"><orgName>Ecole Normale Supérieure de l'Enseignement Technique</orgName>
<orgName type="acronym">ENSET</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
<listRelation><relation active="#struct-351043" type="direct"></relation>
</listRelation>
<tutelles><tutelle active="#struct-351043" type="direct"><org type="institution" xml:id="struct-351043" status="INCOMING"><orgName>Ecole normale supérieure de l'Enseignement Technique</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
<author><name sortKey="Di Martino, Joseph" sort="Di Martino, Joseph" uniqKey="Di Martino J" first="Joseph" last="Di Martino">Joseph Di Martino</name>
<affiliation wicri:level="1"><hal:affiliation type="researchteam" xml:id="struct-205127" status="OLD"><idno type="RNSR">200118295L</idno>
<orgName>Analysis, perception and recognition of speech</orgName>
<orgName type="acronym">PAROLE</orgName>
<date type="end">2014-06-30</date>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/equipes/parole</ref>
</desc>
<listRelation><relation active="#struct-129671" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-423086" type="direct"></relation>
<relation active="#struct-206040" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
<tutelles><tutelle active="#struct-129671" type="direct"><org type="laboratory" xml:id="struct-129671" status="VALID"><idno type="RNSR">198618246Y</idno>
<orgName>INRIA Nancy - Grand Est</orgName>
<desc><address><addrLine>615 rue du Jardin Botanique 54600 Villers-lès-Nancy</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/nancy</ref>
</desc>
<listRelation><relation active="#struct-300009" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-300009" type="indirect"><org type="institution" xml:id="struct-300009" status="VALID"><orgName>Institut National de Recherche en Informatique et en Automatique</orgName>
<orgName type="acronym">Inria</orgName>
<desc><address><addrLine>Domaine de VoluceauRocquencourt - BP 10578153 Le Chesnay Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/en/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-423086" type="direct"><org type="department" xml:id="struct-423086" status="VALID"><orgName>Department of Natural Language Processing & Knowledge Discovery</orgName>
<orgName type="acronym">LORIA - NLPKD</orgName>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr/la-recherche-en/departements/Knowledge-and-Language-Management</ref>
</desc>
<listRelation><relation active="#struct-206040" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-206040" type="indirect"><org type="laboratory" xml:id="struct-206040" status="VALID"><idno type="IdRef">067077927</idno>
<idno type="RNSR">198912571S</idno>
<idno type="IdUnivLorraine">[UL]RSI--</idno>
<orgName>Laboratoire Lorrain de Recherche en Informatique et ses Applications</orgName>
<orgName type="acronym">LORIA</orgName>
<date type="start">2012-01-01</date>
<desc><address><addrLine>Campus Scientifique BP 239 54506 Vandoeuvre-lès-Nancy Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr</ref>
</desc>
<listRelation><relation active="#struct-300009" type="direct"></relation>
<relation active="#struct-413289" type="direct"></relation>
<relation name="UMR7503" active="#struct-441569" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-413289" type="indirect"><org type="institution" xml:id="struct-413289" status="VALID"><idno type="IdRef">157040569</idno>
<idno type="IdUnivLorraine">[UL]100--</idno>
<orgName>Université de Lorraine</orgName>
<orgName type="acronym">UL</orgName>
<date type="start">2012-01-01</date>
<desc><address><addrLine>34 cours Léopold - CS 25233 - 54052 Nancy cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.univ-lorraine.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR7503" active="#struct-441569" type="indirect"><org type="institution" xml:id="struct-441569" status="VALID"><idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
<placeName><settlement type="city">Nancy</settlement>
<settlement type="city">Metz</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Université de Lorraine</orgName>
</affiliation>
</author>
<author><name sortKey="Ibn Elhaj, El Hassane" sort="Ibn Elhaj, El Hassane" uniqKey="Ibn Elhaj E" first="El Hassane" last="Ibn Elhaj">El Hassane Ibn Elhaj</name>
<affiliation wicri:level="1"><hal:affiliation type="institution" xml:id="struct-332892" status="VALID"><orgName>Institut National de Postes et Télécommunications [Rabat]</orgName>
<orgName type="acronym">INPT</orgName>
<desc><address><addrLine>2, av ALLal EL Fassi - Madinat AL Irfane - Rabat</addrLine>
<country key="MA"></country>
</address>
<ref type="url">http://www.inpt.ac.ma</ref>
</desc>
</hal:affiliation>
<country>Maroc</country>
</affiliation>
</author>
<author><name sortKey="Hammouch, Ahmed" sort="Hammouch, Ahmed" uniqKey="Hammouch A" first="Ahmed" last="Hammouch">Ahmed Hammouch</name>
<affiliation wicri:level="1"><hal:affiliation type="laboratory" xml:id="struct-208694" status="INCOMING"><orgName>Ecole Normale Supérieure de l'Enseignement Technique</orgName>
<orgName type="acronym">ENSET</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
<listRelation><relation active="#struct-351043" type="direct"></relation>
</listRelation>
<tutelles><tutelle active="#struct-351043" type="direct"><org type="institution" xml:id="struct-351043" status="INCOMING"><orgName>Ecole normale supérieure de l'Enseignement Technique</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
</titleStmt>
<publicationStmt><idno type="wicri:source">HAL</idno>
<idno type="RBID">Hal:hal-01093309</idno>
<idno type="halId">hal-01093309</idno>
<idno type="halUri">https://hal.inria.fr/hal-01093309</idno>
<idno type="url">https://hal.inria.fr/hal-01093309</idno>
<date when="2014-10-20">2014-10-20</date>
<idno type="wicri:Area/Hal/Corpus">002A37</idno>
<idno type="wicri:Area/Hal/Curation">002A37</idno>
<idno type="wicri:Area/Hal/Checkpoint">000837</idno>
<idno type="wicri:explorRef" wicri:stream="Hal" wicri:step="Checkpoint">000837</idno>
<idno type="wicri:Area/Main/Merge">000904</idno>
<idno type="wicri:Area/Main/Curation">000903</idno>
<idno type="wicri:Area/Main/Exploration">000903</idno>
</publicationStmt>
<sourceDesc><biblStruct><analytic><title xml:lang="en">Improving the recognition of pathological voice using the discriminant HLDA transformation</title>
<author><name sortKey="Lachhab, Othman" sort="Lachhab, Othman" uniqKey="Lachhab O" first="Othman" last="Lachhab">Othman Lachhab</name>
<affiliation wicri:level="1"><hal:affiliation type="laboratory" xml:id="struct-208694" status="INCOMING"><orgName>Ecole Normale Supérieure de l'Enseignement Technique</orgName>
<orgName type="acronym">ENSET</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
<listRelation><relation active="#struct-351043" type="direct"></relation>
</listRelation>
<tutelles><tutelle active="#struct-351043" type="direct"><org type="institution" xml:id="struct-351043" status="INCOMING"><orgName>Ecole normale supérieure de l'Enseignement Technique</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
<author><name sortKey="Di Martino, Joseph" sort="Di Martino, Joseph" uniqKey="Di Martino J" first="Joseph" last="Di Martino">Joseph Di Martino</name>
<affiliation wicri:level="1"><hal:affiliation type="researchteam" xml:id="struct-205127" status="OLD"><idno type="RNSR">200118295L</idno>
<orgName>Analysis, perception and recognition of speech</orgName>
<orgName type="acronym">PAROLE</orgName>
<date type="end">2014-06-30</date>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/equipes/parole</ref>
</desc>
<listRelation><relation active="#struct-129671" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-423086" type="direct"></relation>
<relation active="#struct-206040" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
<tutelles><tutelle active="#struct-129671" type="direct"><org type="laboratory" xml:id="struct-129671" status="VALID"><idno type="RNSR">198618246Y</idno>
<orgName>INRIA Nancy - Grand Est</orgName>
<desc><address><addrLine>615 rue du Jardin Botanique 54600 Villers-lès-Nancy</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/nancy</ref>
</desc>
<listRelation><relation active="#struct-300009" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-300009" type="indirect"><org type="institution" xml:id="struct-300009" status="VALID"><orgName>Institut National de Recherche en Informatique et en Automatique</orgName>
<orgName type="acronym">Inria</orgName>
<desc><address><addrLine>Domaine de VoluceauRocquencourt - BP 10578153 Le Chesnay Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.inria.fr/en/</ref>
</desc>
</org>
</tutelle>
<tutelle active="#struct-423086" type="direct"><org type="department" xml:id="struct-423086" status="VALID"><orgName>Department of Natural Language Processing & Knowledge Discovery</orgName>
<orgName type="acronym">LORIA - NLPKD</orgName>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr/la-recherche-en/departements/Knowledge-and-Language-Management</ref>
</desc>
<listRelation><relation active="#struct-206040" type="direct"></relation>
<relation active="#struct-300009" type="indirect"></relation>
<relation active="#struct-413289" type="indirect"></relation>
<relation name="UMR7503" active="#struct-441569" type="indirect"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-206040" type="indirect"><org type="laboratory" xml:id="struct-206040" status="VALID"><idno type="IdRef">067077927</idno>
<idno type="RNSR">198912571S</idno>
<idno type="IdUnivLorraine">[UL]RSI--</idno>
<orgName>Laboratoire Lorrain de Recherche en Informatique et ses Applications</orgName>
<orgName type="acronym">LORIA</orgName>
<date type="start">2012-01-01</date>
<desc><address><addrLine>Campus Scientifique BP 239 54506 Vandoeuvre-lès-Nancy Cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.loria.fr</ref>
</desc>
<listRelation><relation active="#struct-300009" type="direct"></relation>
<relation active="#struct-413289" type="direct"></relation>
<relation name="UMR7503" active="#struct-441569" type="direct"></relation>
</listRelation>
</org>
</tutelle>
<tutelle active="#struct-413289" type="indirect"><org type="institution" xml:id="struct-413289" status="VALID"><idno type="IdRef">157040569</idno>
<idno type="IdUnivLorraine">[UL]100--</idno>
<orgName>Université de Lorraine</orgName>
<orgName type="acronym">UL</orgName>
<date type="start">2012-01-01</date>
<desc><address><addrLine>34 cours Léopold - CS 25233 - 54052 Nancy cedex</addrLine>
<country key="FR"></country>
</address>
<ref type="url">http://www.univ-lorraine.fr/</ref>
</desc>
</org>
</tutelle>
<tutelle name="UMR7503" active="#struct-441569" type="indirect"><org type="institution" xml:id="struct-441569" status="VALID"><idno type="ISNI">0000000122597504</idno>
<idno type="IdRef">02636817X</idno>
<orgName>Centre National de la Recherche Scientifique</orgName>
<orgName type="acronym">CNRS</orgName>
<date type="start">1939-10-19</date>
<desc><address><country key="FR"></country>
</address>
<ref type="url">http://www.cnrs.fr/</ref>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
<placeName><settlement type="city">Nancy</settlement>
<settlement type="city">Metz</settlement>
<region type="region" nuts="2">Grand Est</region>
<region type="old region" nuts="2">Lorraine (région)</region>
</placeName>
<orgName type="university">Université de Lorraine</orgName>
</affiliation>
</author>
<author><name sortKey="Ibn Elhaj, El Hassane" sort="Ibn Elhaj, El Hassane" uniqKey="Ibn Elhaj E" first="El Hassane" last="Ibn Elhaj">El Hassane Ibn Elhaj</name>
<affiliation wicri:level="1"><hal:affiliation type="institution" xml:id="struct-332892" status="VALID"><orgName>Institut National de Postes et Télécommunications [Rabat]</orgName>
<orgName type="acronym">INPT</orgName>
<desc><address><addrLine>2, av ALLal EL Fassi - Madinat AL Irfane - Rabat</addrLine>
<country key="MA"></country>
</address>
<ref type="url">http://www.inpt.ac.ma</ref>
</desc>
</hal:affiliation>
<country>Maroc</country>
</affiliation>
</author>
<author><name sortKey="Hammouch, Ahmed" sort="Hammouch, Ahmed" uniqKey="Hammouch A" first="Ahmed" last="Hammouch">Ahmed Hammouch</name>
<affiliation wicri:level="1"><hal:affiliation type="laboratory" xml:id="struct-208694" status="INCOMING"><orgName>Ecole Normale Supérieure de l'Enseignement Technique</orgName>
<orgName type="acronym">ENSET</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
<listRelation><relation active="#struct-351043" type="direct"></relation>
</listRelation>
<tutelles><tutelle active="#struct-351043" type="direct"><org type="institution" xml:id="struct-351043" status="INCOMING"><orgName>Ecole normale supérieure de l'Enseignement Technique</orgName>
<desc><address><country key="FR"></country>
</address>
</desc>
</org>
</tutelle>
</tutelles>
</hal:affiliation>
<country>France</country>
</affiliation>
</author>
</analytic>
</biblStruct>
</sourceDesc>
</fileDesc>
<profileDesc><textClass><keywords scheme="mix" xml:lang="en"><term>Automatic Speech Recognition (ASR)</term>
<term>GMM</term>
<term>HLDA</term>
<term>HMM</term>
<term>HTK</term>
<term>MFCC</term>
<term>Pathological voices</term>
</keywords>
</textClass>
</profileDesc>
</teiHeader>
<front><div type="abstract" xml:lang="en">In this paper, we propose a simple and fast method for evaluating the pathological voice (esophageal) by applying the continuous speech recognition in a speaker dependent mode, on our own database of the pathological voice, we call FPSD (French Pathological Speech Database). The recognition system used is implemented using the HTK platform, based on HMM/GMM monophone models. The acoustic vectors are linearly transformed by the HLDA (Heteroscedastic Linear Discriminant Analysis) method to reduce their size in a smaller space with good discriminative properties. The obtained phone recognition rate (63.59 %) is very promising when we know that esophageal voice contains unnatural sounds, difficult to understand.</div>
</front>
</TEI>
<affiliations><list><country><li>France</li>
<li>Maroc</li>
</country>
<region><li>Grand Est</li>
<li>Lorraine (région)</li>
</region>
<settlement><li>Metz</li>
<li>Nancy</li>
</settlement>
<orgName><li>Université de Lorraine</li>
</orgName>
</list>
<tree><country name="France"><noRegion><name sortKey="Lachhab, Othman" sort="Lachhab, Othman" uniqKey="Lachhab O" first="Othman" last="Lachhab">Othman Lachhab</name>
</noRegion>
<name sortKey="Di Martino, Joseph" sort="Di Martino, Joseph" uniqKey="Di Martino J" first="Joseph" last="Di Martino">Joseph Di Martino</name>
<name sortKey="Hammouch, Ahmed" sort="Hammouch, Ahmed" uniqKey="Hammouch A" first="Ahmed" last="Hammouch">Ahmed Hammouch</name>
</country>
<country name="Maroc"><noRegion><name sortKey="Ibn Elhaj, El Hassane" sort="Ibn Elhaj, El Hassane" uniqKey="Ibn Elhaj E" first="El Hassane" last="Ibn Elhaj">El Hassane Ibn Elhaj</name>
</noRegion>
</country>
</tree>
</affiliations>
</record>
Pour manipuler ce document sous Unix (Dilib)
EXPLOR_STEP=$WICRI_ROOT/Wicri/Lorraine/explor/InforLorV4/Data/Main/Exploration
HfdSelect -h $EXPLOR_STEP/biblio.hfd -nk 000903 | SxmlIndent | more
Ou
HfdSelect -h $EXPLOR_AREA/Data/Main/Exploration/biblio.hfd -nk 000903 | SxmlIndent | more
Pour mettre un lien sur cette page dans le réseau Wicri
{{Explor lien |wiki= Wicri/Lorraine |area= InforLorV4 |flux= Main |étape= Exploration |type= RBID |clé= Hal:hal-01093309 |texte= Improving the recognition of pathological voice using the discriminant HLDA transformation }}
This area was generated with Dilib version V0.6.33. |