Robust speech representation of voiced sounds based on synchrony determination with PLLs

We propose to include synchrony effects, known to exist in the auditory system, to represent voiced parts of the speech signal in a robust way. The system decomposes the input signal by means of a bandpass filter bank, and utilizes a bank of phase locked loops (PLLs) to obtain information on the fre...

Ausführliche Beschreibung

Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Pelle, Patricia, Estienne, Claudio, Franco, Horacio
Format: Tagungsbericht
Sprache:eng
Schlagworte:
Online-Zugang:Volltext bestellen
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
container_end_page 5427
container_issue
container_start_page 5424
container_title
container_volume
creator Pelle, Patricia
Estienne, Claudio
Franco, Horacio
description We propose to include synchrony effects, known to exist in the auditory system, to represent voiced parts of the speech signal in a robust way. The system decomposes the input signal by means of a bandpass filter bank, and utilizes a bank of phase locked loops (PLLs) to obtain information on the frequencies present at a specific time. This information about the frequency distribution is transformed into a spectral-like representation based on synchrony effects. Noisy speech recognition experiments are performed using this synchrony-based spectrum, which is transformed into a small set of coefficients by using a transformation similar to that utilized for mel cepstrum features. We show that recognition performance compared to mel cepstrum features is advantageous, when measured over a range of SNR conditions, especially in the high noise level case.
doi_str_mv 10.1109/ICASSP.2011.5947585
format Conference Proceeding
fullrecord <record><control><sourceid>ieee_6IE</sourceid><recordid>TN_cdi_ieee_primary_5947585</recordid><sourceformat>XML</sourceformat><sourcesystem>PC</sourcesystem><ieee_id>5947585</ieee_id><sourcerecordid>5947585</sourcerecordid><originalsourceid>FETCH-LOGICAL-i175t-f441f613a0f9bb63f33c05a3c9bde0b6b8ac29ab114005a30e4d345f2d7448e3</originalsourceid><addsrcrecordid>eNo1UMtqwzAQVF9QN80X5KIfcLprSZZ1LKEvMDQ0OeQWJHuFVRo7WE5L_r4uSecy7AwzsMPYDGGOCObhbfG4Wi3nGSDOlZFaFeqC3aFUWoMSRl-yJBPapGhgc8WmRhf_XgHXLEGVQZqjNLdsGuMnjMgzrZVJ2Oajc4c48Lgnqhre076nSO1gh9C1vPP8uwsV1Tx2h7aO3Nk4HqMTj23V9F175DUN1O9Ce0r8hKHhy7KM9-zG269I0zNP2Pr5ab14Tcv3l_GdMg2o1ZB6KdHnKCx441wuvBAVKCsq42oCl7vCVpmxDlHCnw4kayGVz2otZUFiwman2kBE230fdrY_bs8biV83_lha</addsrcrecordid><sourcetype>Publisher</sourcetype><iscdi>true</iscdi><recordtype>conference_proceeding</recordtype></control><display><type>conference_proceeding</type><title>Robust speech representation of voiced sounds based on synchrony determination with PLLs</title><source>IEEE Electronic Library (IEL) Conference Proceedings</source><creator>Pelle, Patricia ; Estienne, Claudio ; Franco, Horacio</creator><creatorcontrib>Pelle, Patricia ; Estienne, Claudio ; Franco, Horacio</creatorcontrib><description>We propose to include synchrony effects, known to exist in the auditory system, to represent voiced parts of the speech signal in a robust way. The system decomposes the input signal by means of a bandpass filter bank, and utilizes a bank of phase locked loops (PLLs) to obtain information on the frequencies present at a specific time. This information about the frequency distribution is transformed into a spectral-like representation based on synchrony effects. Noisy speech recognition experiments are performed using this synchrony-based spectrum, which is transformed into a small set of coefficients by using a transformation similar to that utilized for mel cepstrum features. We show that recognition performance compared to mel cepstrum features is advantageous, when measured over a range of SNR conditions, especially in the high noise level case.</description><identifier>ISSN: 1520-6149</identifier><identifier>ISBN: 9781457705380</identifier><identifier>ISBN: 1457705389</identifier><identifier>EISSN: 2379-190X</identifier><identifier>EISBN: 1457705397</identifier><identifier>EISBN: 9781457705373</identifier><identifier>EISBN: 9781457705397</identifier><identifier>EISBN: 1457705370</identifier><identifier>DOI: 10.1109/ICASSP.2011.5947585</identifier><language>eng</language><publisher>IEEE</publisher><subject>auditory system ; Histograms ; noise ; Noise measurement ; Phase locked loops ; PLL ; Robustness ; Speech ; speech features ; Time frequency analysis ; Voltage-controlled oscillators</subject><ispartof>2011 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2011, p.5424-5427</ispartof><woscitedreferencessubscribed>false</woscitedreferencessubscribed></display><links><openurl>$$Topenurl_article</openurl><openurlfulltext>$$Topenurlfull_article</openurlfulltext><thumbnail>$$Tsyndetics_thumb_exl</thumbnail><linktohtml>$$Uhttps://ieeexplore.ieee.org/document/5947585$$EHTML$$P50$$Gieee$$H</linktohtml><link.rule.ids>310,311,781,785,790,791,2059,27930,54925</link.rule.ids><linktorsrc>$$Uhttps://ieeexplore.ieee.org/document/5947585$$EView_record_in_IEEE$$FView_record_in_$$GIEEE</linktorsrc></links><search><creatorcontrib>Pelle, Patricia</creatorcontrib><creatorcontrib>Estienne, Claudio</creatorcontrib><creatorcontrib>Franco, Horacio</creatorcontrib><title>Robust speech representation of voiced sounds based on synchrony determination with PLLs</title><title>2011 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)</title><addtitle>ICASSP</addtitle><description>We propose to include synchrony effects, known to exist in the auditory system, to represent voiced parts of the speech signal in a robust way. The system decomposes the input signal by means of a bandpass filter bank, and utilizes a bank of phase locked loops (PLLs) to obtain information on the frequencies present at a specific time. This information about the frequency distribution is transformed into a spectral-like representation based on synchrony effects. Noisy speech recognition experiments are performed using this synchrony-based spectrum, which is transformed into a small set of coefficients by using a transformation similar to that utilized for mel cepstrum features. We show that recognition performance compared to mel cepstrum features is advantageous, when measured over a range of SNR conditions, especially in the high noise level case.</description><subject>auditory system</subject><subject>Histograms</subject><subject>noise</subject><subject>Noise measurement</subject><subject>Phase locked loops</subject><subject>PLL</subject><subject>Robustness</subject><subject>Speech</subject><subject>speech features</subject><subject>Time frequency analysis</subject><subject>Voltage-controlled oscillators</subject><issn>1520-6149</issn><issn>2379-190X</issn><isbn>9781457705380</isbn><isbn>1457705389</isbn><isbn>1457705397</isbn><isbn>9781457705373</isbn><isbn>9781457705397</isbn><isbn>1457705370</isbn><fulltext>true</fulltext><rsrctype>conference_proceeding</rsrctype><creationdate>2011</creationdate><recordtype>conference_proceeding</recordtype><sourceid>6IE</sourceid><sourceid>RIE</sourceid><recordid>eNo1UMtqwzAQVF9QN80X5KIfcLprSZZ1LKEvMDQ0OeQWJHuFVRo7WE5L_r4uSecy7AwzsMPYDGGOCObhbfG4Wi3nGSDOlZFaFeqC3aFUWoMSRl-yJBPapGhgc8WmRhf_XgHXLEGVQZqjNLdsGuMnjMgzrZVJ2Oajc4c48Lgnqhre076nSO1gh9C1vPP8uwsV1Tx2h7aO3Nk4HqMTj23V9F175DUN1O9Ce0r8hKHhy7KM9-zG269I0zNP2Pr5ab14Tcv3l_GdMg2o1ZB6KdHnKCx441wuvBAVKCsq42oCl7vCVpmxDlHCnw4kayGVz2otZUFiwman2kBE230fdrY_bs8biV83_lha</recordid><startdate>201105</startdate><enddate>201105</enddate><creator>Pelle, Patricia</creator><creator>Estienne, Claudio</creator><creator>Franco, Horacio</creator><general>IEEE</general><scope>6IE</scope><scope>6IH</scope><scope>CBEJK</scope><scope>RIE</scope><scope>RIO</scope></search><sort><creationdate>201105</creationdate><title>Robust speech representation of voiced sounds based on synchrony determination with PLLs</title><author>Pelle, Patricia ; Estienne, Claudio ; Franco, Horacio</author></sort><facets><frbrtype>5</frbrtype><frbrgroupid>cdi_FETCH-LOGICAL-i175t-f441f613a0f9bb63f33c05a3c9bde0b6b8ac29ab114005a30e4d345f2d7448e3</frbrgroupid><rsrctype>conference_proceedings</rsrctype><prefilter>conference_proceedings</prefilter><language>eng</language><creationdate>2011</creationdate><topic>auditory system</topic><topic>Histograms</topic><topic>noise</topic><topic>Noise measurement</topic><topic>Phase locked loops</topic><topic>PLL</topic><topic>Robustness</topic><topic>Speech</topic><topic>speech features</topic><topic>Time frequency analysis</topic><topic>Voltage-controlled oscillators</topic><toplevel>online_resources</toplevel><creatorcontrib>Pelle, Patricia</creatorcontrib><creatorcontrib>Estienne, Claudio</creatorcontrib><creatorcontrib>Franco, Horacio</creatorcontrib><collection>IEEE Electronic Library (IEL) Conference Proceedings</collection><collection>IEEE Proceedings Order Plan (POP) 1998-present by volume</collection><collection>IEEE Xplore All Conference Proceedings</collection><collection>IEEE Electronic Library (IEL)</collection><collection>IEEE Proceedings Order Plans (POP) 1998-present</collection></facets><delivery><delcategory>Remote Search Resource</delcategory><fulltext>fulltext_linktorsrc</fulltext></delivery><addata><au>Pelle, Patricia</au><au>Estienne, Claudio</au><au>Franco, Horacio</au><format>book</format><genre>proceeding</genre><ristype>CONF</ristype><atitle>Robust speech representation of voiced sounds based on synchrony determination with PLLs</atitle><btitle>2011 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)</btitle><stitle>ICASSP</stitle><date>2011-05</date><risdate>2011</risdate><spage>5424</spage><epage>5427</epage><pages>5424-5427</pages><issn>1520-6149</issn><eissn>2379-190X</eissn><isbn>9781457705380</isbn><isbn>1457705389</isbn><eisbn>1457705397</eisbn><eisbn>9781457705373</eisbn><eisbn>9781457705397</eisbn><eisbn>1457705370</eisbn><abstract>We propose to include synchrony effects, known to exist in the auditory system, to represent voiced parts of the speech signal in a robust way. The system decomposes the input signal by means of a bandpass filter bank, and utilizes a bank of phase locked loops (PLLs) to obtain information on the frequencies present at a specific time. This information about the frequency distribution is transformed into a spectral-like representation based on synchrony effects. Noisy speech recognition experiments are performed using this synchrony-based spectrum, which is transformed into a small set of coefficients by using a transformation similar to that utilized for mel cepstrum features. We show that recognition performance compared to mel cepstrum features is advantageous, when measured over a range of SNR conditions, especially in the high noise level case.</abstract><pub>IEEE</pub><doi>10.1109/ICASSP.2011.5947585</doi><tpages>4</tpages></addata></record>
fulltext fulltext_linktorsrc
identifier ISSN: 1520-6149
ispartof 2011 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2011, p.5424-5427
issn 1520-6149
2379-190X
language eng
recordid cdi_ieee_primary_5947585
source IEEE Electronic Library (IEL) Conference Proceedings
subjects auditory system
Histograms
noise
Noise measurement
Phase locked loops
PLL
Robustness
Speech
speech features
Time frequency analysis
Voltage-controlled oscillators
title Robust speech representation of voiced sounds based on synchrony determination with PLLs
url https://sfx.bib-bvb.de/sfx_tum?ctx_ver=Z39.88-2004&ctx_enc=info:ofi/enc:UTF-8&ctx_tim=2024-12-14T20%3A20%3A29IST&url_ver=Z39.88-2004&url_ctx_fmt=infofi/fmt:kev:mtx:ctx&rfr_id=info:sid/primo.exlibrisgroup.com:primo3-Article-ieee_6IE&rft_val_fmt=info:ofi/fmt:kev:mtx:book&rft.genre=proceeding&rft.atitle=Robust%20speech%20representation%20of%20voiced%20sounds%20based%20on%20synchrony%20determination%20with%20PLLs&rft.btitle=2011%20IEEE%20International%20Conference%20on%20Acoustics,%20Speech%20and%20Signal%20Processing%20(ICASSP)&rft.au=Pelle,%20Patricia&rft.date=2011-05&rft.spage=5424&rft.epage=5427&rft.pages=5424-5427&rft.issn=1520-6149&rft.eissn=2379-190X&rft.isbn=9781457705380&rft.isbn_list=1457705389&rft_id=info:doi/10.1109/ICASSP.2011.5947585&rft_dat=%3Cieee_6IE%3E5947585%3C/ieee_6IE%3E%3Curl%3E%3C/url%3E&rft.eisbn=1457705397&rft.eisbn_list=9781457705373&rft.eisbn_list=9781457705397&rft.eisbn_list=1457705370&disable_directlink=true&sfx.directlink=off&sfx.report_link=0&rft_id=info:oai/&rft_id=info:pmid/&rft_ieee_id=5947585&rfr_iscdi=true