Estimating vocal tract length by minimizing non-uniformity of cross-sectional area
Previous approaches to estimation of vocal tract length (VTL) differ in what information about the speaker is assumed to be known and which formants are treated as better predictors of VTL. However, they are alike in modeling formant frequencies as deviations from the resonances of a uniform tube, a...
Gespeichert in:
1. Verfasser: | |
---|---|
Format: | Tagungsbericht |
Sprache: | eng |
Online-Zugang: | Volltext |
Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
container_end_page | |
---|---|
container_issue | 1 |
container_start_page | |
container_title | |
container_volume | 35 |
creator | Flego, Stefon |
description | Previous approaches to estimation of vocal tract length (VTL) differ in what information about the speaker is assumed to be known and which formants are treated as better predictors of VTL. However, they are alike in modeling formant frequencies as deviations from the resonances of a uniform tube, and in allocating equal credibility to all vowel spectra as predictors of length. The latter may be problematic, as phonotactic asymmetries in vowel quality in a data set can draw formants’ mean frequencies away from their underlying resonances, skewing VTL estimates. Herein, an additional parameter is proposed privileging vowel spectra that approximate the resonance characteristics of a uniform tube. The metric for this proximity is standard variance in Phi (SigmaPhi) across a vowel spectrum. In this study, formant data from 32 participants were analyzed using the estimators compared in Lammert & Narayanan (2015). Each estimator was run using frequencies from all vowel spectra, as well as from spectra with low values of SigmaPhi. As the threshold for maximum SigmaPhi was lowered, VTL estimates became tighter and the estimators typically converged on a length. This approach requires no labeling or identification of vowel type, and is therefore easily replicable across languages and speakers. |
doi_str_mv | 10.1121/2.0001000 |
format | Conference Proceeding |
fullrecord | <record><control><sourceid>scitation</sourceid><recordid>TN_cdi_scitation_primary_10_1121_2_0001000</recordid><sourceformat>XML</sourceformat><sourcesystem>PC</sourcesystem><sourcerecordid>poma</sourcerecordid><originalsourceid>FETCH-LOGICAL-s1050-dedc5f92d7dffce915867d99b71908ff7383a53521dacd7ffdfcddd42b7abf693</originalsourceid><addsrcrecordid>eNp9UEtLAzEYDIJgrR78BzkLW_Nl3UeOUuoDCoIoeAvZJF-N7CYliYX117uLBW8ehrnMDDNDyBWwFQCHG75ijMGEE7IAUYqiZez9jJyn9MlYDbyuFuRlk7IbVHZ-Rw9Bq57mqHSmvfW7_EG7kQ7Ou8F9zwIffPHlHYY4uDzSgFTHkFKRrM4u-MmsolUX5BRVn-zlkZfk7X7zun4sts8PT-u7bZGAVaww1ugKBTeNQdRWQNXWjRGia0CwFrEp21JVZcXBKG0aRIPaGHPLu0Z1WItySa5_c5N2Wc0F5D5OW-IoDyFKLo_r5d7gf2Jgcr7rz1D-ALqpYO8</addsrcrecordid><sourcetype>Enrichment Source</sourcetype><iscdi>true</iscdi><recordtype>conference_proceeding</recordtype></control><display><type>conference_proceeding</type><title>Estimating vocal tract length by minimizing non-uniformity of cross-sectional area</title><source>AIP Journals Complete</source><source>Elektronische Zeitschriftenbibliothek - Frei zugängliche E-Journals</source><creator>Flego, Stefon</creator><creatorcontrib>Flego, Stefon</creatorcontrib><description>Previous approaches to estimation of vocal tract length (VTL) differ in what information about the speaker is assumed to be known and which formants are treated as better predictors of VTL. However, they are alike in modeling formant frequencies as deviations from the resonances of a uniform tube, and in allocating equal credibility to all vowel spectra as predictors of length. The latter may be problematic, as phonotactic asymmetries in vowel quality in a data set can draw formants’ mean frequencies away from their underlying resonances, skewing VTL estimates. Herein, an additional parameter is proposed privileging vowel spectra that approximate the resonance characteristics of a uniform tube. The metric for this proximity is standard variance in Phi (SigmaPhi) across a vowel spectrum. In this study, formant data from 32 participants were analyzed using the estimators compared in Lammert & Narayanan (2015). Each estimator was run using frequencies from all vowel spectra, as well as from spectra with low values of SigmaPhi. As the threshold for maximum SigmaPhi was lowered, VTL estimates became tighter and the estimators typically converged on a length. This approach requires no labeling or identification of vowel type, and is therefore easily replicable across languages and speakers.</description><identifier>EISSN: 1939-800X</identifier><identifier>DOI: 10.1121/2.0001000</identifier><identifier>CODEN: PMARCW</identifier><language>eng</language><ispartof>Proceedings of Meetings on Acoustics, 2018, Vol.35 (1)</ispartof><rights>Acoustical Society of America</rights><lds50>peer_reviewed</lds50><oa>free_for_read</oa><woscitedreferencessubscribed>false</woscitedreferencessubscribed></display><links><openurl>$$Topenurl_article</openurl><openurlfulltext>$$Topenurlfull_article</openurlfulltext><thumbnail>$$Tsyndetics_thumb_exl</thumbnail><linktohtml>$$Uhttps://pubs.aip.org/poma/article-lookup/doi/10.1121/2.0001000$$EHTML$$P50$$Gscitation$$H</linktohtml><link.rule.ids>208,309,310,776,780,785,786,790,4498,23909,23910,25118,27902,76127</link.rule.ids></links><search><creatorcontrib>Flego, Stefon</creatorcontrib><title>Estimating vocal tract length by minimizing non-uniformity of cross-sectional area</title><title>Proceedings of Meetings on Acoustics</title><description>Previous approaches to estimation of vocal tract length (VTL) differ in what information about the speaker is assumed to be known and which formants are treated as better predictors of VTL. However, they are alike in modeling formant frequencies as deviations from the resonances of a uniform tube, and in allocating equal credibility to all vowel spectra as predictors of length. The latter may be problematic, as phonotactic asymmetries in vowel quality in a data set can draw formants’ mean frequencies away from their underlying resonances, skewing VTL estimates. Herein, an additional parameter is proposed privileging vowel spectra that approximate the resonance characteristics of a uniform tube. The metric for this proximity is standard variance in Phi (SigmaPhi) across a vowel spectrum. In this study, formant data from 32 participants were analyzed using the estimators compared in Lammert & Narayanan (2015). Each estimator was run using frequencies from all vowel spectra, as well as from spectra with low values of SigmaPhi. As the threshold for maximum SigmaPhi was lowered, VTL estimates became tighter and the estimators typically converged on a length. This approach requires no labeling or identification of vowel type, and is therefore easily replicable across languages and speakers.</description><issn>1939-800X</issn><fulltext>true</fulltext><rsrctype>conference_proceeding</rsrctype><creationdate>2018</creationdate><recordtype>conference_proceeding</recordtype><sourceid/><recordid>eNp9UEtLAzEYDIJgrR78BzkLW_Nl3UeOUuoDCoIoeAvZJF-N7CYliYX117uLBW8ehrnMDDNDyBWwFQCHG75ijMGEE7IAUYqiZez9jJyn9MlYDbyuFuRlk7IbVHZ-Rw9Bq57mqHSmvfW7_EG7kQ7Ou8F9zwIffPHlHYY4uDzSgFTHkFKRrM4u-MmsolUX5BRVn-zlkZfk7X7zun4sts8PT-u7bZGAVaww1ugKBTeNQdRWQNXWjRGia0CwFrEp21JVZcXBKG0aRIPaGHPLu0Z1WItySa5_c5N2Wc0F5D5OW-IoDyFKLo_r5d7gf2Jgcr7rz1D-ALqpYO8</recordid><startdate>20181105</startdate><enddate>20181105</enddate><creator>Flego, Stefon</creator><scope/></search><sort><creationdate>20181105</creationdate><title>Estimating vocal tract length by minimizing non-uniformity of cross-sectional area</title><author>Flego, Stefon</author></sort><facets><frbrtype>5</frbrtype><frbrgroupid>cdi_FETCH-LOGICAL-s1050-dedc5f92d7dffce915867d99b71908ff7383a53521dacd7ffdfcddd42b7abf693</frbrgroupid><rsrctype>conference_proceedings</rsrctype><prefilter>conference_proceedings</prefilter><language>eng</language><creationdate>2018</creationdate><toplevel>peer_reviewed</toplevel><toplevel>online_resources</toplevel><creatorcontrib>Flego, Stefon</creatorcontrib></facets><delivery><delcategory>Remote Search Resource</delcategory><fulltext>fulltext</fulltext></delivery><addata><au>Flego, Stefon</au><format>book</format><genre>proceeding</genre><ristype>CONF</ristype><atitle>Estimating vocal tract length by minimizing non-uniformity of cross-sectional area</atitle><btitle>Proceedings of Meetings on Acoustics</btitle><date>2018-11-05</date><risdate>2018</risdate><volume>35</volume><issue>1</issue><eissn>1939-800X</eissn><coden>PMARCW</coden><abstract>Previous approaches to estimation of vocal tract length (VTL) differ in what information about the speaker is assumed to be known and which formants are treated as better predictors of VTL. However, they are alike in modeling formant frequencies as deviations from the resonances of a uniform tube, and in allocating equal credibility to all vowel spectra as predictors of length. The latter may be problematic, as phonotactic asymmetries in vowel quality in a data set can draw formants’ mean frequencies away from their underlying resonances, skewing VTL estimates. Herein, an additional parameter is proposed privileging vowel spectra that approximate the resonance characteristics of a uniform tube. The metric for this proximity is standard variance in Phi (SigmaPhi) across a vowel spectrum. In this study, formant data from 32 participants were analyzed using the estimators compared in Lammert & Narayanan (2015). Each estimator was run using frequencies from all vowel spectra, as well as from spectra with low values of SigmaPhi. As the threshold for maximum SigmaPhi was lowered, VTL estimates became tighter and the estimators typically converged on a length. This approach requires no labeling or identification of vowel type, and is therefore easily replicable across languages and speakers.</abstract><doi>10.1121/2.0001000</doi><tpages>9</tpages><oa>free_for_read</oa></addata></record> |
fulltext | fulltext |
identifier | EISSN: 1939-800X |
ispartof | Proceedings of Meetings on Acoustics, 2018, Vol.35 (1) |
issn | 1939-800X |
language | eng |
recordid | cdi_scitation_primary_10_1121_2_0001000 |
source | AIP Journals Complete; Elektronische Zeitschriftenbibliothek - Frei zugängliche E-Journals |
title | Estimating vocal tract length by minimizing non-uniformity of cross-sectional area |
url | https://sfx.bib-bvb.de/sfx_tum?ctx_ver=Z39.88-2004&ctx_enc=info:ofi/enc:UTF-8&ctx_tim=2025-01-28T14%3A04%3A16IST&url_ver=Z39.88-2004&url_ctx_fmt=infofi/fmt:kev:mtx:ctx&rfr_id=info:sid/primo.exlibrisgroup.com:primo3-Article-scitation&rft_val_fmt=info:ofi/fmt:kev:mtx:book&rft.genre=proceeding&rft.atitle=Estimating%20vocal%20tract%20length%20by%20minimizing%20non-uniformity%20of%20cross-sectional%20area&rft.btitle=Proceedings%20of%20Meetings%20on%20Acoustics&rft.au=Flego,%20Stefon&rft.date=2018-11-05&rft.volume=35&rft.issue=1&rft.eissn=1939-800X&rft.coden=PMARCW&rft_id=info:doi/10.1121/2.0001000&rft_dat=%3Cscitation%3Epoma%3C/scitation%3E%3Curl%3E%3C/url%3E&disable_directlink=true&sfx.directlink=off&sfx.report_link=0&rft_id=info:oai/&rft_id=info:pmid/&rfr_iscdi=true |