ODAQ: Open Dataset of Audio Quality

Research into the prediction and analysis of perceived audio quality is hampered by the scarcity of openly available datasets of audio signals accompanied by corresponding subjective quality scores. To address this problem, we present the Open Dataset of Audio Quality (ODAQ), a new dataset containin...

Ausführliche Beschreibung

Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Torcoli, Matteo, Wu, Chih-Wei, Dick, Sascha, Williams, Phillip A, Halimeh, Mhd Modar, Wolcott, William, Habets, Emanuel A. P
Format: Artikel
Sprache:eng
Online-Zugang:Volltext bestellen
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
container_end_page
container_issue
container_start_page
container_title
container_volume
creator Torcoli, Matteo
Wu, Chih-Wei
Dick, Sascha
Williams, Phillip A
Halimeh, Mhd Modar
Wolcott, William
Habets, Emanuel A. P
description Research into the prediction and analysis of perceived audio quality is hampered by the scarcity of openly available datasets of audio signals accompanied by corresponding subjective quality scores. To address this problem, we present the Open Dataset of Audio Quality (ODAQ), a new dataset containing the results of a MUSHRA listening test conducted with expert listeners from 2 international laboratories. ODAQ contains 240 audio samples and corresponding quality scores. Each audio sample is rated by 26 listeners. The audio samples are stereo audio signals sampled at 44.1 or 48 kHz and are processed by a total of 6 method classes, each operating at different quality levels. The processing method classes are designed to generate quality degradations possibly encountered during audio coding and source separation, and the quality levels for each method class span the entire quality range. The diversity of the processing methods, the large span of quality levels, the high sampling frequency, and the pool of international listeners make ODAQ particularly suited for further research into subjective and objective audio quality. The dataset is released with permissive licenses, and the software used to conduct the listening test is also made publicly available.
doi_str_mv 10.48550/arxiv.2401.00197
format Article
fullrecord <record><control><sourceid>arxiv_GOX</sourceid><recordid>TN_cdi_arxiv_primary_2401_00197</recordid><sourceformat>XML</sourceformat><sourcesystem>PC</sourcesystem><sourcerecordid>2401_00197</sourcerecordid><originalsourceid>FETCH-LOGICAL-a677-98cc1bd540346ef3b6c9fa0300b9287d85a8ee5a4c1fcb821c446bc666f4c4483</originalsourceid><addsrcrecordid>eNotzskKwjAUheFsXIj6AK4MuG69aYam7krrBEIR3JebNIGCE7UV-_aOq_OvDh8hUwah0FLCAptn_QgjASwEYEk8JPMiTw9LWtzchebY4t219Opp2lX1lR46PNVtPyYDj6e7m_x3RI7r1THbBvtis8vSfYAqjoNEW8tMJQVwoZznRtnEI3AAk0Q6rrRE7ZxEYZm3RkfMCqGMVUp58U7NR2T2u_0qy1tTn7Hpy4-2_Gr5C68uN9s</addsrcrecordid><sourcetype>Open Access Repository</sourcetype><iscdi>true</iscdi><recordtype>article</recordtype></control><display><type>article</type><title>ODAQ: Open Dataset of Audio Quality</title><source>arXiv.org</source><creator>Torcoli, Matteo ; Wu, Chih-Wei ; Dick, Sascha ; Williams, Phillip A ; Halimeh, Mhd Modar ; Wolcott, William ; Habets, Emanuel A. P</creator><creatorcontrib>Torcoli, Matteo ; Wu, Chih-Wei ; Dick, Sascha ; Williams, Phillip A ; Halimeh, Mhd Modar ; Wolcott, William ; Habets, Emanuel A. P</creatorcontrib><description>Research into the prediction and analysis of perceived audio quality is hampered by the scarcity of openly available datasets of audio signals accompanied by corresponding subjective quality scores. To address this problem, we present the Open Dataset of Audio Quality (ODAQ), a new dataset containing the results of a MUSHRA listening test conducted with expert listeners from 2 international laboratories. ODAQ contains 240 audio samples and corresponding quality scores. Each audio sample is rated by 26 listeners. The audio samples are stereo audio signals sampled at 44.1 or 48 kHz and are processed by a total of 6 method classes, each operating at different quality levels. The processing method classes are designed to generate quality degradations possibly encountered during audio coding and source separation, and the quality levels for each method class span the entire quality range. The diversity of the processing methods, the large span of quality levels, the high sampling frequency, and the pool of international listeners make ODAQ particularly suited for further research into subjective and objective audio quality. The dataset is released with permissive licenses, and the software used to conduct the listening test is also made publicly available.</description><identifier>DOI: 10.48550/arxiv.2401.00197</identifier><language>eng</language><creationdate>2023-12</creationdate><rights>http://arxiv.org/licenses/nonexclusive-distrib/1.0</rights><oa>free_for_read</oa><woscitedreferencessubscribed>false</woscitedreferencessubscribed></display><links><openurl>$$Topenurl_article</openurl><openurlfulltext>$$Topenurlfull_article</openurlfulltext><thumbnail>$$Tsyndetics_thumb_exl</thumbnail><link.rule.ids>228,230,780,885</link.rule.ids><linktorsrc>$$Uhttps://arxiv.org/abs/2401.00197$$EView_record_in_Cornell_University$$FView_record_in_$$GCornell_University$$Hfree_for_read</linktorsrc><backlink>$$Uhttps://doi.org/10.48550/arXiv.2401.00197$$DView paper in arXiv$$Hfree_for_read</backlink></links><search><creatorcontrib>Torcoli, Matteo</creatorcontrib><creatorcontrib>Wu, Chih-Wei</creatorcontrib><creatorcontrib>Dick, Sascha</creatorcontrib><creatorcontrib>Williams, Phillip A</creatorcontrib><creatorcontrib>Halimeh, Mhd Modar</creatorcontrib><creatorcontrib>Wolcott, William</creatorcontrib><creatorcontrib>Habets, Emanuel A. P</creatorcontrib><title>ODAQ: Open Dataset of Audio Quality</title><description>Research into the prediction and analysis of perceived audio quality is hampered by the scarcity of openly available datasets of audio signals accompanied by corresponding subjective quality scores. To address this problem, we present the Open Dataset of Audio Quality (ODAQ), a new dataset containing the results of a MUSHRA listening test conducted with expert listeners from 2 international laboratories. ODAQ contains 240 audio samples and corresponding quality scores. Each audio sample is rated by 26 listeners. The audio samples are stereo audio signals sampled at 44.1 or 48 kHz and are processed by a total of 6 method classes, each operating at different quality levels. The processing method classes are designed to generate quality degradations possibly encountered during audio coding and source separation, and the quality levels for each method class span the entire quality range. The diversity of the processing methods, the large span of quality levels, the high sampling frequency, and the pool of international listeners make ODAQ particularly suited for further research into subjective and objective audio quality. The dataset is released with permissive licenses, and the software used to conduct the listening test is also made publicly available.</description><fulltext>true</fulltext><rsrctype>article</rsrctype><creationdate>2023</creationdate><recordtype>article</recordtype><sourceid>GOX</sourceid><recordid>eNotzskKwjAUheFsXIj6AK4MuG69aYam7krrBEIR3JebNIGCE7UV-_aOq_OvDh8hUwah0FLCAptn_QgjASwEYEk8JPMiTw9LWtzchebY4t219Opp2lX1lR46PNVtPyYDj6e7m_x3RI7r1THbBvtis8vSfYAqjoNEW8tMJQVwoZznRtnEI3AAk0Q6rrRE7ZxEYZm3RkfMCqGMVUp58U7NR2T2u_0qy1tTn7Hpy4-2_Gr5C68uN9s</recordid><startdate>20231230</startdate><enddate>20231230</enddate><creator>Torcoli, Matteo</creator><creator>Wu, Chih-Wei</creator><creator>Dick, Sascha</creator><creator>Williams, Phillip A</creator><creator>Halimeh, Mhd Modar</creator><creator>Wolcott, William</creator><creator>Habets, Emanuel A. P</creator><scope>GOX</scope></search><sort><creationdate>20231230</creationdate><title>ODAQ: Open Dataset of Audio Quality</title><author>Torcoli, Matteo ; Wu, Chih-Wei ; Dick, Sascha ; Williams, Phillip A ; Halimeh, Mhd Modar ; Wolcott, William ; Habets, Emanuel A. P</author></sort><facets><frbrtype>5</frbrtype><frbrgroupid>cdi_FETCH-LOGICAL-a677-98cc1bd540346ef3b6c9fa0300b9287d85a8ee5a4c1fcb821c446bc666f4c4483</frbrgroupid><rsrctype>articles</rsrctype><prefilter>articles</prefilter><language>eng</language><creationdate>2023</creationdate><toplevel>online_resources</toplevel><creatorcontrib>Torcoli, Matteo</creatorcontrib><creatorcontrib>Wu, Chih-Wei</creatorcontrib><creatorcontrib>Dick, Sascha</creatorcontrib><creatorcontrib>Williams, Phillip A</creatorcontrib><creatorcontrib>Halimeh, Mhd Modar</creatorcontrib><creatorcontrib>Wolcott, William</creatorcontrib><creatorcontrib>Habets, Emanuel A. P</creatorcontrib><collection>arXiv.org</collection></facets><delivery><delcategory>Remote Search Resource</delcategory><fulltext>fulltext_linktorsrc</fulltext></delivery><addata><au>Torcoli, Matteo</au><au>Wu, Chih-Wei</au><au>Dick, Sascha</au><au>Williams, Phillip A</au><au>Halimeh, Mhd Modar</au><au>Wolcott, William</au><au>Habets, Emanuel A. P</au><format>journal</format><genre>article</genre><ristype>JOUR</ristype><atitle>ODAQ: Open Dataset of Audio Quality</atitle><date>2023-12-30</date><risdate>2023</risdate><abstract>Research into the prediction and analysis of perceived audio quality is hampered by the scarcity of openly available datasets of audio signals accompanied by corresponding subjective quality scores. To address this problem, we present the Open Dataset of Audio Quality (ODAQ), a new dataset containing the results of a MUSHRA listening test conducted with expert listeners from 2 international laboratories. ODAQ contains 240 audio samples and corresponding quality scores. Each audio sample is rated by 26 listeners. The audio samples are stereo audio signals sampled at 44.1 or 48 kHz and are processed by a total of 6 method classes, each operating at different quality levels. The processing method classes are designed to generate quality degradations possibly encountered during audio coding and source separation, and the quality levels for each method class span the entire quality range. The diversity of the processing methods, the large span of quality levels, the high sampling frequency, and the pool of international listeners make ODAQ particularly suited for further research into subjective and objective audio quality. The dataset is released with permissive licenses, and the software used to conduct the listening test is also made publicly available.</abstract><doi>10.48550/arxiv.2401.00197</doi><oa>free_for_read</oa></addata></record>
fulltext fulltext_linktorsrc
identifier DOI: 10.48550/arxiv.2401.00197
ispartof
issn
language eng
recordid cdi_arxiv_primary_2401_00197
source arXiv.org
title ODAQ: Open Dataset of Audio Quality
url https://sfx.bib-bvb.de/sfx_tum?ctx_ver=Z39.88-2004&ctx_enc=info:ofi/enc:UTF-8&ctx_tim=2025-01-04T21%3A16%3A04IST&url_ver=Z39.88-2004&url_ctx_fmt=infofi/fmt:kev:mtx:ctx&rfr_id=info:sid/primo.exlibrisgroup.com:primo3-Article-arxiv_GOX&rft_val_fmt=info:ofi/fmt:kev:mtx:journal&rft.genre=article&rft.atitle=ODAQ:%20Open%20Dataset%20of%20Audio%20Quality&rft.au=Torcoli,%20Matteo&rft.date=2023-12-30&rft_id=info:doi/10.48550/arxiv.2401.00197&rft_dat=%3Carxiv_GOX%3E2401_00197%3C/arxiv_GOX%3E%3Curl%3E%3C/url%3E&disable_directlink=true&sfx.directlink=off&sfx.report_link=0&rft_id=info:oai/&rft_id=info:pmid/&rfr_iscdi=true