OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset

We introduce OpenDebateEvidence, a comprehensive dataset for argument mining and summarization sourced from the American Competitive Debate community. This dataset includes over 3.5 million documents with rich metadata, making it one of the most extensive collections of debate evidence. OpenDebateEv...

Ausführliche Beschreibung

Gespeichert in:

Bibliographische Detailangaben
Hauptverfasser:	Roush, Allen, Shabazz, Yusuf, Balaji, Arvind, Zhang, Peter, Mezza, Stefano, Zhang, Markus, Basu, Sanjay, Vishwanath, Sriram, Fatemi, Mehdi, Shwartz-Ziv, Ravid
Format:	Artikel
Sprache:	eng
Schlagworte:	Computer Science - Artificial Intelligence Computer Science - Computation and Language Computer Science - Learning
Online-Zugang:	Volltext bestellen
Tags:	Tag hinzufügen Keine Tags, Fügen Sie den ersten Tag hinzu!

container_end_page
container_issue
container_start_page
container_title
container_volume
creator	Roush, Allen Shabazz, Yusuf Balaji, Arvind Zhang, Peter Mezza, Stefano Zhang, Markus Basu, Sanjay Vishwanath, Sriram Fatemi, Mehdi Shwartz-Ziv, Ravid
description	We introduce OpenDebateEvidence, a comprehensive dataset for argument mining and summarization sourced from the American Competitive Debate community. This dataset includes over 3.5 million documents with rich metadata, making it one of the most extensive collections of debate evidence. OpenDebateEvidence captures the complexity of arguments in high school and college debates, providing valuable resources for training and evaluation. Our extensive experiments demonstrate the efficacy of fine-tuning state-of-the-art large language models for argumentative abstractive summarization across various methods, models, and datasets. By providing this comprehensive resource, we aim to advance computational argumentation and support practical applications for debaters, educators, and researchers. OpenDebateEvidence is publicly available to support further research and innovation in computational argumentation. Access it here: https://huggingface.co/datasets/Yusuf5/OpenCaselist
doi_str_mv	10.48550/arxiv.2406.14657
format	Article
fullrecord	<record><control><sourceid>arxiv_GOX</sourceid><recordid>TN_cdi_arxiv_primary_2406_14657</recordid><sourceformat>XML</sourceformat><sourcesystem>PC</sourcesystem><sourcerecordid>2406_14657</sourcerecordid><originalsourceid>FETCH-LOGICAL-a677-4116123cd16ceaaaf998e639a343a6ad86ec608b723a8352e674eabe973367403</originalsourceid><addsrcrecordid>eNotz81uwjAQBGBfeqigD9BT_QJJ46yzdrhFQH8kIg5wjzbOgiwRFyUmavv0pbSnmdNoPiEeVZZqWxTZMw2ffkpznWGqNBbmXtTbM4cVtxR5PfmOg-OFrGRN4-gnTnaOTiyr4XjpOURZ--DDUVLo5O7S9zT4b4r-I8gVRRo5zsXdgU4jP_znTOxf1vvlW7LZvr4vq01CaEyilUKVg-sUOiaiQ1laRigJNBBSZ5EdZrY1OZCFImc0mqnl0gBcawYz8fQ3e_M058Ffr3w1v67m5oIf9NBH1g</addsrcrecordid><sourcetype>Open Access Repository</sourcetype><iscdi>true</iscdi><recordtype>article</recordtype></control><display><type>article</type><title>OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset</title><source>arXiv.org</source><creator>Roush, Allen ; Shabazz, Yusuf ; Balaji, Arvind ; Zhang, Peter ; Mezza, Stefano ; Zhang, Markus ; Basu, Sanjay ; Vishwanath, Sriram ; Fatemi, Mehdi ; Shwartz-Ziv, Ravid</creator><creatorcontrib>Roush, Allen ; Shabazz, Yusuf ; Balaji, Arvind ; Zhang, Peter ; Mezza, Stefano ; Zhang, Markus ; Basu, Sanjay ; Vishwanath, Sriram ; Fatemi, Mehdi ; Shwartz-Ziv, Ravid</creatorcontrib><description>We introduce OpenDebateEvidence, a comprehensive dataset for argument mining and summarization sourced from the American Competitive Debate community. This dataset includes over 3.5 million documents with rich metadata, making it one of the most extensive collections of debate evidence. OpenDebateEvidence captures the complexity of arguments in high school and college debates, providing valuable resources for training and evaluation. Our extensive experiments demonstrate the efficacy of fine-tuning state-of-the-art large language models for argumentative abstractive summarization across various methods, models, and datasets. By providing this comprehensive resource, we aim to advance computational argumentation and support practical applications for debaters, educators, and researchers. OpenDebateEvidence is publicly available to support further research and innovation in computational argumentation. Access it here: https://huggingface.co/datasets/Yusuf5/OpenCaselist</description><identifier>DOI: 10.48550/arxiv.2406.14657</identifier><language>eng</language><subject>Computer Science - Artificial Intelligence ; Computer Science - Computation and Language ; Computer Science - Learning</subject><creationdate>2024-06</creationdate><rights>http://creativecommons.org/licenses/by/4.0</rights><oa>free_for_read</oa><woscitedreferencessubscribed>false</woscitedreferencessubscribed></display><links><openurl>$$Topenurl_article</openurl><openurlfulltext>$$Topenurlfull_article</openurlfulltext><thumbnail>$$Tsyndetics_thumb_exl</thumbnail><link.rule.ids>228,230,776,881</link.rule.ids><linktorsrc>$$Uhttps://arxiv.org/abs/2406.14657$$EView_record_in_Cornell_University$$FView_record_in_$$GCornell_University$$Hfree_for_read</linktorsrc><backlink>$$Uhttps://doi.org/10.48550/arXiv.2406.14657$$DView paper in arXiv$$Hfree_for_read</backlink></links><search><creatorcontrib>Roush, Allen</creatorcontrib><creatorcontrib>Shabazz, Yusuf</creatorcontrib><creatorcontrib>Balaji, Arvind</creatorcontrib><creatorcontrib>Zhang, Peter</creatorcontrib><creatorcontrib>Mezza, Stefano</creatorcontrib><creatorcontrib>Zhang, Markus</creatorcontrib><creatorcontrib>Basu, Sanjay</creatorcontrib><creatorcontrib>Vishwanath, Sriram</creatorcontrib><creatorcontrib>Fatemi, Mehdi</creatorcontrib><creatorcontrib>Shwartz-Ziv, Ravid</creatorcontrib><title>OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset</title><description>We introduce OpenDebateEvidence, a comprehensive dataset for argument mining and summarization sourced from the American Competitive Debate community. This dataset includes over 3.5 million documents with rich metadata, making it one of the most extensive collections of debate evidence. OpenDebateEvidence captures the complexity of arguments in high school and college debates, providing valuable resources for training and evaluation. Our extensive experiments demonstrate the efficacy of fine-tuning state-of-the-art large language models for argumentative abstractive summarization across various methods, models, and datasets. By providing this comprehensive resource, we aim to advance computational argumentation and support practical applications for debaters, educators, and researchers. OpenDebateEvidence is publicly available to support further research and innovation in computational argumentation. Access it here: https://huggingface.co/datasets/Yusuf5/OpenCaselist</description><subject>Computer Science - Artificial Intelligence</subject><subject>Computer Science - Computation and Language</subject><subject>Computer Science - Learning</subject><fulltext>true</fulltext><rsrctype>article</rsrctype><creationdate>2024</creationdate><recordtype>article</recordtype><sourceid>GOX</sourceid><recordid>eNotz81uwjAQBGBfeqigD9BT_QJJ46yzdrhFQH8kIg5wjzbOgiwRFyUmavv0pbSnmdNoPiEeVZZqWxTZMw2ffkpznWGqNBbmXtTbM4cVtxR5PfmOg-OFrGRN4-gnTnaOTiyr4XjpOURZ--DDUVLo5O7S9zT4b4r-I8gVRRo5zsXdgU4jP_znTOxf1vvlW7LZvr4vq01CaEyilUKVg-sUOiaiQ1laRigJNBBSZ5EdZrY1OZCFImc0mqnl0gBcawYz8fQ3e_M058Ffr3w1v67m5oIf9NBH1g</recordid><startdate>20240620</startdate><enddate>20240620</enddate><creator>Roush, Allen</creator><creator>Shabazz, Yusuf</creator><creator>Balaji, Arvind</creator><creator>Zhang, Peter</creator><creator>Mezza, Stefano</creator><creator>Zhang, Markus</creator><creator>Basu, Sanjay</creator><creator>Vishwanath, Sriram</creator><creator>Fatemi, Mehdi</creator><creator>Shwartz-Ziv, Ravid</creator><scope>AKY</scope><scope>GOX</scope></search><sort><creationdate>20240620</creationdate><title>OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset</title><author>Roush, Allen ; Shabazz, Yusuf ; Balaji, Arvind ; Zhang, Peter ; Mezza, Stefano ; Zhang, Markus ; Basu, Sanjay ; Vishwanath, Sriram ; Fatemi, Mehdi ; Shwartz-Ziv, Ravid</author></sort><facets><frbrtype>5</frbrtype><frbrgroupid>cdi_FETCH-LOGICAL-a677-4116123cd16ceaaaf998e639a343a6ad86ec608b723a8352e674eabe973367403</frbrgroupid><rsrctype>articles</rsrctype><prefilter>articles</prefilter><language>eng</language><creationdate>2024</creationdate><topic>Computer Science - Artificial Intelligence</topic><topic>Computer Science - Computation and Language</topic><topic>Computer Science - Learning</topic><toplevel>online_resources</toplevel><creatorcontrib>Roush, Allen</creatorcontrib><creatorcontrib>Shabazz, Yusuf</creatorcontrib><creatorcontrib>Balaji, Arvind</creatorcontrib><creatorcontrib>Zhang, Peter</creatorcontrib><creatorcontrib>Mezza, Stefano</creatorcontrib><creatorcontrib>Zhang, Markus</creatorcontrib><creatorcontrib>Basu, Sanjay</creatorcontrib><creatorcontrib>Vishwanath, Sriram</creatorcontrib><creatorcontrib>Fatemi, Mehdi</creatorcontrib><creatorcontrib>Shwartz-Ziv, Ravid</creatorcontrib><collection>arXiv Computer Science</collection><collection>arXiv.org</collection></facets><delivery><delcategory>Remote Search Resource</delcategory><fulltext>fulltext_linktorsrc</fulltext></delivery><addata><au>Roush, Allen</au><au>Shabazz, Yusuf</au><au>Balaji, Arvind</au><au>Zhang, Peter</au><au>Mezza, Stefano</au><au>Zhang, Markus</au><au>Basu, Sanjay</au><au>Vishwanath, Sriram</au><au>Fatemi, Mehdi</au><au>Shwartz-Ziv, Ravid</au><format>journal</format><genre>article</genre><ristype>JOUR</ristype><atitle>OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset</atitle><date>2024-06-20</date><risdate>2024</risdate><abstract>We introduce OpenDebateEvidence, a comprehensive dataset for argument mining and summarization sourced from the American Competitive Debate community. This dataset includes over 3.5 million documents with rich metadata, making it one of the most extensive collections of debate evidence. OpenDebateEvidence captures the complexity of arguments in high school and college debates, providing valuable resources for training and evaluation. Our extensive experiments demonstrate the efficacy of fine-tuning state-of-the-art large language models for argumentative abstractive summarization across various methods, models, and datasets. By providing this comprehensive resource, we aim to advance computational argumentation and support practical applications for debaters, educators, and researchers. OpenDebateEvidence is publicly available to support further research and innovation in computational argumentation. Access it here: https://huggingface.co/datasets/Yusuf5/OpenCaselist</abstract><doi>10.48550/arxiv.2406.14657</doi><oa>free_for_read</oa></addata></record>
fulltext	fulltext_linktorsrc
identifier	DOI: 10.48550/arxiv.2406.14657
ispartof
issn
language	eng
recordid	cdi_arxiv_primary_2406_14657
source	arXiv.org
subjects	Computer Science - Artificial Intelligence Computer Science - Computation and Language Computer Science - Learning
title	OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset
url	https://sfx.bib-bvb.de/sfx_tum?ctx_ver=Z39.88-2004&ctx_enc=info:ofi/enc:UTF-8&ctx_tim=2025-02-03T11%3A01%3A19IST&url_ver=Z39.88-2004&url_ctx_fmt=infofi/fmt:kev:mtx:ctx&rfr_id=info:sid/primo.exlibrisgroup.com:primo3-Article-arxiv_GOX&rft_val_fmt=info:ofi/fmt:kev:mtx:journal&rft.genre=article&rft.atitle=OpenDebateEvidence:%20A%20Massive-Scale%20Argument%20Mining%20and%20Summarization%20Dataset&rft.au=Roush,%20Allen&rft.date=2024-06-20&rft_id=info:doi/10.48550/arxiv.2406.14657&rft_dat=%3Carxiv_GOX%3E2406_14657%3C/arxiv_GOX%3E%3Curl%3E%3C/url%3E&disable_directlink=true&sfx.directlink=off&sfx.report_link=0&rft_id=info:oai/&rft_id=info:pmid/&rfr_iscdi=true