Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts

Large language models (LLMs) have shown impressive capabilities in tasks such as machine translation, text summarization, question answering, and solving complex mathematical problems. However, their primary training on data-rich languages like English limits their performance in low-resource langua...

Ausführliche Beschreibung

Gespeichert in:

Bibliographische Detailangaben
Hauptverfasser:	Oğuz, Metehan, Ciftci, Yusuf Umut, Bakman, Yavuz Faruk
Format:	Artikel
Sprache:	eng
Schlagworte:	Computer Science - Computation and Language
Online-Zugang:	Volltext bestellen
Tags:	Tag hinzufügen Keine Tags, Fügen Sie den ersten Tag hinzu!

container_end_page
container_issue
container_start_page
container_title
container_volume
creator	Oğuz, Metehan Ciftci, Yusuf Umut Bakman, Yavuz Faruk
description	Large language models (LLMs) have shown impressive capabilities in tasks such as machine translation, text summarization, question answering, and solving complex mathematical problems. However, their primary training on data-rich languages like English limits their performance in low-resource languages. This study addresses this gap by focusing on the Indexical Shift problem in Turkish. The Indexical Shift problem involves resolving pronouns in indexical shift contexts, a grammatical challenge not present in high-resource languages like English. We present the first study examining indexical shift in any language, releasing a Turkish dataset specifically designed for this purpose. Our Indexical Shift Dataset consists of 156 multiple-choice questions, each annotated with necessary linguistic details, to evaluate LLMs in a few-shot setting. We evaluate recent multilingual LLMs, including GPT-4, GPT-3.5, Cohere-AYA, Trendyol-LLM, and Turkcell-LLM, using this dataset. Our analysis reveals that even advanced models like GPT-4 struggle with the grammatical nuances of indexical shift in Turkish, achieving only moderate performance. These findings underscore the need for focused research on the grammatical challenges posed by low-resource languages. We released the dataset and code \href{https://anonymous.4open.science/r/indexical_shift_llm-E1B4} {here}.
doi_str_mv	10.48550/arxiv.2406.05569
format	Article
fullrecord	<record><control><sourceid>arxiv_GOX</sourceid><recordid>TN_cdi_arxiv_primary_2406_05569</recordid><sourceformat>XML</sourceformat><sourcesystem>PC</sourcesystem><sourcerecordid>2406_05569</sourcerecordid><originalsourceid>FETCH-LOGICAL-a679-b4c5767c49bb45b063728f5f3ab46112f0a89c150e9d906b71c48f19daef91a93</originalsourceid><addsrcrecordid>eNpNkMtOhDAYhdm4MKMP4Mr_AQRbaAt1N8EbCUajGJekLe3QOLSGMgZd-uTOxYWrk3wn5yy-KDrDKCEFpehSjLP9TFKCWIIoZfw4-rn2UNcPAZ618itnvzUM-gLeeu2gAhvA-WlLrmAZgg5h0G4Cbw6TV9fpMUzCddatdrTZjO829FBti9kqsYan0Tu_cQGs-0dfemsmKL2b9DyFk-jIiHXQp3-5iJrbm6a8j-vHu6pc1rFgOY8lUTRnuSJcSkIlYlmeFoaaTEjCME4NEgVXmCLNO46YzLEihcG8E9pwLHi2iM4Pt3sJ7cdoBzF-tTsZ7V5G9gtuKFpT</addsrcrecordid><sourcetype>Open Access Repository</sourcetype><iscdi>true</iscdi><recordtype>article</recordtype></control><display><type>article</type><title>Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts</title><source>arXiv.org</source><creator>Oğuz, Metehan ; Ciftci, Yusuf Umut ; Bakman, Yavuz Faruk</creator><creatorcontrib>Oğuz, Metehan ; Ciftci, Yusuf Umut ; Bakman, Yavuz Faruk</creatorcontrib><description>Large language models (LLMs) have shown impressive capabilities in tasks such as machine translation, text summarization, question answering, and solving complex mathematical problems. However, their primary training on data-rich languages like English limits their performance in low-resource languages. This study addresses this gap by focusing on the Indexical Shift problem in Turkish. The Indexical Shift problem involves resolving pronouns in indexical shift contexts, a grammatical challenge not present in high-resource languages like English. We present the first study examining indexical shift in any language, releasing a Turkish dataset specifically designed for this purpose. Our Indexical Shift Dataset consists of 156 multiple-choice questions, each annotated with necessary linguistic details, to evaluate LLMs in a few-shot setting. We evaluate recent multilingual LLMs, including GPT-4, GPT-3.5, Cohere-AYA, Trendyol-LLM, and Turkcell-LLM, using this dataset. Our analysis reveals that even advanced models like GPT-4 struggle with the grammatical nuances of indexical shift in Turkish, achieving only moderate performance. These findings underscore the need for focused research on the grammatical challenges posed by low-resource languages. We released the dataset and code \href{https://anonymous.4open.science/r/indexical_shift_llm-E1B4} {here}.</description><identifier>DOI: 10.48550/arxiv.2406.05569</identifier><language>eng</language><subject>Computer Science - Computation and Language</subject><creationdate>2024-06</creationdate><rights>http://creativecommons.org/licenses/by/4.0</rights><oa>free_for_read</oa><woscitedreferencessubscribed>false</woscitedreferencessubscribed></display><links><openurl>$$Topenurl_article</openurl><openurlfulltext>$$Topenurlfull_article</openurlfulltext><thumbnail>$$Tsyndetics_thumb_exl</thumbnail><link.rule.ids>228,230,780,885</link.rule.ids><linktorsrc>$$Uhttps://arxiv.org/abs/2406.05569$$EView_record_in_Cornell_University$$FView_record_in_$$GCornell_University$$Hfree_for_read</linktorsrc><backlink>$$Uhttps://doi.org/10.48550/arXiv.2406.05569$$DView paper in arXiv$$Hfree_for_read</backlink></links><search><creatorcontrib>Oğuz, Metehan</creatorcontrib><creatorcontrib>Ciftci, Yusuf Umut</creatorcontrib><creatorcontrib>Bakman, Yavuz Faruk</creatorcontrib><title>Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts</title><description>Large language models (LLMs) have shown impressive capabilities in tasks such as machine translation, text summarization, question answering, and solving complex mathematical problems. However, their primary training on data-rich languages like English limits their performance in low-resource languages. This study addresses this gap by focusing on the Indexical Shift problem in Turkish. The Indexical Shift problem involves resolving pronouns in indexical shift contexts, a grammatical challenge not present in high-resource languages like English. We present the first study examining indexical shift in any language, releasing a Turkish dataset specifically designed for this purpose. Our Indexical Shift Dataset consists of 156 multiple-choice questions, each annotated with necessary linguistic details, to evaluate LLMs in a few-shot setting. We evaluate recent multilingual LLMs, including GPT-4, GPT-3.5, Cohere-AYA, Trendyol-LLM, and Turkcell-LLM, using this dataset. Our analysis reveals that even advanced models like GPT-4 struggle with the grammatical nuances of indexical shift in Turkish, achieving only moderate performance. These findings underscore the need for focused research on the grammatical challenges posed by low-resource languages. We released the dataset and code \href{https://anonymous.4open.science/r/indexical_shift_llm-E1B4} {here}.</description><subject>Computer Science - Computation and Language</subject><fulltext>true</fulltext><rsrctype>article</rsrctype><creationdate>2024</creationdate><recordtype>article</recordtype><sourceid>GOX</sourceid><recordid>eNpNkMtOhDAYhdm4MKMP4Mr_AQRbaAt1N8EbCUajGJekLe3QOLSGMgZd-uTOxYWrk3wn5yy-KDrDKCEFpehSjLP9TFKCWIIoZfw4-rn2UNcPAZ618itnvzUM-gLeeu2gAhvA-WlLrmAZgg5h0G4Cbw6TV9fpMUzCddatdrTZjO829FBti9kqsYan0Tu_cQGs-0dfemsmKL2b9DyFk-jIiHXQp3-5iJrbm6a8j-vHu6pc1rFgOY8lUTRnuSJcSkIlYlmeFoaaTEjCME4NEgVXmCLNO46YzLEihcG8E9pwLHi2iM4Pt3sJ7cdoBzF-tTsZ7V5G9gtuKFpT</recordid><startdate>20240608</startdate><enddate>20240608</enddate><creator>Oğuz, Metehan</creator><creator>Ciftci, Yusuf Umut</creator><creator>Bakman, Yavuz Faruk</creator><scope>AKY</scope><scope>GOX</scope></search><sort><creationdate>20240608</creationdate><title>Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts</title><author>Oğuz, Metehan ; Ciftci, Yusuf Umut ; Bakman, Yavuz Faruk</author></sort><facets><frbrtype>5</frbrtype><frbrgroupid>cdi_FETCH-LOGICAL-a679-b4c5767c49bb45b063728f5f3ab46112f0a89c150e9d906b71c48f19daef91a93</frbrgroupid><rsrctype>articles</rsrctype><prefilter>articles</prefilter><language>eng</language><creationdate>2024</creationdate><topic>Computer Science - Computation and Language</topic><toplevel>online_resources</toplevel><creatorcontrib>Oğuz, Metehan</creatorcontrib><creatorcontrib>Ciftci, Yusuf Umut</creatorcontrib><creatorcontrib>Bakman, Yavuz Faruk</creatorcontrib><collection>arXiv Computer Science</collection><collection>arXiv.org</collection></facets><delivery><delcategory>Remote Search Resource</delcategory><fulltext>fulltext_linktorsrc</fulltext></delivery><addata><au>Oğuz, Metehan</au><au>Ciftci, Yusuf Umut</au><au>Bakman, Yavuz Faruk</au><format>journal</format><genre>article</genre><ristype>JOUR</ristype><atitle>Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts</atitle><date>2024-06-08</date><risdate>2024</risdate><abstract>Large language models (LLMs) have shown impressive capabilities in tasks such as machine translation, text summarization, question answering, and solving complex mathematical problems. However, their primary training on data-rich languages like English limits their performance in low-resource languages. This study addresses this gap by focusing on the Indexical Shift problem in Turkish. The Indexical Shift problem involves resolving pronouns in indexical shift contexts, a grammatical challenge not present in high-resource languages like English. We present the first study examining indexical shift in any language, releasing a Turkish dataset specifically designed for this purpose. Our Indexical Shift Dataset consists of 156 multiple-choice questions, each annotated with necessary linguistic details, to evaluate LLMs in a few-shot setting. We evaluate recent multilingual LLMs, including GPT-4, GPT-3.5, Cohere-AYA, Trendyol-LLM, and Turkcell-LLM, using this dataset. Our analysis reveals that even advanced models like GPT-4 struggle with the grammatical nuances of indexical shift in Turkish, achieving only moderate performance. These findings underscore the need for focused research on the grammatical challenges posed by low-resource languages. We released the dataset and code \href{https://anonymous.4open.science/r/indexical_shift_llm-E1B4} {here}.</abstract><doi>10.48550/arxiv.2406.05569</doi><oa>free_for_read</oa></addata></record>
fulltext	fulltext_linktorsrc
identifier	DOI: 10.48550/arxiv.2406.05569
ispartof
issn
language	eng
recordid	cdi_arxiv_primary_2406_05569
source	arXiv.org
subjects	Computer Science - Computation and Language
title	Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
url	https://sfx.bib-bvb.de/sfx_tum?ctx_ver=Z39.88-2004&ctx_enc=info:ofi/enc:UTF-8&ctx_tim=2025-01-15T00%3A01%3A23IST&url_ver=Z39.88-2004&url_ctx_fmt=infofi/fmt:kev:mtx:ctx&rfr_id=info:sid/primo.exlibrisgroup.com:primo3-Article-arxiv_GOX&rft_val_fmt=info:ofi/fmt:kev:mtx:journal&rft.genre=article&rft.atitle=Do%20LLMs%20Recognize%20me,%20When%20I%20is%20not%20me:%20Assessment%20of%20LLMs%20Understanding%20of%20Turkish%20Indexical%20Pronouns%20in%20Indexical%20Shift%20Contexts&rft.au=O%C4%9Fuz,%20Metehan&rft.date=2024-06-08&rft_id=info:doi/10.48550/arxiv.2406.05569&rft_dat=%3Carxiv_GOX%3E2406_05569%3C/arxiv_GOX%3E%3Curl%3E%3C/url%3E&disable_directlink=true&sfx.directlink=off&sfx.report_link=0&rft_id=info:oai/&rft_id=info:pmid/&rfr_iscdi=true