arrow
返回

Large Language Models as Kuwaiti Annotators

delete2025-02-08
delete0
delete
OA
AI
H
Hana Alostad *
DOI:10.3390/bdcc9020033delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Stance detection for low-resource languages, such as the Kuwaiti dialect, poses a significant challenge in natural language processing (NLP) due to the scarcity of annotated datasets and specialized tools. This study addresses these limitations by evaluating the effectiveness of open large language models (LLMs) in automating stance detection through zero-shot and few-shot prompt engineering, with a focus on the potential of open-source models to achieve performance levels comparable to those of closed-source alternatives. We also highlight the critical distinctions between zero- and few-shot learning, emphasizing their significance for addressing the challenges posed by low-resource languages. Our evaluation involved testing 11 LLMs on a manually labeled dataset of social media posts, including GPT-4o, Gemini Pro 1.5, Mistral-Large, Jais-30B, and AYA-23. As expected, closed-source models such as GPT-4o, Gemini Pro 1.5, and Mistral-Large demonstrated superior performance, achieving maximum F1 scores of 95.4%, 95.0%, and 93.2%, respectively, in few-shot scenarios with English as the prompt template language. However, open-source models such as Jais-30B and AYA-23 achieved competitive results, with maximum F1 scores of 93.0% and 93.1%, respectively, under the same conditions. Furthermore, statistical analysis using ANOVA and Tukey's HSD post hoc tests revealed no significant differences in overall performance among GPT-4o, Gemini Pro 1.5, Mistral-Large, Jais-30B, and AYA-23. This finding underscores the potential of open-source LLMs as cost-effective and privacy-preserving alternatives for low-resource language annotation. This is the first study comparing LLMs for stance detection in the Kuwaiti dialect. Our findings highlight the importance of prompt design and model consistency in improving the quality of annotations and pave the way for NLP solutions for under-represented Arabic dialects.
Keyword:
data annotation
large language models
zero-shot
few-shot
Arabic
Kuwaiti dialect

期刊

B
Big Data and Cognitive Computing
IF:
4.4
论文数:
1.3K
被引数:
2.4K

机构

暂无机构信息
引用论文

引用论文

err分享
err收藏
Social media and attitudes towards a COVID-19 vaccination: A systematic review of the literature
err2022-06-01
err149
errOAAI
errCascini, Fidelia; Pantovic, Ana; Al-Ajlouni, Yazan A.; Failla, Giovanna; Puleo, Valeria; Melnyk, Andriy; Lontano, Alberto; Ricciardi, Walter
err分享
err收藏
err分享
err收藏
err
IF0
err
err0
PREAI
err
err分享
err收藏
The Use of Wood Chips for Revitalization of Degraded Forest Soil on Young Scots Pine Plantation
err2020-06-17
err0
errOAAI
errAndrzej Klimek; Stanisław Rolbiecki; Roman Rolbiecki; Grzegorz Gackowski; Piotr Stachowski; Barbara Jagosz
err分享
err收藏
59. Variability in response to 1 Hz repetitive TMS
err2016-04-01
err0
PREAI
errG. Strigaro; M. Hamada; R. Cantello; J.C. Rothwell
err分享
err收藏
学者 查看更多内容