paper-with-me

홈 › Papers

Small Wins Big: Comparing Large Language Models and Domain Fine-Tuned Models for Sarcasm Detection in Code-Mixed Hinglish Text

2026-02-25 · Bitan Majumder, Anirban Sen arxiv

Sarcasm detection in multilingual and code-mixed environments remains a challenging task for natural language processing models due to structural variations, informal expressions, and low-resource linguistic availability. This study compares four large language models, Llama 3.1, Mistral, Gemma 3, and Phi-4, with a fine-tuned DistilBERT model for sarcasm detection in code-mixed Hinglish text. The results indicate that the smaller, sequentially fine-tuned DistilBERT model achieved the highest overall accuracy of 84%, outperforming all of the LLMs in zero and few-shot set ups, using minimal LLM generated code-mixed data used for fine-tuning. These findings indicate that domain-adaptive fine-tuning of smaller transformer based models may significantly improve sarcasm detection over general LLM inference, in low-resource and data scarce settings.

📄 PDF Abstract BibTeX arXiv:2602.21933

Code (0)

등록된 구현이 없습니다.

Tasks

Sarcasm Detection

Similar Papers 제목 키워드 기반

Human Still Wins over LLM: An Empirical Study of Active Learning on Domain-Specific Annotation Tasks

2023-11-16 · Yuxuan Lu, Bingsheng Yao, Shao Zhang, Yun Wang 외

Large Language Models (LLMs) have demonstrated considerable advances, and several claims have been made about their exceeding human performance. However, in real-world tasks, domain knowledge is often required. Low-resou…

Active Learning

SimBench: A Rule-Based Multi-Turn Interaction Benchmark for Evaluating an LLM's Ability to Generate Digital Twins

2024-08-21 · Jingquan Wang, Harry Zhang, Huzaifa Mustafa Unjhawala, Peter Negrut 외

We introduce SimBench, a benchmark designed to evaluate the proficiency of student large language models (S-LLMs) in generating digital twins (DTs) that can be used in simulators for virtual testing. Given a collection o…

Benchmarking

A Diffusion-Model Subpopulation Digital Twin for Mobile Health Deployment: A Case Study on the HeartSteps Intervention

2026-07-23 · Ziping Xu, Yuyi Chang, Chenshun Ni, Nithin Sugavanam 외 arxiv

Mobile-health interventions increasingly use online learning and decision making algorithms to personalize when to nudge users toward healthier behavior, but a poorly designed algorithm can burden and disengage participa…

Decision Making

Large Language Models for Explainable Decisions in Dynamic Digital Twins

2024-05-23 · Nan Zhang, Christian Vergara-Marcillo, Georgios Diamantopoulos, Jingran Shen 외

Dynamic data-driven Digital Twins (DDTs) can enable informed decision-making and provide an optimisation platform for the underlying system. By leveraging principles of Dynamic Data-Driven Applications Systems (DDDAS), D…

Decision Making

Psychometric Comparability of LLM-Based Digital Twins

2025-12-22 · Yufei Zhang, Zhihao Ma arxiv

Large language models (LLMs) act as digital twins for human respondents, yet their psychometric comparability remains uncertain. We propose a construct validity framework spanning construct representation and the nomothe…