paper-with-me

홈 › Papers

CS-Sum: A Benchmark for Code-Switching Dialogue Summarization and the Limits of Large Language Models

2025-05-19 · Sathya Krishnan Suresh, Tanmay Surana, Lim Zhi Hao, Eng Siong Chng

Code-switching (CS) poses a significant challenge for Large Language Models (LLMs), yet its comprehensibility remains underexplored in LLMs. We introduce CS-Sum, to evaluate the comprehensibility of CS by the LLMs through CS dialogue to English summarization. CS-Sum is the first benchmark for CS dialogue summarization across Mandarin-English (EN-ZH), Tamil-English (EN-TA), and Malay-English (EN-MS), with 900-1300 human-annotated dialogues per language pair. Evaluating ten LLMs, including open and closed-source models, we analyze performance across few-shot, translate-summarize, and fine-tuning (LoRA, QLoRA on synthetic data) approaches. Our findings show that though the scores on automated metrics are high, LLMs make subtle mistakes that alter the complete meaning of the dialogue. To this end, we introduce 3 most common type of errors that LLMs make when handling CS input. Error rates vary across CS pairs and LLMs, with some LLMs showing more frequent errors on certain language pairs, underscoring the need for specialized training on code-switched data.

📄 PDF Abstract BibTeX arXiv:2505.13559

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PingPong: A Natural Benchmark for Multi-Turn Code-Switching Dialogues

2026-01-24 · Mohammad Rifqi Farhansyah, Hanif Muhammad Zhafran, Farid Adilazuarda, Shamsuddeen Hassan Muhammad 외 arxiv

Code-switching is a widespread practice among the world's multilingual majority, yet few benchmarks accurately reflect its complexity in everyday communication. We present PingPong, a benchmark for natural multi-party co…

Question Answering

Enhancing Semantic Understanding with Self-supervised Methods for Abstractive Dialogue Summarization

2022-09-01 · Hyunjae Lee, Jaewoong Yun, Hyunjin Choi, Seongho Joe 외

Contextualized word embeddings can lead to state-of-the-art performances in natural language understanding. Recently, a pre-trained deep contextualized text encoder such as BERT has shown its potential in improving natur…

Abstractive Dialogue SummarizationAbstractive Text SummarizationArticlesDecoder+5

CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition

2025-02-26 · Jiaming Zhou, Yujie Guo, Shiwan Zhao, Haoqin Sun 외

Code-switching (CS), the alternation between two or more languages within a single conversation, presents significant challenges for automatic speech recognition (ASR) systems. Existing Mandarin-English code-switching da…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

An Exploratory Study on Long Dialogue Summarization: What Works and What's Next

2021-09-10 · Yusen Zhang, Ansong Ni, Tao Yu, Rui Zhang 외

Dialogue summarization helps readers capture salient information from long conversations in meetings, interviews, and TV series. However, real-world dialogues pose a great challenge to current summarization models, as th…

ArticlesRetrieval

An Exploratory Study on Long Dialogue Summarization: What Works and What’s Next

2021-11-01 · Findings (EMNLP) 2021 11 · Yusen Zhang, Ansong Ni, Tao Yu, Rui Zhang 외

Dialogue summarization helps readers capture salient information from long conversations in meetings, interviews, and TV series. However, real-world dialogues pose a great challenge to current summarization models, as th…

ArticlesRetrieval