paper-with-me

홈 › Papers

Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works

2024-04-19 · Xinfeng Yuan, Siyu Yuan, Yuhan Cui, Tianhe Lin, Xintao Wang, Rui Xu, Jiangjie Chen, Deqing Yang

Large language models (LLMs) have demonstrated impressive performance and spurred numerous AI applications, in which role-playing agents (RPAs) are particularly popular, especially for fictional characters. The prerequisite for these RPAs lies in the capability of LLMs to understand characters from fictional works. Previous efforts have evaluated this capability via basic classification tasks or characteristic imitation, failing to capture the nuanced character understanding with LLMs. In this paper, we propose evaluating LLMs' character understanding capability via the character profiling task, i.e., summarizing character profiles from corresponding materials, a widely adopted yet understudied practice for RPA development. Specifically, we construct the CroSS dataset from literature experts and assess the generated profiles by comparing them with ground truth references and evaluating their applicability in downstream tasks. Our experiments, which cover various summarization methods and LLMs, have yielded promising results. These results strongly validate the character understanding capability of LLMs. Resources are available at https://github.com/Joanna0123/character_profiling.

📄 PDF Abstract BibTeX arXiv:2404.12726

Code (1)

joanna0123/character_profiling 공식 구현

Similar Papers 제목 키워드 기반

Beyond Single Character: Evaluating MLLMs for Sentence-Level Oracle Bone Inscription Understanding

2026-06-30 · Ziqi Li, Zijian Chen, Tingzhu Chen, Guangtao Zhai arxiv

Existing AI-assisted oracle bone inscription (OBI) visual recognition and understanding studies mainly focus on character-level, ignoring the long-form textual coherence and contextual dependencies embedded in complete d…

PACUTE: Phonology-, Affix-, and Character-level Understanding of Tokens for Filipino

2026-06-13 · Jann Railey Montalan, David Demitri Africa, Jimson Paulo Layacan, Richell Isaiah Flores 외 arxiv

Large language models (LLMs) process text as sequences of subword tokens, which can obscure the character-level and morphological structure that underlies word formation. This limitation is most acute for languages with …

The Impact of Visual Information in Chinese Characters: Evaluating Large Models' Ability to Recognize and Utilize Radicals

2024-10-11 · Xiaofeng Wu, Karl Stratos, Wei Xu

The glyphic writing system of Chinese incorporates information-rich visual features in each character, such as radicals that provide hints about meaning or pronunciation. However, there has been no investigation into whe…

Part-Of-Speech Tagging

YOMI-Bench: A Benchmark for Evaluating Kanji Reading and Phonological Understanding of LLMs for Japanese

2026-07-01 · Ryota Mibayashi, Hiroya Takamura, Hitomi Yanaka arxiv

We propose YOMI-Bench, a benchmark for evaluating kanji reading and phonological understanding of large language models (LLMs) for Japanese. In Japanese, a single kanji character often has multiple possible readings, mak…

TurkBench: A Benchmark for Evaluating Turkish Large Language Models

2026-01-11 · Çağrı Toraman, Ahmet Kaan Sever, Ayse Aysu Cengiz, Elif Ecem Arslan 외 arxiv

With the recent surge in the development of large language models, the need for comprehensive and language-specific evaluation benchmarks has become critical. While significant progress has been made in evaluating Englis…

Instruction Following