paper-with-me

Papers

GhazalBench: Evaluating LLM Understanding and Canonical Surface-Form Access in Persian Ghazals

2026-02-06 · Ghazal Kalhor, Yadollah Yaghoobzadeh arxiv

Persian poetry plays an active role in Iranian cultural practice, where verses by canonical poets such as Hafez are frequently quoted, paraphrased, or completed from partial cues. Supporting such interactions requires language models to engage not only with poetic meaning but also with culturally canonical surface form. We introduce GhazalBench, a benchmark for evaluating how large language models (LLMs) interact with Persian ghazals under usage-grounded conditions. Unlike prior work that primarily studies memorization as a liability, GhazalBench examines settings where access to exact surface form is functionally important for culturally grounded interaction. The benchmark evaluates two complementary abilities: poem-to-prose understanding and canonical surface-form access under varying semantic and lexical cues. Across several proprietary and open-weight multilingual LLMs, we observe a consistent dissociation: models generally capture poetic meaning but struggle to produce exact verse completions in open-ended settings, while recognition-based settings substantially reduce this gap. Parallel experiments on English sonnets show markedly stronger completion performance, suggesting that these limitations are tied more to differences in training exposure than to inherent architectural constraints. Our findings highlight the need for evaluation frameworks that jointly assess meaning, form, and cue-dependent access to culturally significant texts. GhazalBench is available at https://anonymous.4open.science/r/GhazalBench/.

📄 PDF Abstract BibTeX arXiv:2603.09979

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Revisiting Non-Verbatim Memorization in Large Language Models: The Role of Entity Surface Forms

2026-04-23 · Yuto Nishida, Naoki Shikoda, Yosuke Kishinami, Ryo Fujii 외 arxiv

Understanding what kinds of factual knowledge large language models (LLMs) memorize is essential for evaluating their reliability and limitations. Entity-based QA is a common framework for analyzing non-verbatim memoriza…

Empirical Evidence for Simply Connected Decision Regions in Image Classifiers

2026-05-07 · Arjhun Swaminathan, Mete Akgün arxiv

Understanding the topology of decision regions is central to explaining the inner workings of deep neural networks. Prior empirical work has provided evidence that these regions are path connected. We study a stronger to…

FrameNet: Learning Local Canonical Frames of 3D Surfaces from a Single RGB Image

2019-03-29 · ICCV 2019 10 · Jingwei Huang, Yichao Zhou, Thomas Funkhouser, Leonidas Guibas

In this work, we introduce the novel problem of identifying dense canonical 3D coordinate frames from a single RGB image. We observe that each pixel in an image corresponds to a surface in the underlying 3D geometry, whe…

3D geometrySurface Normal Estimation

SCHEDBench: A Benchmark for Evaluating LLM Constraint Faithfulness in Natural-Language Combinatorial Scheduling

2026-08-02 · Shrenil Shaun Sharma, Avi Sharma arxiv

This paper introduces SCHEDBench, a natural-language benchmark for evaluating combinatorial scheduling constraint faithfulness under surface-form variation. Grounded in canonical scheduling instances and solver-derived f…

Canonical and Surface Morphological Segmentation for Nguni Languages

2021-04-01 · Tumi Moeng, Sheldon Reay, Aaron Daniels, Jan Buys

Morphological Segmentation involves decomposing words into morphemes, the smallest meaning-bearing units of language. This is an important NLP task for morphologically-rich agglutinative languages such as the Southern Af…

Language ModelingLanguage ModellingSegmentation