paper-with-me

홈 › Papers

Validity of LLMs as data annotators: AMALIA on authority

2026-07-09 · Manuel Pita arxiv

A national language model offers a linguistic community its own instrument for measuring what its citizens say and value. Portugal's AMALIA, a publicly funded 9B-parameter model for European Portuguese, appears competitive on agreement alone: asked to code the moral foundation of authority, it agrees with trained human coders to within six F1 points of open models eight to thirteen times its size. Yet agreement is reliability, not validity. For theoretical constructs that must be inferred rather than read from surface features, the question is whether the model follows the construct's theory or reaches the right code by correlated shortcuts. We test this with the recovery gap: the loss in performance when a holistic prompt is decomposed into the codebook's atomic clauses and recombined by the theory's explicit rule. If calibration closes that gap, some portability should survive across models and languages; where it does not, the construct-model instrument is the likely locus of failure. We ask whether a calibrated English instrument transfers to AMALIA-9B and to European Portuguese. For one construct and one corpus, it does not. Decomposition recovers only about half of AMALIA's holistic performance, and error analysis suggests reliance on surface correlates, especially moral outrage near authority figures. An open multilingual LLM closes the gap on the same Portuguese corpus under the same instructions, pointing away from the corpus as the main explanation. AMALIA can still screen and pre-code at scale, but it cannot yet measure this construct well enough to stand alone. The study is a single counterexample, not a verdict on national models; it argues that sovereign-LLM benchmark batteries should test not only agreement with human coders, but the evidential route by which that agreement is warranted.

📄 PDF Abstract BibTeX arXiv:2607.08731

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

AMALIA Technical Report: A Fully Open Source Large Language Model for European Portuguese

2026-03-27 · Afonso Simplício, Gonçalo Vinagre, Miguel Moura Ramos, Diogo Tavares 외 arxiv

Despite rapid progress in open large language models (LLMs), European Portuguese (pt-PT) remains underrepresented in both training data and native evaluation, with machine-translated benchmarks likely missing the variant…

Three Models of RLHF Annotation: Extension, Evidence, and Authority

2026-04-28 · Steve Coyne arxiv

Preference-based alignment methods, most prominently Reinforcement Learning with Human Feedback (RLHF), use the judgments of human annotators to shape large language model behaviour. However, the normative role of these …

Reinforcement Learning

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model

2026-06-17 · Diogo Glória-Silva, João Cardeira, Manuel Letras da Luz, Afonso Simplício 외 arxiv

Large Vision and Language Models (LVLMs) have advanced rapidly, yet European Portuguese (pt-PT) remains systematically underserved by existing open-source multimodal models, which either conflate it with Brazilian Portug…

LLMs Prompted for Legal Context Object More: Overrefusal from Small On-Premises LLMs in Criminal Legal Context

2026-06-23 · Anastasiia Kucherenko, François Brouchoud, Dimitri Percia David, Andrei Kucharavy arxiv

While the validity of LLMs' use in the legal context remains subject to ethical and legal debate, legal professionals are already experimenting with personal LLMs, if only for translation and reformulation. However, even…

Evaluating Alignment of Behavioral Dispositions in LLMs

2026-02-11 · Amir Taubenfeld, Zorik Gekhman, Lior Nezry, Omri Feldman 외 arxiv

As LLMs integrate into our daily lives, understanding their behavior becomes essential. In this work, we focus on behavioral dispositions$-$the underlying tendencies that shape responses in social contexts$-$and introduc…