paper-with-me

Papers

ValueBench: Towards Comprehensively Evaluating Value Orientations and Understanding of Large Language Models

2024-06-06 · Yuanyi Ren, Haoran Ye, Hanjun Fang, Xin Zhang, Guojie Song

Large Language Models (LLMs) are transforming diverse fields and gaining increasing influence as human proxies. This development underscores the urgent need for evaluating value orientations and understanding of LLMs to ensure their responsible integration into public-facing applications. This work introduces ValueBench, the first comprehensive psychometric benchmark for evaluating value orientations and value understanding in LLMs. ValueBench collects data from 44 established psychometric inventories, encompassing 453 multifaceted value dimensions. We propose an evaluation pipeline grounded in realistic human-AI interactions to probe value orientations, along with novel tasks for evaluating value understanding in an open-ended value space. With extensive experiments conducted on six representative LLMs, we unveil their shared and distinctive value orientations and exhibit their ability to approximate expert conclusions in value-related extraction and generation tasks. ValueBench is openly accessible at https://github.com/Value4AI/ValueBench.

📄 PDF Abstract BibTeX arXiv:2406.04214

Code (1)

value4ai/valuebench 공식 구현 pytorch

Similar Papers 제목 키워드 기반

LocalValueBench: A Collaboratively Built and Extensible Benchmark for Evaluating Localized Value Alignment and Ethical Safety in Large Language Models

2024-07-27 · Gwenyth Isobel Meadows, Nicholas Wai Long Lau, Eva Adelina Susanto, Chi Lok Yu 외

The proliferation of large language models (LLMs) requires robust evaluation of their alignment with local values and ethical standards, especially as existing benchmarks often reflect the cultural, legal, and ideologica…

Prompt Engineering

Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

2026-05-11 · Haonan Dong, Qiguan Feng, Kehan Jiang, Haoran Ye 외 arxiv

Autonomous agents have rapidly matured as task executors and seen widespread deployment via harnesses such as OpenClaw. Safety concerns have rightly drawn growing research attention, and beneath them lie the values silen…

Value Portrait: Understanding Values of LLMs with Human-aligned Benchmark

2025-05-02 · Jongwook Han, Dongmin Choi, Woojung Song, Eun-Ju Lee 외

The importance of benchmarks for assessing the values of language models has been pronounced due to the growing need of more authentic, human-aligned responses. However, existing benchmarks rely on human or machine annot…

Model-Free Assessment of Simulator Fidelity via Quantile Curves

2025-12-04 · Garud Iyengar, Yu-Shiou Willy Lin, Kaizheng Wang arxiv

As generative AI models are increasingly used to simulate real-world systems, quantifying the ``sim-to-real'' gap is critical. For each input setting of interest -- which we call a \emph{scenario}, such as a survey quest…

Human Values Matter: Investigating How Misalignment Shapes Collective Behaviors in LLM Agent Communities

2026-04-07 · Xiangxu Zhang, Jiamin Wang, Qinlin Zhao, Hanze Guo 외 arxiv

As LLMs become increasingly integrated into human society, evaluating their orientations on human values from social science has drawn growing attention. Nevertheless, it is still unclear why human values matter for LLMs…