paper-with-me

Papers

Measuring AI Alignment with Human Flourishing

2025-07-10 · Elizabeth Hilliard, Akshaya Jagadeesh, Alex Cook, Steele Billings, Nicholas Skytland, Alicia Llewellyn, Jackson Paull, Nathan Paull, Nolan Kurylo, Keatra Nesbitt, Robert Gruenewald, Anthony Jantzi, Omar Chavez arxiv

This paper introduces the Flourishing AI Benchmark (FAI Benchmark), a novel evaluation framework that assesses AI alignment with human flourishing across seven dimensions: Character and Virtue, Close Social Relationships, Happiness and Life Satisfaction, Meaning and Purpose, Mental and Physical Health, Financial and Material Stability, and Faith and Spirituality. Unlike traditional benchmarks that focus on technical capabilities or harm prevention, the FAI Benchmark measures AI performance on how effectively models contribute to the flourishing of a person across these dimensions. The benchmark evaluates how effectively LLM AI systems align with current research models of holistic human well-being through a comprehensive methodology that incorporates 1,229 objective and subjective questions. Using specialized judge Large Language Models (LLMs) and cross-dimensional evaluation, the FAI Benchmark employs geometric mean scoring to ensure balanced performance across all flourishing dimensions. Initial testing of 28 leading language models reveals that while some models approach holistic alignment (with the highest-scoring models achieving 72/100), none are acceptably aligned across all dimensions, particularly in Faith and Spirituality, Character and Virtue, and Meaning and Purpose. This research establishes a framework for developing AI systems that actively support human flourishing rather than merely avoiding harm, offering significant implications for AI development, ethics, and evaluation.

📄 PDF Abstract BibTeX arXiv:2507.07787

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Positive Alignment: Artificial Intelligence for Human Flourishing

2026-05-11 · Ruben Laukkonen, Seb Krier, Chloé Bakalar, Shamil Chandaria 외 arxiv

Existing alignment research is dominated by concerns about safety and preventing harm: safeguards, controllability, and compliance. This paradigm of alignment parallels early psychology's focus on mental illness: necessa…

Evaluating Artificial Intelligence Through a Christian Understanding of Human Flourishing

2026-04-03 · Nicholas Skytland, Lauren Parsons, Alicia Llewellyn, Steele Billings 외 arxiv

Artificial intelligence (AI) alignment is fundamentally a formation problem, not only a safety problem. As Large Language Models (LLMs) increasingly mediate moral deliberation and spiritual inquiry, they do more than pro…

From Human to Machine Psychology: A Conceptual Framework for Understanding Well-Being in Large Language Model

2025-06-14 · G. R. Lau, W. Y. Low

As large language models (LLMs) increasingly simulate human cognition and behavior, researchers have begun to investigate their psychological properties. Yet, what it means for such models to flourish, a core construct i…

Language ModelingLanguage ModellingLarge Language Model

The Human Flourishing Geographic Index: A County-Level Dataset for the United States, 2013--2023

2025-11-05 · Stefano M. Iacus, Devika Jain, Andrea Nasuto, Giuseppe Porro 외 arxiv

Quantifying human flourishing, a multidimensional construct including happiness, health, purpose, virtue, relationships, and financial stability, is critical for understanding societal well-being beyond economic indicato…

Optimising for Flourishing: Flourishing Metrics and Return on Flourishing as Success Criteria for Artificial Intelligence and Post-AGI Economic Systems

2026-07-31 · Keyun Ruan, Jonathan D. Teubner, John M. Bremen arxiv

Current evaluation frameworks for artificial intelligence focus mainly on capability, safety, and proxies such as adoption, engagement, efficiency, productivity, and financial return. These criteria are necessary but ins…