paper-with-me

홈 › Papers

Just aware enough: Evaluating awareness across artificial systems

2026-01-21 · Nadine Meertens, Suet Lee, Ophelia Deroy arxiv

Recent debates on artificial intelligence increasingly emphasise questions of AI consciousness and moral status, yet there remains little agreement on how such properties should be evaluated. In this paper, we argue that awareness offers a more productive and methodologically tractable alternative. We introduce a practical method for evaluating awareness across diverse systems, where awareness is understood as encompassing a system's abilities to process, store and use information in the service of goal-directed action. Central to this approach is the claim that any evaluation aiming to capture the diversity of artificial systems must be domain-sensitive, deployable at any scale, multidimensional, and enable the prediction of task performance, while generalising to the level of abilities for the sake of comparison. Given these four desiderata, we outline a structured approach to evaluating and comparing awareness profiles across artificial systems with differing architectures, scales, and operational domains. By shifting the focus from artificial consciousness to being just aware enough, this approach aims to facilitate principled assessment, support design and oversight, and enable more constructive scientific and public discourse.

📄 PDF Abstract BibTeX arXiv:2601.14901

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Fair Enough? A map of the current limitations of the requirements to have fair algorithms

2023-11-21 · Daniele Regoli, Alessandro Castelnovo, Nicole Inverardi, Gabriele Nanino 외

In recent years, the increase in the usage and efficiency of Artificial Intelligence and, more in general, of Automated Decision-Making systems has brought with it an increasing and welcome awareness of the risks associa…

Decision MakingFairness

Evaluating Cultural and Social Awareness of LLM Web Agents

2024-10-30 · Haoyi Qiu, Alexander R. Fabbri, Divyansh Agarwal, Kung-Hsiang Huang 외

As large language models (LLMs) expand into performing as agents for real-world applications beyond traditional NLP tasks, evaluating their robustness becomes increasingly important. However, existing benchmarks often ov…

BenchmarkingNavigate

Evaluating and Improving Cultural Awareness of Reward Models for LLM Alignment

2025-09-26 · Hongbin Zhang, Kehai Chen, Xuefeng Bai, Yang Xiang 외 arxiv

Reward models (RMs) are crucial for aligning large language models (LLMs) with diverse cultures. Consequently, evaluating their cultural awareness is essential for further advancing global alignment of LLMs. However, exi…

Reinforcement Learning

CIAware-Bench: Benchmarking Control Intervention Awareness Across Frontier LLMs

2026-06-09 · Joachim Schaeffer, Thomas Jiralerspong, Alexander Panfilov, Guillaume Lajoie 외 arxiv

AI control protocols oversee untrusted models by monitoring their actions and modifying potentially unsafe steps, often using a trusted model. This partially tampers with the untrusted model's trajectory. If the trusted …

Binary Classification

Are LLMs Aware that Some Questions are not Open-ended?

2024-10-01 · Dongjie Yang, Hai Zhao

Large Language Models (LLMs) have shown the impressive capability of answering questions in a wide range of scenarios. However, when LLMs face different types of questions, it is worth exploring whether LLMs are aware th…

Text Generation