paper-with-me

Papers

AGITB: A Signal-Level Benchmark for Evaluating Artificial General Intelligence

2025-04-06 · Matej Šprogar

Despite remarkable progress in machine learning, current AI systems continue to fall short of true human-like intelligence. While Large Language Models (LLMs) excel in pattern recognition and response generation, they lack genuine understanding - an essential hallmark of Artificial General Intelligence (AGI). Existing AGI evaluation methods fail to offer a practical, gradual, and informative metric. This paper introduces the Artificial General Intelligence Test Bed (AGITB), comprising twelve rigorous tests that form a signal-processing-level foundation for the potential emergence of cognitive capabilities. AGITB evaluates intelligence through a model's ability to predict binary signals across time without relying on symbolic representations or pretraining. Unlike high-level tests grounded in language or perception, AGITB focuses on core computational invariants reflective of biological intelligence, such as determinism, sensitivity, and generalisation. The test bed assumes no prior bias, operates independently of semantic meaning, and ensures unsolvability through brute force or memorization. While humans pass AGITB by design, no current AI system has met its criteria, making AGITB a compelling benchmark for guiding and recognizing progress toward AGI.

📄 PDF Abstract BibTeX arXiv:2504.04430

Code (1)

matejsprogar/agitb 공식 구현

Tasks

MemorizationResponse Generation

Similar Papers 제목 키워드 기반

ARC-AGI-2: A New Challenge for Frontier AI Reasoning Systems

2025-05-17 · Francois Chollet, Mike Knoop, Gregory Kamradt, Bryan Landers 외

The Abstraction and Reasoning Corpus for Artificial General Intelligence (ARC-AGI), introduced in 2019, established a challenging benchmark for evaluating the general fluid intelligence of artificial systems via a set of…

ARC

Continual learning benefits from multiple sleep mechanisms: NREM, REM, and Synaptic Downscaling

2022-09-09 · Brian S. Robinson, Clare W. Lau, Alexander New, Shane M. Nichols 외

Learning new tasks and skills in succession without losing prior learning (i.e., catastrophic forgetting) is a computational challenge for both artificial and biological neural networks, yet artificial systems struggle t…

Continual Learningimage-classificationImage Classification

Improved Explanatory Efficacy on Human Affect and Workload through Interactive Process in Artificial Intelligence

2019-12-13 · Byung Hyung Kim, Seunghun Koh, Sejoon Huh, Sungho Jo 외

Despite recent advances in the field of explainable artificial intelligence systems, a concrete quantitative measure for evaluating the usability of such systems is nonexistent. Ensuring the success of an explanatory int…

EEGElectroencephalogram (EEG)Explainable artificial intelligenceRecommendation Systems

TS-Skill: A Benchmark for Evaluating Analytical Skills in Time-Series Question Answering

2026-05-23 · Liying Han, Kang Yang, Oliver Wang, Jason Wu 외 arxiv

Large language models (LLMs) and time-series language models (TSLMs) are increasingly applied to time-series question answering (TSQA). Unlike text-only QA, TSQA requires models to ground answers in temporal signals whos…

Question GenerationQuestion Answering

Voice Signal Processing for Machine Learning. The Case of Speaker Isolation

2024-03-29 · Radan Ganchev

The widespread use of automated voice assistants along with other recent technological developments have increased the demand for applications that process audio signals and human voice in particular. Voice recognition t…