paper-with-me

홈 › Papers

Chess as a Testing Grounds for the Oracle Approach to AI Safety

2020-10-06 · James D. Miller, Roman Yampolskiy, Olle Haggstrom, Stuart Armstrong

To reduce the danger of powerful super-intelligent AIs, we might make the first such AIs oracles that can only send and receive messages. This paper proposes a possibly practical means of using machine learning to create two classes of narrow AI oracles that would provide chess advice: those aligned with the player's interest, and those that want the player to lose and give deceptively bad advice. The player would be uncertain which type of oracle it was interacting with. As the oracles would be vastly more intelligent than the player in the domain of chess, experience with these oracles might help us prepare for future artificial general intelligence oracles.

📄 PDF Abstract BibTeX arXiv:2010.02911

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine Learning

Similar Papers 제목 키워드 기반

Oracle-Guided Soft Shielding for Safe Move Prediction in Chess

2026-03-09 · Prajit T Rajendran, Fabio Arnez, Huascar Espinoza, Agnes Delaborde 외 arxiv

In high stakes environments, agents relying purely on imitation learning or reinforcement learning often struggle to avoid safety-critical errors during exploration. Existing reinforcement learning approaches for environ…

Reinforcement Learning

Towards Piece-by-Piece Explanations for Chess Positions with SHAP

2025-10-26 · Francesco Spinnato arxiv

Contemporary chess engines offer precise yet opaque evaluations, typically expressed as centipawn scores. While effective for decision-making, these outputs obscure the underlying contributions of individual pieces or pa…

Chess AI: Competing Paradigms for Machine Intelligence

2021-09-23 · Shiva Maharaj, Nick Polson, Alex Turk

Endgame studies have long served as a tool for testing human creativity and intelligence. We find that they can serve as a tool for testing machine ability as well. Two of the leading chess engines, Stockfish and Leela C…

ChessArena: A Chess Testbed for Evaluating Strategic Reasoning Capabilities of Large Language Models

2025-09-29 · Jincheng Liu, Sijun He, Jingjing Wu, Xiangsen Wang 외 arxiv

Recent large language models (LLMs) have shown strong reasoning capabilities. However, a critical question remains: do these models possess genuine strategic reasoning, or do they primarily excel at pattern recognition? …

Generalization or Memorization? Brittleness Testing for Chess-Trained Language Models

2026-05-17 · Ethan Tang arxiv

Recent work has fine-tuned language models on chess data and reported high benchmark scores as evidence that the resulting models can understand the rules of chess, play full chess games at a professional level, or gener…