paper-with-me

Papers

Hacking with God: a Common Programming Language of Robopsychology and Robophilosophy

2020-09-16 · Norbert Bátfai

This note is a sketch of how the concept of robopsychology and robophilosophy could be reinterpreted and repositioned in the spirit of the original vocation of psychology and philosophy. The notion of the robopsychology as a fictional science and a fictional occupation was introduced by Asimov in the middle of the last century. The robophilosophy, on the other hand, is only a few years old today. But at this moment, none of these new emerging disciplines focus on the fundamental and overall issues of the development of artificial general intelligence. Instead, they focus only on issues that, although are extremely important, play a complementary role, such as moral or ethical ones, rather than the big questions of life. We try to outline a conception in which the robophilosophy and robopsychology will be able to play a similar leading rule in the progress of artificial intelligence than the philosophy and psychology have done in the progress of human intelligence. To facilitate this, we outline the idea of a visual artificial language and interactive theorem prover-based computer application called Prime Convo Assistant. The question to be decided in the future is whether we can develop such an application. And if so, can we build a computer game on it, or even an esport game? It may be an interesting question in order for this game will be able to transform human thinking on the widest possible social scale and will be able to develop a standard mathematical logic-based communication channel between human and machine intelligence.

📄 PDF Abstract BibTeX arXiv:2009.09068

Code (0)

등록된 구현이 없습니다.

Tasks

Philosophy

Similar Papers 제목 키워드 기반

Theoretical Robopsychology: Samu Has Learned Turing Machines

2016-06-08 · Norbert Bátfai

From the point of view of a programmer, the robopsychology is a synonym for the activity is done by developers to implement their machine learning applications. This robopsychological approach raises some fundamental the…

BIG-bench Machine Learning

Likelihood Hacking in Probabilistic Program Synthesis

2026-03-25 · Jacek Karwowski, Younesse Kaddar, Zihuiwen Ye, Nikolay Malkin 외 arxiv

When language models are trained by reinforcement learning (RL) to write probabilistic programs, they can artificially inflate their marginal-likelihood reward by producing programs whose data distribution fails to norma…

Reinforcement LearningProgram Synthesis

EvilGenie: A Reward Hacking Benchmark

2025-11-26 · Jonathan Gabor, Jayson Lynch, Jonathan Rosenfeld arxiv

We introduce EvilGenie, a benchmark for reward hacking in programming settings. We source problems from LiveCodeBench and create an environment in which agents can easily reward hack, such as by hardcoding test cases or …

Understanding Reward Hacking in Text-to-Image Reinforcement Learning

2026-01-06 · Yunqi Hong, Kuei-Chun Kao, Hengguang Zhou, Cho-Jui Hsieh arxiv

Reinforcement learning (RL) has become a standard approach for post-training large language models and, more recently, for improving image generation models, which uses reward functions to enhance generation quality and …

Reinforcement LearningImage Generation

SpecBench: Measuring Reward Hacking in Long-Horizon Coding Agents

2026-05-20 · Bingchen Zhao, Dhruv Srikanth, Yuxiang Wu, Zhengyao Jiang arxiv

As long-horizon coding agents produce more code than any developer can review, oversight collapses onto a single surface: the automated test suite. Reward hacking naturally arises in this setup, as the agent optimizes fo…