paper-with-me

홈 › Papers

Keke AI Competition: Solving puzzle levels in a dynamically changing mechanic space

2022-09-11 · M Charity, Julian Togelius

The Keke AI Competition introduces an artificial agent competition for the game Baba is You - a Sokoban-like puzzle game where players can create rules that influence the mechanics of the game. Altering a rule can cause temporary or permanent effects for the rest of the level that could be part of the solution space. The nature of these dynamic rules and the deterministic aspect of the game creates a challenge for AI to adapt to a variety of mechanic combinations in order to solve a level. This paper describes the framework and evaluation metrics used to rank submitted agents and baseline results from sample tree search agents.

📄 PDF Abstract BibTeX arXiv:2209.04911

Code (0)

등록된 구현이 없습니다.

Tasks

Sokoban

Similar Papers 제목 키워드 기반

From Rosetta to Match-Up: A Paired Corpus of Linguistic Puzzles with Human and LLM Benchmarks

2026-05-13 · Neh Majmudar, Anne Huang, Jinfan Frank Hu, Elena Filatova arxiv

In this paper, we examine linguistic puzzles used in high school linguistics competitions, focusing on two common formats: Rosetta Stone and Match-Up. We propose a systematic procedure for converting existing Rosetta Sto…

From Frustration to Fun: An Adaptive Problem-Solving Puzzle Game Powered by Genetic Algorithm

2025-09-28 · Matthew McConnell, Richard Zhao arxiv

This paper explores adaptive problem solving with a game designed to support the development of problem-solving skills. Using an adaptive, AI-powered puzzle game, our adaptive problem-solving system dynamically generates…

The 2017 AIBIRDS Competition

2018-03-14 · Matthew Stephenson, Jochen Renz, Xiaoyu Ge, Peng Zhang

This paper presents an overview of the sixth AIBIRDS competition, held at the 26th International Joint Conference on Artificial Intelligence. This competition tasked participants with developing an intelligent agent whic…

Deep Reinforcement LearningReinforcement Learning

EnigmaEval: A Benchmark of Long Multimodal Reasoning Challenges

2025-02-13 · Clinton J. Wang, Dean Lee, Cristina Menghini, Johannes Mols 외

As language models master existing reasoning benchmarks, we need new challenges to evaluate their cognitive frontiers. Puzzle-solving events are rich repositories of challenging multimodal problems that test a wide range…

Humanity's Last ExamMultimodal Reasoning

VGRP-Bench: Visual Grid Reasoning Puzzle Benchmark for Large Vision-Language Models

2025-03-29 · Yufan Ren, Konstantinos Tertikas, Shalini Maiti, Junlin Han 외

Large Vision-Language Models (LVLMs) struggle with puzzles, which require precise perception, rule comprehension, and logical reasoning. Assessing and enhancing their performance in this domain is crucial, as it reflects…

Logical Reasoning