paper-with-me

홈 › Papers

TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning

2026-02-02 · Hayeong Lee, JunHyeok Oh, Byung-Jun Lee arxiv

The design of environments plays a critical role in shaping the development and evaluation of cooperative multi-agent reinforcement learning (MARL) algorithms. While existing benchmarks highlight critical challenges, they often lack the modularity required to design custom evaluation scenarios. We introduce the Totally Accelerated Battle Simulator in JAX (TABX), a high-throughput sandbox designed for reconfigurable multi-agent tasks. TABX provides granular control over environmental parameters, permitting a systematic investigation into emergent agent behaviors and algorithmic trade-offs across a diverse spectrum of task complexities. Leveraging JAX for hardware-accelerated execution on GPUs, TABX enables massive parallelization and significantly reduces computational overhead. By providing a fast, extensible, and easily customized framework, TABX facilitates the study of MARL agents in complex structured domains and serves as a scalable foundation for future research. Our code is available at: https://github.com/ku-dmlab/TABX.

📄 PDF Abstract BibTeX arXiv:2602.01665

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement Learning

Similar Papers 제목 키워드 기반

fastabx: A library for efficient computation of ABX discriminability

2025-05-05 · Maxime Poli, Emmanuel Chemla, Emmanuel Dupoux

We introduce fastabx, a high-performance Python library for building ABX discrimination tasks. ABX is a measure of the separation between generic categories of interest. It has been used extensively to evaluate phonetic …

Representation Learning

TabXEval: Why this is a Bad Table? An eXhaustive Rubric for Table Evaluation

2025-05-28 · Vihang Pancholi, Jainit Bafna, Tejas Anvekar, Manish Shrivastava 외

Evaluating tables qualitatively & quantitatively presents a significant challenge, as traditional metrics often fail to capture nuanced structural and content discrepancies. To address this, we introduce a novel, methodi…

Specificity

ToolSandbox: A Stateful, Conversational, Interactive Evaluation Benchmark for LLM Tool Use Capabilities

2024-08-08 · Jiarui Lu, Thomas Holleis, Yizhe Zhang, Bernhard Aumayer 외

Recent large language models (LLMs) advancements sparked a growing research interest in tool assisted LLMs solving real-world challenges, which calls for comprehensive evaluation of tool-use capabilities. While previous …

Reinforcement Learning approach for Real Time Strategy Games Battle city and S3

2016-02-16 · Harshit Sethy, Amit Patel

In this paper we proposed reinforcement learning algorithms with the generalized reward function. In our proposed method we use Q-learning and SARSA algorithms with generalised reward function to train the reinforcement …

Q-LearningReal-Time Strategy Gamesreinforcement-learningReinforcement Learning+1

Why Channel-Centric Models are not Enough to Predict End-to-End Performance in Private 5G: A Measurement Campaign and Case Study

2026-03-09 · Nils Jörgensen arxiv

Communication-aware robot planning requires accurate predictions of wireless network performance. Current approaches rely on channel-level metrics such as received signal strength and signal-to-noise ratio, assuming thes…