paper-with-me

홈 › Papers

Static Sandboxes Are Inadequate: Modeling Societal Complexity Requires Open-Ended Co-Evolution in LLM-Based Multi-Agent Simulations

2025-10-15 · Jinkun Chen, Sher Badshah, Xuemin Yu, Sijia Han arxiv

What if artificial agents could not just communicate, but also evolve, adapt, and reshape their worlds in ways we cannot fully predict? With llm now powering multi-agent systems and social simulations, we are witnessing new possibilities for modeling open-ended, ever-changing environments. Yet, most current simulations remain constrained within static sandboxes, characterized by predefined tasks, limited dynamics, and rigid evaluation criteria. These limitations prevent them from capturing the complexity of real-world societies. In this paper, we argue that static, task-specific benchmarks are fundamentally inadequate and must be rethought. We critically review emerging architectures that blend llm with multi-agent dynamics, highlight key hurdles such as balancing stability and diversity, evaluating unexpected behaviors, and scaling to greater complexity, and introduce a fresh taxonomy for this rapidly evolving field. Finally, we present a research roadmap centered on open-endedness, continuous co-evolution, and the development of resilient, socially aligned AI ecosystems. We call on the community to move beyond static paradigms and help shape the next generation of adaptive, socially-aware multi-agent simulations.

📄 PDF Abstract BibTeX arXiv:2510.13982

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Operationalising AI Regulatory Sandboxes under the EU AI Act: The Triple Challenge of Capacity, Coordination and Attractiveness to Providers

2025-09-07 · Deirdre Ahern arxiv

The EU AI Act provides a rulebook for all AI systems being put on the market or into service in the European Union. This article investigates the requirement under the AI Act that Member States establish national AI regu…

Quantifying Frontier LLM Capabilities for Container Sandbox Escape

2026-03-01 · Rahul Marchand, Art O Cathain, Jerome Wynne, Philippos Maximos Giavridis 외 arxiv

Large language models (LLMs) increasingly act as autonomous agents, using tools to execute code, read and write files, and access networks, creating novel security risks. To mitigate these risks, agents are commonly depl…

Science sandboxes measure the scientific capability of AI agents

2026-08-31 · Arya S. Rao, Rodrigo I. Castro, Sager J. Gosai, Kenneth B. Hsu 외 arxiv

Scientific progress depends not only on finding solutions, but on learning the rules that explain why they work and using that understanding to design better experiments. We introduce science sandboxes, a framework for s…

Dynamic Frequency-Based Fingerprinting Attacks against Modern Sandbox Environments

2024-04-16 · Debopriya Roy Dipta, Thore Tiemann, Berk Gulmezoglu, Eduard Marin 외

The cloud computing landscape has evolved significantly in recent years, embracing various sandboxes to meet the diverse demands of modern cloud applications. These sandboxes encompass container-based technologies like D…

Cloud ComputingCPU

EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis

2026-01-09 · Xiaoshuai Song, Haofei Chang, Guanting Dong, Yutao Zhu 외 arxiv

Large language models (LLMs) are expected to be trained to act as agents in various real-world environments, but this process relies on rich and varied tool-interaction sandboxes. However, access to real systems is often…

Reinforcement Learning