paper-with-me

홈 › Papers

Agentic Design of Compositional Machines

2025-10-16 · Wenqian Zhang, Weiyang Liu, Zhen Liu arxiv

The design of complex machines stands as both a marker of human intelligence and a foundation of engineering practice. Given recent advances in large language models (LLMs), we ask whether they, too, can learn to create. We approach this question through the lens of compositional machine design: a task in which machines are assembled from standardized components to meet functional demands like locomotion or manipulation in a simulated physical environment. With this simplification, machine design is expressed as writing XML-like code that explicitly specifies pairwise part connections. To support this investigation, we introduce BesiegeField, a testbed built on the machine-building game Besiege, which enables part-based construction, physical simulation and reward-driven evaluation. Using BesiegeField, we benchmark state-of-the-art LLMs with agentic workflows and identify key capabilities required for success, including spatial reasoning, strategic assembly, and instruction-following. As current open-source models fall short, we explore reinforcement learning (RL) as a path to improvement: we curate a cold-start dataset, conduct RL finetuning experiments, and highlight open challenges at the intersection of language, machine design, and physical reasoning.

📄 PDF Abstract BibTeX arXiv:2510.14980

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningSpatial Reasoning

Similar Papers 제목 키워드 기반

Ground-Compose-Reinforce: Grounding Language in Agentic Behaviours using Limited Data

2025-07-14 · Andrew C. Li, Toryn Q. Klassen, Andrew Wang, Parand A. Alamdari 외 arxiv

Grounding language in perception and action is a key challenge when building situated agents that can interact with humans, or other agents, via language. In the past, addressing this challenge has required manually desi…

Explain Before You Answer: A Survey on Compositional Visual Reasoning

2025-08-24 · Fucai Ke, Joy Hsu, Zhixi Cai, Zixian Ma 외 arxiv

Compositional visual reasoning has emerged as a key research frontier in multimodal AI, aiming to endow machines with the human-like ability to decompose visual scenes, ground intermediate concepts, and perform multi-ste…

Multimodal ReasoningVisual Reasoning

ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning

2025-10-16 · Roger Creus Castanyer, Faisal Mohamed, Pablo Samuel Castro, Cyrus Neary 외 arxiv

Reinforcement learning (RL) algorithms are highly sensitive to reward function specification, which remains a central challenge limiting their broad applicability. We present ARM-FM: Automated Reward Machines via Foundat…

Zero-shot GeneralizationReinforcement Learning

A Compositional Sheaf-Theoretic Framework for Event-Based Systems (Extended Version)

2020-05-10 · Gioele Zardini, David I. Spivak, Andrea Censi, Emilio Frazzoli

A compositional sheaf-theoretic framework for the modeling of complex event-based systems is presented. We show that event-based systems are machines, with inputs and outputs, and that they can be composed with machines …

WebAggregator: Enhancing Compositional Reasoning Capabilities of Deep Research Agent Foundation Models

2025-10-16 · Rui Wang, Ce Zhang, Jun-Yu Ma, Jianshu Zhang 외 arxiv

The hallmark of Deep Research agents lies in compositional reasoning, the capacity to aggregate distributed, heterogeneous information into coherent logical insights. However, current agentic systems are often retrieval-…