paper-with-me

Papers

Grounding LLMs in Scientific Discovery via Embodied Actions

2026-02-24 · Bo Zhang, Jinfeng Zhou, Yuxuan Chen, Jianing Yin, Minlie Huang, Hongning Wang arxiv

Large Language Models (LLMs) have shown significant potential in scientific discovery but struggle to bridge the gap between theoretical reasoning and verifiable physical simulation. Existing solutions operate in a passive "execute-then-response" loop and thus lacks runtime perception, obscuring agents to transient anomalies (e.g., numerical instability or diverging oscillations). To address this limitation, we propose EmbodiedAct, a framework that transforms established scientific software into active embodied agents by grounding LLMs in embodied actions with a tight perception-execution loop. We instantiate EmbodiedAct within MATLAB and evaluate it on complex engineering design and scientific modeling tasks. Extensive experiments show that EmbodiedAct significantly outperforms existing baselines, achieving SOTA performance by ensuring satisfactory reliability and stability in long-horizon simulations and enhanced accuracy in scientific modeling.

📄 PDF Abstract BibTeX arXiv:2602.20639

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Embodied Science: Closing the Discovery Loop with Agentic Embodied AI

2026-03-20 · Xiang Zhuang, Chenyi Zhou, Kehua Feng, Zhihui Zhu 외 arxiv

Artificial intelligence has demonstrated remarkable capability in predicting scientific properties, yet scientific discovery remains an inherently physical, long-horizon pursuit governed by experimental cycles. Most curr…

Auto-Bench: An Automated Benchmark for Scientific Discovery in LLMs

2025-02-21 · Tingting Chen, Srinivas Anumasa, Beibei Lin, Vedant Shah 외

Given the remarkable performance of Large Language Models (LLMs), an important question arises: Can LLMs conduct human-like scientific research and discover new knowledge, and act as an AI scientist? Scientific discovery…

scientific discoveryvalid

Intelligence Requires Grounding But Not Embodiment

2026-01-24 · Marcus Ma, Shrikanth Narayanan arxiv

Recent advances in LLMs have reignited scientific debate over whether embodiment is necessary for intelligence. We present the argument that intelligence requires grounding, a phenomenon entailed by embodiment, but not e…

LLM and Simulation as Bilevel Optimizers: A New Paradigm to Advance Physical Scientific Discovery

2024-05-16 · Pingchuan Ma, Tsun-Hsuan Wang, Minghao Guo, Zhiqing Sun 외

Large Language Models have recently gained significant attention in scientific discovery for their extensive knowledge and advanced reasoning capabilities. However, they encounter challenges in effectively simulating obs…

Bilevel Optimizationscientific discovery

ExpVid: A Benchmark for Experiment Video Understanding & Reasoning

2025-10-13 · Yicheng Xu, Yue Wu, Jiashuo Yu, Ziang Yan 외 arxiv

Multimodal Large Language Models (MLLMs) hold promise for accelerating scientific discovery by interpreting complex experimental procedures. However, their true capabilities are poorly understood, as existing benchmarks …

Visual Grounding