paper-with-me

홈 › Papers

An Auditable AI Agent Loop for Empirical Economics: A Case Study in Forecast Combination

2026-03-18 · Minchul Shin arxiv

AI coding agents, general purpose assistants that write and execute code, make empirical specification search fast and cheap, but they also widen hidden researcher degrees of freedom. This paper adapts an open-source agent-loop architecture to an empirical economics workflow and adds a post-search holdout evaluation. In a forecast-combination illustration, independent agent searches find methods that improve on benchmarks from the original study. Logged search and holdout evaluation together make adaptive specification search more transparent and help distinguish robust improvements from sample-specific discoveries.

📄 PDF Abstract BibTeX arXiv:2603.17381

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

HLER: Human-in-the-Loop Economic Research via Multi-Agent Pipelines for Empirical Discovery

2026-03-08 · Chen Zhu, Xiaolu Wang arxiv

Large language models (LLMs) have enabled agent-based systems that aim to automate scientific research workflows. Most existing approaches focus on fully autonomous discovery, where AI systems generate research ideas, co…

Automated Interpretability and Feature Discovery in Language Models with Agents

2026-05-02 · Arnau Marin-Llobet, Javier Ferrando arxiv

We introduce an autonomous multiagent framework for mechanistic interpretability that automates both explaining and finding internal features in large language models. The system runs two coupled loops: (1) explanation r…

Dr. Claw: An AI Scientist Workspace for Vibe Research

2026-08-31 · Dingjie Song, Hanrong Zhang, Dawei Liu, Yixin Liu 외 hf

Command-line coding agents (e.g., Claude Code, Gemini CLI) can already read and write files and sustain long sessions, yet end-to-end research still fragments across chat tools, IDEs, terminals, and writing environments,…

Regimes: An Auditable, Held-Out-Gated Improvement Loop Demonstrated on LongMemEval with ActiveGraph

2026-06-08 · Yohei Nakajima arxiv

Autonomous improvement loops are hard to trust because the improvement process is usually external scaffolding bolted onto the agent: failures go unlogged, diagnoses cannot be replayed, and promote-or-discard decisions l…

An Auditable Agent Platform For Automated Molecular Optimisation

2025-08-05 · Atabey Ünlü, Phil Rohr, Ahmet Celebi arxiv

Drug discovery frequently loses momentum when data, expertise, and tools are scattered, slowing design cycles. To shorten this loop we built a hierarchical, tool using agent framework that automates molecular optimisatio…

Drug Discovery