paper-with-me

홈 › Papers

The 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems

2026-02-19 · Leon Staufer, Kevin Feng, Kevin Wei, Luke Bailey, Yawen Duan, Mick Yang, A. Pinar Ozisik, Stephen Casper, Noam Kolt arxiv

Agentic AI systems are increasingly capable of performing professional and personal tasks with limited human involvement. However, tracking these developments is difficult because the AI agent ecosystem is complex, rapidly evolving, and inconsistently documented, posing obstacles to both researchers and policymakers. To address these challenges, this paper presents the 2025 AI Agent Index. The Index documents information regarding the origins, design, capabilities, ecosystem, and safety features of 30 state-of-the-art AI agents based on publicly available information and email correspondence with developers. In addition to documenting information about individual agents, the Index illuminates broader trends in the development of agents, their capabilities, and the level of transparency of developers. Notably, we find different transparency levels among agent developers and observe that most developers share little information about safety, evaluations, and societal impacts. The 2025 AI Agent Index is available online at https://aiagentindex.mit.edu

📄 PDF Abstract BibTeX arXiv:2602.17753

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The AI Agent Index

2025-02-03 · Stephen Casper, Luke Bailey, Rosco Hunter, Carson Ezell 외

Leading AI developers and startups are increasingly deploying agentic AI systems that can plan and execute complex tasks with limited human involvement. However, there is currently no structured framework for documenting…

AI AgentManagement

The Doctor Will (Still) See You Now: On the Structural Limits of Agentic AI in Healthcare

2026-02-06 · Gabriela Aránguiz Dias, Kiana Jafari, Allie Griffith, Carolina Aránguiz Dias 외 arxiv

Across healthcare, agentic artificial intelligence (AI) systems are increasingly promoted as capable of autonomous action, yet in practice they currently operate under near-total human oversight due to safety, regulatory…

AI Transparency Atlas: Framework, Scoring, and Real-Time Model Card Evaluation Pipeline

2025-12-13 · Akhmadillo Mamirov, Faiaz Azmain, Hanyu Wang arxiv

AI model documentation is fragmented across platforms and inconsistent in structure, preventing policymakers, auditors, and users from reliably assessing safety claims, data provenance, and version-level changes. We anal…

TechOps: Technical Documentation Templates for the AI Act

2025-08-12 · Laura Lucaj, Alex Loosley, Hakan Jonsson, Urs Gasser 외 arxiv

Operationalizing the EU AI Act requires clear technical documentation to ensure AI systems are transparent, traceable, and accountable. Existing documentation templates for AI systems do not fully cover the entire AI lif…

MI9: An Integrated Runtime Governance Framework for Agentic AI

2025-08-05 · Charles L. Wang, Trisha Singhal, Ameya Kelkar, Jason Tuo arxiv

Agentic AI systems capable of reasoning, planning, and executing actions present fundamentally distinct governance challenges compared to traditional AI models. Unlike conventional AI, these systems exhibit emergent and …