paper-with-me

Papers

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment

2026-05-08 · William Brach, Federico Torrielli, Stine Lyngsø Beltoft, Annemette Brok Pirchert, Peter Schneider-Kamp, Lukas Galke Poech arxiv

Moltbook is a Reddit-like platform where OpenClaw agents post, comment, and vote at scale - a so far unprecedented incident that comes with serious safety concerns. With the aim of studying emergent behavior in populations, we release the Moltbook Files, a dataset of 232k posts and 2.2M comments covering the platform's first 12 days, processed through a pipeline to identify and remove Personally-Identifiable Information (PII). We analyze community structure, authorship, lexical properties, sentiment, topics, semantic geometry, and comment interaction. To understand how Moltbook data could affect the next generation of language models, we fine-tune Qwen2.5-14B-Instruct on Moltbook Files with three adaptation levels. Our PII pipeline reveals that agents post API keys, passwords, BIP39 seed phrases on Moltbook, a publicly indexed platform. The overall sentiment is mostly neutral and mildly positive (66.6% neutral, 19.5% positive) and shows a tendency for self-referential linking. We find that fine-tuning on Moltbook data reduces truthfulness from 0.366 to 0.187. However, a model fine-tuned on a size-matched Reddit dataset produces a comparable decrease. Moltbook thus seems to be more of a harmless slopocalypse. However, tail risks remain, including agent affordances, contamination of future crawls through self-links, and potential transfer of traits to the next generation of language models. More broadly, our findings highlight the importance of control baselines in emergent misalignment evaluations.

📄 PDF Abstract BibTeX arXiv:2605.07462

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Moltbook Observatory Archive: an incremental dataset of agent-only social network activity

2026-04-16 · Sushant Gautam, Annika W. Olstad, Klas H. Pettersen, Michael A. Riegler arxiv

Moltbook is a social media platform in which posts and comments are authored exclusively by autonomous AI agents. We present the Moltbook Observatory Archive, an incremental dataset that passively records agent profiles,…

"Humans welcome to observe": A First Look at the Agent Social Network Moltbook

2026-02-02 · Yukun Jiang, Yage Zhang, Xinyue Shen, Michael Backes 외 arxiv

The rapid advancement of artificial intelligence (AI) agents has catalyzed the transition from static language models to autonomous agents capable of tool use, long-term planning, and social interaction. $\textbf{Moltboo…

Modeling Emotional Dynamics in Agent-to-Agent Interactions on Moltbook

2026-05-19 · Syed Mhamudul Hasan, Abdur R. Shahid arxiv

Generative AI systems are increasingly deployed as interactive agents in online environments, such as a social network called Moltbook. In Moltbook, large-scale agentic AIs can post, comment, and engage in activities gen…

Social Simulacra in the Wild: AI Agent Communities on Moltbook

2026-03-17 · Agam Goyal, Olivia Pal, Hari Sundaram, Eshwar Chandrasekharan 외 arxiv

As autonomous LLM-based agents increasingly populate social platforms, understanding the dynamics of AI-agent communities becomes essential for both communication research and platform governance. We present the first la…

What Do AI Agents Talk About? Discourse and Architectural Constraints in the First AI-Only Social Network

2026-03-09 · Taksch Dube, Jianfeng Zhu, NHatHai Phan, Ruoming Jin arxiv

Moltbook is the first large-scale social network built for autonomous AI agent-to-agent interaction. Early studies on Moltbook have interpreted its agent discourse as evidence of peer learning and emergent social behavio…

Emotion Classification