paper-with-me

Papers

LifeSide: Benchmarking Agents as Lifelong Digital Companions

2026-06-03 · Yuqian Wu, Zhijie Deng, Wei Chen, Junwei Li, Yutian Jiang, Junle Chen, Zhengjun Huang, Qingxiang Liu, Jing Tang, Jiaheng Wei, Yuxuan Liang arxiv

Lifelong digital companions must integrate cross-session cues, continually update their understanding of users, and adapt to shifting privacy boundaries. Existing evaluations fail to capture this, testing memory recall and short-term empathy in isolation. To bridge this gap, we introduce \benchmark, a benchmark centered on multi-session \textit{Memory-Emotion-Environment} loops. By modeling users as persistent worlds with layered profiles and event trajectories, \benchmark uses multi-agent simulation to project environmental dynamics into dialogue, preserving the critical gap between latent thoughts and observable expressions. Evaluating 2,000 personas and 111K tasks across memory tracking, user understanding, privacy control, and emotional companionship, our experiment results reveal a stark reality: even models that saturate current memory benchmarks fail to sustain accurate user understanding and true companionship over long horizons.

📄 PDF Abstract BibTeX arXiv:2606.04660

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

KnowMe-Bench: Benchmarking Person Understanding for Lifelong Digital Companions

2026-01-08 · Tingyu Wu, Zhisheng Chen, Ziyan Weng, Shuhe Wang 외 arxiv

Existing long-horizon memory benchmarks mostly use multi-turn dialogues or synthetic user histories, which makes retrieval performance an imperfect proxy for person understanding. We present \BenchName, a publicly releas…

Deco: Extending Personal Physical Objects into Pervasive AI Companion through a Dual-Embodiment Framework

2026-05-05 · Zhihan Jiang, Mengyuan Millie Wu, Ruishi Zou, Shiyu Xu 외 arxiv

Individuals frequently form deep attachments to physical objects (e.g., plush toys) that usually cannot sense or respond to their emotions. While AI companions offer responsiveness and personalization, they exist indepen…

Digital Companionship: Overlapping Uses of AI Companions and AI Assistants

2025-09-16 · Aikaterina Manoli, Janet V. T. Pauketat, Ali Ladak, Hayoun Noh 외 arxiv

Large language models are increasingly used for both task-based assistance and social companionship, yet research has typically focused on one or the other. Drawing on a survey (N = 202) and 30 interviews with high-engag…

H2HTalk: Evaluating Large Language Models as Emotional Companion

2025-07-04 · Boyang Wang, Yalun Wu, Hongcheng Guo, Zhoujun Li arxiv

As digital emotional support needs grow, Large Language Model companions offer promising authentic, always-available empathy, though rigorous evaluation lags behind model advancement. We present Heart-to-Heart Talk (H2HT…

Emotional Intelligence

SaliMory: Orchestrating Cognitive Memory for Conversational Agents

2026-06-02 · Kai Zhang, Xinyuan Zhang, Hongda Jiang, Shiun-Zu Kuo 외 arxiv

Conversational agents that serve as lifelong companions must maintain persistent memory across all interactions. However, simply expanding context windows with raw retrieval degrades reasoning quality, while training mem…

Reinforcement Learning