paper-with-me

홈 › Papers

WebChain: A Large-Scale Human-Annotated Dataset of Real-World Web Interaction Traces

2026-03-05 · Sicheng Fan, Rui Wan, Yifei Leng, Gaoning Liang, Li Ling, Yanyi Shang, Dehan Kong arxiv

We introduce WebChain, the largest open-source dataset of human-annotated trajectories on real-world websites, designed to accelerate reproducible research in web agents. It contains 31,725 trajectories and 318k steps, featuring a core Triple Alignment of visual, structural, and action data to provide rich, multi-modal supervision. The data is collected via a scalable pipeline that ensures coverage of complex, high-value tasks often missed by synthetic methods. Leveraging this dataset, we propose a Dual Mid-Training recipe that decouples spatial grounding from planning, achieving state-of-the-art performance on our proposed WebChainBench and other public GUI benchmarks. Our work provides the data and insights necessary to build and rigorously evaluate the next generation of scalable web agents.

📄 PDF Abstract BibTeX arXiv:2603.05295

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos

2026-04-17 · Deepak Kumar, Abhishek Pratap Singh, Puneet Kumar, Xiaobai Li 외 arxiv

Understanding affective dynamics in real-world social systems is fundamental to modeling and analyzing human-human interactions in complex environments. Group affect emerges from intertwined human-human interactions, con…

InfiniHuman: Infinite 3D Human Creation with Precise Control

2025-10-13 · Yuxuan Xue, Xianghui Xie, Margaret Kostyrko, Gerard Pons-Moll arxiv

Generating realistic and controllable 3D human avatars is a long-standing challenge, particularly when covering broad attribute ranges such as ethnicity, age, clothing styles, and detailed body shapes. Capturing and anno…

Image Generation

Distilling Human-Aligned Privacy Sensitivity Assessment from Large Language Models

2026-03-31 · Gabriel Loiseau, Damien Sileo, Damien Riquet, Maxime Meyer 외 arxiv

Accurate privacy evaluation of textual data remains a critical challenge in privacy-preserving natural language processing. Recent work has shown that large language models (LLMs) can serve as reliable privacy evaluators…

UMDFaces: An Annotated Face Dataset for Training Deep Networks

2016-11-04 · Ankan Bansal, Anirudh Nanduri, Carlos Castillo, Rajeev Ranjan 외

Recent progress in face detection (including keypoint detection), and recognition is mainly being driven by (i) deeper convolutional neural network architectures, and (ii) larger datasets. However, most of the large data…

Face DetectionFace RecognitionKeypoint Detection

Cheems: A Practical Guidance for Building and Evaluating Chinese Reward Models from Scratch

2025-02-24 · Xueru Wen, Jie Lou, Zichao Li, Yaojie Lu 외

Reward models (RMs) are crucial for aligning large language models (LLMs) with human preferences. However, most RM research is centered on English and relies heavily on synthetic resources, which leads to limited and les…