paper-with-me

Papers

Combining Data-driven Supervision with Human-in-the-loop Feedback for Entity Resolution

2021-11-20 · Wenpeng Yin, Shelby Heinecke, Jia Li, Nitish Shirish Keskar, Michael Jones, Shouzhong Shi, Stanislav Georgiev, Kurt Milich, Joseph Esposito, Caiming Xiong

The distribution gap between training datasets and data encountered in production is well acknowledged. Training datasets are often constructed over a fixed period of time and by carefully curating the data to be labeled. Thus, training datasets may not contain all possible variations of data that could be encountered in real-world production environments. Tasked with building an entity resolution system - a model that identifies and consolidates data points that represent the same person - our first model exhibited a clear training-production performance gap. In this case study, we discuss our human-in-the-loop enabled, data-centric solution to closing the training-production performance divergence. We conclude with takeaways that apply to data-centric learning at large.

📄 PDF Abstract BibTeX arXiv:2111.10497

Code (0)

등록된 구현이 없습니다.

Tasks

Entity Resolution

Similar Papers 제목 키워드 기반

A Human-in/on-the-Loop Framework for Accessible Text Generation

2026-03-19 · Lourdes Moreno, Paloma Martínez arxiv

Plain Language and Easy-to-Read formats in text simplification are essential for cognitive accessibility. Yet current automatic simplification and evaluation pipelines remain largely automated, metric-driven, and fail to…

Text SimplificationText Generation

Beyond Self-Play: Hierarchical Reasoning for Continuous Motion in Closed-Loop Traffic Simulation

2026-05-09 · Weifan Zhang, Xiaofeng Zhao, Adel Bazzi, Mingrui Li 외 arxiv

Closed-loop traffic simulation requires agents that are both scalable and behaviorally realistic. Recent self-play reinforcement learning approaches demonstrate strong scalability, but their equilibrium strategies fail t…

Multi-agent Reinforcement Learning

IMPACT-Scribe: Interactive Temporal Action Segmentation with Boundary Scribbles and Query Planning

2026-05-03 · Qian Yin, Di Wen, Kunyu Peng, David Schneider 외 arxiv

Dense temporal annotation of procedural activity videos is vital for action understanding and embodied intelligence but remains labor-intensive due to reactive tools. Each correction is treated as an isolated edit, limit…

Action UnderstandingAction Segmentation

Task-oriented grasping for dexterous robots using postural synergies and reinforcement learning

2026-02-24 · Dimitrios Dimou, José Santos-Victor, Plinio Moreno arxiv

In this paper, we address the problem of task-oriented grasping for humanoid robots, emphasizing the need to align with human social norms and task-specific objectives. Existing methods, employ a variety of open-loop and…

Reinforcement Learning

A knowledge-augmented dataset of high-risk driving scenarios with LLM annotations for autonomous driving

2026-07-08 · Heye Huang, Jingguang Li, Zhiyuan Zhou, Paul Liang 외 arxiv

Safe autonomous driving requires both rapid responses to common high-risk events and deeper reasoning over rare, extreme long-tail scenarios in traffic safety. These scenarios are severely under-represented in naturalist…

Autonomous Driving