paper-with-me

홈 › Papers

A Bitter Lesson for Data Filtering

2026-05-19 · Christopher Mohri, John Duchi, Tatsunori Hashimoto arxiv

We investigate data filtering for large model pretraining via new scaling studies that target the high compute, data-scarce regime. In spite of an apparently common belief that filtering data to include only high-quality information is essential, our experiments suggest that with enough compute, the best data filter is no data filter. We find that sufficiently trained large parameter models not only tolerate low-quality and distractor data, but in fact benefit from nominally ``poor'' data.

📄 PDF Abstract BibTeX arXiv:2605.19407

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Pan-Cancer mitotic figures detection and domain generalization: MIDOG 2025 Challenge

2025-08-28 · Zhuoyan Shen, Esther Bär, Maria Hawkins, Konstantin Bräutigam 외 arxiv

This report details our submission to the Mitotic Domain Generalization (MIDOG) 2025 challenge, which addresses the critical task of mitotic figure detection in histopathology for cancer prognostication. Following the "B…

Domain Generalization

Learning the Bitter Lesson: Empirical Evidence from 20 Years of CVPR Proceedings

2024-10-12 · Mojtaba Yousefi, Jack Collins

This study examines the alignment of \emph{Conference on Computer Vision and Pattern Recognition} (CVPR) research with the principles of the "bitter lesson" proposed by Rich Sutton. We analyze two decades of CVPR abstrac…

Simple Supervision Is Hard to Beat: A Bitter Lesson from Sparse Target Labels in Domain-Adaptive Object Detection

2026-06-29 · Lijun Zhang, Ruinian Xu, Mudit Agrawal arxiv

Source-free domain adaptive object detection adapts a source-trained detector to an unlabeled target domain, typically through teacher-student self-training with pseudo-labels. We revisit this setting when a small, unifo…

Object Detection

What the F*ck Is Artificial General Intelligence?

2025-03-31 · Michael Timothy Bennett

Artificial general intelligence (AGI) is an established field of research. Yet Melanie Mitchell and others have questioned if the term still has meaning. AGI has been subject to so much hype and speculation it has become…

The Bitter Lesson of Diffusion Language Models for Agentic Workflows: A Comprehensive Reality Check

2026-01-19 · Qingyu Lu, Liang Ding, Kanjian Zhang, Jinxia Zhang 외 arxiv

The pursuit of real-time agentic interaction has driven interest in Diffusion-based Large Language Models (dLLMs) as alternatives to auto-regressive backbones, promising to break the sequential latency bottleneck. Howeve…