paper-with-me

홈 › Papers

Synthetic Benchmarks Overstate Forward-Forward Scaling: Real-Data Limits of Layer-Local Training

2026-06-04 · Yucheng Chen arxiv

Forward-Forward (FF) learning [Hinton, 2022] replaces backpropagation with strictly layer-local goodness updates. Recent FF-CNN work has narrowed the gap to BP on 32x32 benchmarks, raising the question of whether layer-local training is becoming a viable alternative at realistic scale. To probe this rigorously, we develop DTG-FF -- dynamic temperature goodness, decoupled normalization, and multi-layer fusion -- as an instrument that sets FF-family state of the art across nine real-data benchmarks (91.8% CIFAR-10 and the first FF baseline at ImageNet-100 224x224), and use it to audit how far layer-local training actually scales. (1) Real-data scaling. Under identical recipe and backbone, an architecture-matched BP-DeepSup baseline beats DTG-FF by 2.40/5.93 pp on CIFAR-10/CIFAR-100, and the gap widens with class count. At 224x224 the same instrument reaches only 49.4% -- the first FF baseline at this scale, versus typical BP above 75% [Tian et al., 2020] -- exposing a real-data ceiling invisible at 32x32. (2) Synthetic vs. real K-conflict. DTG-FF increasingly outperforms BP as class count K grows on synthetic teacher-student tasks, yet on real images the FF-BP gap reverses sign and widens with K. A within-dataset CIFAR-100 coarse vs. fine probe isolates label-hierarchy from image distribution: synthetic K-sweeps confound output dimensionality with fine-grained discrimination difficulty and thereby overstate FF transferability. (3) Systems audit. FF can be implemented without storing depth-wide activations, but on commodity 8 GB hardware standard BP+gradient-accumulation reaches 4.18 GB / 157 imgs/s versus DTG-FF's 7.90 GB / 138 imgs/s, so a memory-based justification for FF at this scale is not supported under fair baselines.

📄 PDF Abstract BibTeX arXiv:2606.06539

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

UI-Oceanus: Scaling GUI Agents with Synthetic Environmental Dynamics

2026-02-11 · Mengzhou Wu, Yuzhe Guo, Yuan Cao, Haochuan Lu 외 arxiv

Scaling generalist GUI agents is hindered by the data scalability bottleneck of expensive human demonstrations and the "distillation ceiling" of synthetic teacher supervision. To transcend these limitations, we propose U…

Algorithmic Compliance and Regulatory Loss in Digital Assets

2026-03-04 · Khem Raj Bhatt, Krishna Sharma arxiv

We study the deployment performance of machine learning based enforcement systems used in cryptocurrency anti money laundering (AML). Using forward looking and rolling evaluations on Bitcoin transaction data, we show tha…

SynFlow: Scaling Up LiDAR Scene Flow Estimation with Synthetic Data

2026-04-10 · Qingwen Zhang, Xiaomeng Zhu, Chenhan Jiang, Patric Jensfelt arxiv

Reliable 3D dynamic perception requires models that can anticipate motion beyond predefined categories, yet progress is hindered by the scarcity of dense, high-quality motion annotations. While self-supervision on unlabe…

Scene Flow Estimation

Epistemic Integrity in Large Language Models

2024-11-10 · Bijean Ghafouri, Shahrad Mohammadzadeh, James Zhou, Pratheeksha Nair 외

Large language models are increasingly relied upon as sources of information, but their propensity for generating false or misleading statements with high confidence poses risks for users and society. In this paper, we c…

Common 7B Language Models Already Possess Strong Math Capabilities

2024-03-07 · Chen Li, Weiqi Wang, Jingcheng Hu, Yixuan Wei 외

Mathematical capabilities were previously believed to emerge in common language models only at a very large scale or require extensive math-related pre-training. This paper shows that the LLaMA-2 7B model with common pre…

GSM8KMath