paper-with-me

홈 › Papers

Continued AI Scaling Requires Repeated Efficiency Doublings

2026-03-30 · Chien-Ping Lu arxiv

This paper argues that continued AI scaling requires repeated efficiency doublings. Classical AI scaling laws remain useful because they make progress predictable despite diminishing returns, but the compute variable in those laws is best read as logical compute, not as a record of one fixed physical implementation. Practical burden therefore depends on the efficiency with which physical resources realize that compute. Under that interpretation, diminishing returns mean rising operational burden, not merely a flatter curve. Sustained progress then requires recurrent gains in hardware, algorithms, and systems that keep additional logical compute feasible at acceptable cost. The relevant analogy is Moore's Law, understood less as a theorem than as an organizing expectation of repeated efficiency improvement. AI does not yet have a single agreed cadence for such gains, but recent evidence suggests trends that are at least Moore-like and sometimes faster. The paper's claim is therefore simple: if AI scaling is to remain active, repeated efficiency doublings are not optional. They are required.

📄 PDF Abstract BibTeX arXiv:2603.28507

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Modernizing Amdahl's Law: How AI Scaling Laws Shape Computer Architecture

2026-03-21 · Chien-Ping Lu arxiv

Classical Amdahl's Law conceptualized the limit of speedup for an era of fixed serial-parallel decomposition and homogeneous replication. Modern heterogeneous systems need a different conceptual framework: constrained re…

The Data Efficiency Frontier of Financial Foundation Models: Scaling Laws from Continued Pretraining

2025-12-13 · Jesse Ponnock arxiv

Domain-adaptive pretraining (DAPT) offers a practical path to specializing large language models for high-value domains without full retraining. We conduct an early-stage scaling-law analysis of continued pretraining on …

Domain Adaptation

Structured Singular Value of a Repeated Complex Full-Block Uncertainty

2022-11-11 · Talha Mushtaq, Diganta Bhattacharjee, Peter Seiler, Maziar S. Hemati

The structured singular value (SSV), or mu, is used to assess the robust stability and performance of an uncertain linear time-invariant system. Existing algorithms compute upper and lower bounds on the SSV for structure…

Computational Efficiency

The interplay between domain specialization and model size

2025-01-03 · Roseval Malaquias Junior, Ramon Pires, Thales Sales Almeida, Kenzo Sakiyama 외

Scaling laws for language models have often focused on finding the optimal model size and token count for training from scratch. However, achieving this optimal balance requires significant compute resources due to the e…

Do We Truly Need So Many Samples? Multi-LLM Repeated Sampling Efficiently Scales Test-Time Compute

2025-04-01 · Jianhao Chen, Zishuo Xun, Bocheng Zhou, Han Qi 외

This paper presents a simple, effective, and cost-efficient strategy to improve LLM performance by scaling test-time compute. Our strategy builds upon the repeated-sampling-then-voting framework, with a novel twist: inco…