paper-with-me

홈 › Papers

Compute Requirements for Algorithmic Innovation in Frontier AI Models

2025-07-13 · Peter Barnett arxiv

Algorithmic innovation in the pretraining of large language models has driven a massive reduction in the total compute required to reach a given level of capability. In this paper we empirically investigate the compute requirements for developing algorithmic innovations. We catalog 36 pre-training algorithmic innovations used in Llama 3 and DeepSeek-V3. For each innovation we estimate both the total FLOP used in development and the FLOP/s of the hardware utilized. Innovations using significant resources double in their requirements each year. We then use this dataset to investigate the effect of compute caps on innovation. Our analysis suggests that compute caps alone are unlikely to dramatically slow AI algorithmic progress. Even stringent compute caps -- such as capping total operations to the compute used to train GPT-2 or capping hardware capacity to 8 H100 GPUs -- could still have allowed for half of the cataloged innovations.

📄 PDF Abstract BibTeX arXiv:2507.10618

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Algorithmic progress in computer vision

2022-12-10 · Ege Erdil, Tamay Besiroglu

We investigate algorithmic progress in image classification on ImageNet, perhaps the most well-known test bed for computer vision. We estimate a model, informed by work on neural scaling laws, and infer a decomposition o…

Attributeimage-classificationImage Classification

Position: Require Frontier AI Labs To Release Small "Analog" Models

2025-10-15 · Shriyash Upadhyay, Chaithanya Bandi, Narmeen Oozeer, Philip Quirke arxiv

Recent proposals for regulating frontier AI models have sparked concerns about the cost of safety regulation, and most such regulations have been shelved due to the safety-innovation tradeoff. This paper argues for an al…

FrontierCS: Evolving Challenges for Evolving Intelligence

2025-12-17 · Qiuyang Mang, Wenhao Chai, Zhifei Li, Huanzhi Mao 외 arxiv

We introduce FrontierCS, a benchmark of 156 open-ended problems across diverse areas of computer science, designed and reviewed by experts, including CS PhDs and top-tier competitive programming participants and problem …

LLM-e Guess: Can LLMs Capabilities Advance Without Hardware Progress?

2025-05-07 · Teddy Foley, Spencer Guo, Henry Josephson, Anqi Qu 외

This paper examines whether large language model (LLM) capabilities can continue to advance without additional compute by analyzing the development and role of algorithms used in state-of-the-art LLMs. Motivated by regul…

Large Language ModelMixture-of-Experts

Lists of Top Artists to Watch computed algorithmically

2021-11-30 · Tomasz Imielinski

Lists of top artists to watch are periodically published by various art world media publications. These lists are selected editorially and reflect the subjective opinions of their creators. We show an application of rank…