paper-with-me

Papers

Position: Don't Just "Fix it in Post": A Science of AI Must Study Training Dynamics

2026-06-03 · Stella Biderman, Mohammad Aflah Khan, Niloofar Mireshghallah, Catherine Arnett, Fazl Barez, Naomi Saphra arxiv

What would it mean to have a scientific understanding of AI? Models are not static objects: they are snapshots of time-evolving processes shaped by data, objectives, architectures, and optimization dynamics. Yet much of AI research treats models as fixed artifacts, analyzing behaviors after training rather than asking why they emerge. This position paper argues that a science of AI must move beyond post-hoc fixes and study the training dynamics that produce model behavior. Such a science should support progressively stronger forms of understanding: predicting outcomes from early training signals, intervening when trajectories go wrong, and ultimately designing training procedures that more reliably produce desired properties. Scaling laws have made prediction routine for loss; the challenge is extending this success to capabilities, biases, robustness, and safety-relevant behaviors. We articulate requirements for such theories grounded in the history and philosophy of science, examine progress in mechanistic interpretability, fairness, memorization, and simplicity bias, and identify concrete open problems.

📄 PDF Abstract BibTeX arXiv:2606.06533

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Data Science as Political Action: Grounding Data Science in a Politics of Justice

2018-11-06 · Ben Green

In response to public scrutiny of data-driven algorithms, the field of data science has adopted ethics training and principles. Although ethics can help data scientists reflect on certain normative aspects of their work,…

Ethics

Emancipatory Information Retrieval

2025-01-31 · Bhaskar Mitra

Our world today is facing a confluence of several mutually reinforcing crises each of which intersects with concerns of social justice and emancipation. This paper is a provocation for the role of computer-mediated infor…

Information RetrievalRetrieval

Faithful or Just Plausible? Evaluating the Faithfulness of Closed-Source LLMs in Medical Reasoning

2026-03-14 · Halimat Afolabi, Zainab Afolabi, Elizabeth Friel, Jude Roberts 외 arxiv

Closed-source large language models (LLMs), such as ChatGPT and Gemini, are increasingly consulted for medical advice, yet their explanations may appear plausible while failing to reflect the model's underlying reasoning…

Can Bioinformatics Be Considered as an Experimental Biological Science?

2016-07-17

The objective of this short report is to reconsider the subject of bioinformatics as just being a tool of experimental biological science. To do that, we introduce three examples to show how bioinformatics could be consi…

PosterOmni: Generalized Artistic Poster Creation via Task Distillation and Unified Reward Feedback

2026-02-12 · Sixiang Chen, Jianyu Lai, Jialin Gao, Hengyu Shi 외 arxiv

Image-to-poster generation is a high-demand task requiring not only local adjustments but also high-level design understanding. Models must generate text, layout, style, and visual elements while preserving semantic fide…