paper-with-me

홈 › Papers

Ladders in Chaos: When, How, (and Perhaps Why) Does Test-Time Scaling Improve LLM Machine Translation

2026-08-28 · Di Wu, Sergey Troshin, Christof Monz, Antske Fokkens, Vlad Niculae arxiv

Two forms of test-time scaling for Large Language Models (LLMs) have emerged as effective and widely adopted paradigms: sequential, in which later answer attempts depend on earlier ones, and parallel, such as i.i.d. sampling with reranking. In this study, we investigate their properties in translation. First, our study shows that sequential sampling has a higher performance ceiling, providing a more diverse and effective pool of samples, particularly under smaller sampling budgets. Second, we interrogate the nature of test-time scaling through a multidimensional manual analysis. Human analysis of the Best-of-N translations demonstrates that sequential sampling substantially improves translation fluency and naturalness, but can degrade accuracy when inference budgets are large. Finally, we suggest an explanation of the mechanism through which sequential scaling improves machine translation. Our controlled analysis partially attributes the success of sequential self-improvement to the model's access to a larger target-side context. Ablation experiments on sequential sampling demonstrate its robustness across different sampling temperatures, while also revealing sensitivity to context construction, suggesting directions for future improvement.

📄 PDF Abstract BibTeX arXiv:2608.28496

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Translation

Similar Papers 제목 키워드 기반

Word Ladders: A Mobile Application for Semantic Data Collection

2024-03-29 · Marianna Marcella Bolognesi, Claudia Collacciani, Andrea Ferrari, Francesca Genovese 외

Word Ladders is a free mobile application for Android and iOS, developed for collecting linguistic data, specifically lists of words related to each other through semantic relations of categorical inclusion, within the A…

LadderSym: A Multimodal Interleaved Transformer for Music Practice Error Detection

2025-09-16 · Benjamin Shiue-Hal Chou, Purvish Jajal, Nick John Eliopoulos, James C. Davis 외 arxiv

Music learners can greatly benefit from tools that accurately detect errors in their practice. Existing approaches typically compare audio recordings to music scores using heuristics or learnable models. This paper intro…

Reinforcement Learning

Constructing Per-Shot Bitrate Ladders using Visual Information Fidelity

2024-08-04 · Krishna Srikar Durbha, Alan C. Bovik

Adaptive video streaming allows for the construction of bitrate ladders that deliver perceptually optimized visual quality to viewers under bandwidth constraints. Two common approaches to adaptation are per-title encodin…

Pricing path-dependent Bermudan options using Wiener chaos expansion: an embarrassingly parallel approach

2019-01-17 · Jérôme Lelong

In this work, we propose a new policy iteration algorithm for pricing Bermudan options when the payoff process cannot be written as a function of a lifted Markov process. Our approach is based on a modification of the we…

regression

Learning Sparse Causal Models is not NP-hard

2013-09-26 · Tom Claassen, Joris Mooij, Tom Heskes

This paper shows that causal model discovery is not an NP-hard problem, in the sense that for sparse graphs bounded by node degree k the sound and complete causal model can be obtained in worst case order N^{2(k+2)} inde…

Causal DiscoveryModel DiscoverySelection bias