paper-with-me

홈 › Papers

When Does Complexity Conditioning Help a Frozen Sentence Embedding? A Controlled Study of Per-Sentence and Pair-Level Difficulty Adaptation

2026-06-02 · Suhwan Hwang arxiv

A common intuition is that sentence embeddings should adapt to the difficulty of the input. We test this intuition in a controlled, multi-seed setting: a lightweight post-encoder adapter attaches to a frozen Qwen3-Embedding-0.6B encoder, accessing only its final pooled embedding, and is evaluated on four paraphrase and semantic-similarity tasks (PAWS, MRPC, QQP, STS-B). The naive form of the idea fails: surface-based per-sentence complexity is nearly uncorrelated with frozen-baseline error (Pearson approximately 0.05) and provides no advantage over constant or shuffled controls, while degrading a saturated baseline. Even when the target is aligned to a non-circular pair-difficulty signal, the per-sentence gate still cannot reliably capture difficulty because difficulty is primarily a property of the pair, not the individual sentence. In contrast, a small pair-level residual gated by a held-out cross-encoder difficulty signal yields consistent gains on the larger and graded tasks, including +0.022 Spearman on STS-B and +0.037 on QQP, while remaining anchored to the frozen baseline across all seeds. Because this useful form operates on sentence pairs rather than individual sentences, the resulting model is best understood as a lightweight re-ranker over cached frozen embeddings, not a replacement single-vector embedding; we make no state-of-the-art claim. Our contribution is a controlled account of when difficulty-aware adaptation helps and when it fails, together with a pre-training diagnostic that predicts the available headroom.

📄 PDF Abstract BibTeX arXiv:2606.03244

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ERank in Latent Space as an Image-Complexity and Richness Measure

2026-07-21 · Maksim Smirnov, Grigory Kononov, Anastasiia Linich, Egor Surkov 외 arxiv

We propose the effective rank (ERank) of the channel covariance of an image's deep feature map as a per-sample, label-free measure of visual richness, computed from a single forward pass through a frozen pretrained encod…

Selection Without Signal, Recovery Through Expression: A Measurement Study of Post-Hoc Falsification Operators for Frozen Small Code Models

2026-06-15 · Mehmet Iscan arxiv

Frozen small code models (<=1.5B parameters, run locally without fine-tuning) suit offline and privacy-constrained use, but often emit plausible-but-wrong programs. A natural remedy is a post-hoc operator that selects, v…

Monotonicity Testing of High-Dimensional Distributions with Subcube Conditioning

2025-02-22 · Deeparnab Chakrabarty, Xi Chen, Simeon Ristic, C. Seshadhri 외

We study monotonicity testing of high-dimensional distributions on $\{-1,1\}^n$ in the model of subcube conditioning, suggested and studied by Canonne, Ron, and Servedio~\cite{CRS15} and Bhattacharyya and Chakraborty~\ci…

When Does In-Context Search Help? A Sampling-Complexity Theory of Reflection-Driven Reasoning

2026-07-07 · Yotam Wolf, Noam Wies, Amnon Shashua arxiv

Training large language models (LLMs) with extended reasoning has enabled in-context search, in which models iteratively generate, critique, and revise solution attempts. We provide a theoretical analysis of in-context s…

Reinforcement Learning

Investigating Transferability in Pretrained Language Models

2020-04-30 · Findings of the Association for Computational Linguistics 2020 · Alex Tamkin, Trisha Singh, Davide Giovanardi, Noah Goodman

How does language model pretraining help transfer learning? We consider a simple ablation technique for determining the impact of each pretrained layer on transfer task performance. This method, partial reinitialization,…

Language ModelingLanguage ModellingTransfer Learning