paper-with-me

홈 › Papers

On The Effectiveness-Fluency Trade-Off In LLM Conditioning: A Systematic Study

2026-06-10 · Iuri Macocco, Pau Rodríguez, Arno Blaas, Luca Zappella, Marco Baroni, Xavier Suau arxiv

Controlling the output of Large Language Models (LLMs) is a central challenge for their reliable deployment, yet a clear understanding of the involved trade-offs remains elusive. Current approaches to conditioning are often evaluated with a narrow focus on their effectiveness at injecting or removing a target concept, neglecting generation quality. We systematically investigate a range of conditioning methods in both injection and removal scenarios. We find that efficient steering methods frequently achieve conditioning at a steep cost to fluency. Furthermore, we identify a critical yet previously overlooked interaction with the training paradigm: activation steering methods are far less effective on instruction-tuned models than on their base counterparts. Simple prompting and full-fledged supervised fine-tuning, on the other hand, are viable options for concept injection, but are not as good at concept removal. Finally, cheaply computed textual metrics highly correlate to costly LLM-as-judge scores, and provide insights on the behavior of conditioning methods.

📄 PDF Abstract BibTeX arXiv:2606.12234

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Hierarchical Spoken Language Dysfluency Modeling

2024-01-18 · Jiachen Lian, Gopala Anumanchipalli

Speech disfluency modeling is the bottleneck for both speech therapy and language learning. However, there is no effective AI solution to systematically tackle this problem. We solidify the concept of disfluent speech an…

A Comparative Study of Controllability, Explainability, and Performance in Dysfluency Detection Models

2025-08-25 · Eric Zhang, Li Wei, Sarah Chen, Michael Wang arxiv

Recent advances in dysfluency detection have introduced a variety of modeling paradigms, ranging from lightweight object-detection inspired networks (YOLOStutter) to modular interpretable frameworks (UDM). While performa…

ETC-NLG: End-to-end Topic-Conditioned Natural Language Generation

2020-08-25 · Ginevra Carbone, Gabriele Sarti

Plug-and-play language models (PPLMs) enable topic-conditioned natural language generation by pairing large pre-trained generators with attribute models used to steer the predicted token distribution towards the selected…

AttributeComputational EfficiencyConditional Text GenerationText Generation+1

Revisiting LLM Adaptation for 3D CT Report Generation: A Study of Scaling and Diagnostic Priors

2026-06-15 · Vanshali Sharma, Andrea M. Bejar, Halil Ertugrul Aktas, Quoc-Huy Trinh 외 arxiv

Recent advances in multimodal learning, including large language models (LLMs) and vision-language models (VLMs), have demonstrated strong adaptability to natural images. However, extending their use to the medical domai…

Computational EfficiencyDomain Generalization

Simpson's Paradox and the Accuracy-Fluency Tradeoff in Translation

2024-02-20 · Zheng Wei Lim, Ekaterina Vylomova, Trevor Cohn, Charles Kemp

A good translation should be faithful to the source and should respect the norms of the target language. We address a theoretical puzzle about the relationship between these objectives. On one hand, intuition and some pr…

SentenceTranslation