paper-with-me

Papers

Multi-Objective Linguistic Control of Large Language Models

2024-06-23 · Dang Nguyen, Jiuhai Chen, Tianyi Zhou

Large language models (LLMs), despite their breakthroughs on many challenging benchmark tasks, lean to generate verbose responses and lack the controllability of output complexity, which is usually preferred by human users in practice. In this paper, we study how to precisely control multiple linguistic complexities of LLM output by finetuning using off-the-shelf data. To this end, we propose multi-control tuning (MCTune), which includes multiple linguistic complexity values of ground-truth responses as controls in the input for instruction tuning. We finetune LLaMA2-7B on Alpaca-GPT4 and WizardLM datasets. Evaluations on widely used benchmarks demonstrate that our method does not only improve LLMs' multi-complexity controllability substantially but also retains or even enhances the quality of the responses as a side benefit.

📄 PDF Abstract BibTeX arXiv:2406.16229

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

How does the pre-training objective affect what large language models learn about linguistic properties?

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Several pre-training objectives, such as masked language modeling (MLM), have been proposed to pre-train language models (e.g. BERT) with the aim of learning better language representations. However, to the best of our k…

Language ModelingLanguage ModellingMasked Language Modeling

How does the pre-training objective affect what large language models learn about linguistic properties?

2022-03-20 · ACL 2022 5 · Ahmed Alajrami, Nikolaos Aletras

Several pre-training objectives, such as masked language modeling (MLM), have been proposed to pre-train language models (e.g. BERT) with the aim of learning better language representations. However, to the best of our k…

Language ModelingLanguage ModellingMasked Language Modeling

Voice Impression Control in Zero-Shot TTS

2025-06-06 · Keinichi Fujita, Shota Horiguchi, Yusuke Ijima

Para-/non-linguistic information in speech is pivotal in shaping the listeners' impression. Although zero-shot text-to-speech (TTS) has achieved high speaker fidelity, modulating subtle para-/non-linguistic information t…

Language ModelingLanguage ModellingLarge Language Modeltext-to-speech+1

Do Audio-Language Models Understand Linguistic Variations?

2024-10-21 · Ramaneswaran Selvakumar, Sonal Kumar, Hemant Kumar Giri, Nishit Anand 외

Open-vocabulary audio language models (ALMs), like Contrastive Language Audio Pretraining (CLAP), represent a promising new paradigm for audio-text retrieval using natural language queries. In this paper, for the first t…

Contrastive LearningNatural Language QueriesRetrievalText Retrieval+1

Dynamic Multi-Reward Weighting for Multi-Style Controllable Generation

2024-02-21 · Karin de Langis, Ryan Koo, Dongyeop Kang

Textual style expresses a diverse set of information, including interpersonal dynamics (e.g., formality) and the author's emotions or attitudes (e.g., disgust). An open question is how language models can be explicitly c…

Multi-Objective Reinforcement LearningReinforcement Learning (RL)