paper-with-me

홈 › Papers

Two to Tango: Coupled Task-Reference Selection for Safe LLM Fine-tuning

2026-06-01 · Xinrui Chen, Jianhao Zhang, Ou Wu, Di Gao arxiv

Fine-tuning safety aligned large language models (LLMs) on downstream data improves adaptation but may erode learned safety behavior. Existing methods use fixed safety examples, global constraints, or one-sided task filtering. Our diagnostics show task updates expose different safety constraints, motivating joint selection of relevant references and compatible task samples. We propose DualSelect, a coupled framework for task and reference selection that refreshes task conditioned safety references before filtering whole task samples compatible with the induced reference direction. Under a minimax view, DualSelect selects safety references with high preservation loss and task conflict, together with compatible task samples, through entropy-regularized scoring surrogates, lazy reference refresh, and gradient correction. On 1B-8B LLMs, DualSelect preserves safety without losing task utility; using the REDORCA judge, it improves Safety Avg. over the strongest baseline by at least 5.10 points and remains highest in Safety Avg. across judges with moderate overhead. This view extends to retention focused continual learning.

📄 PDF Abstract BibTeX arXiv:2606.09866

Code (0)

등록된 구현이 없습니다.

Tasks

Continual Learning

Similar Papers 제목 키워드 기반

TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation

2024-10-05 · Haiyang Liu, Xingchao Yang, Tomoya Akiyama, Yuantian Huang 외

We present TANGO, a framework for generating co-speech body-gesture videos. Given a few-minute, single-speaker reference video and target speech audio, TANGO produces high-fidelity videos with synchronized body gestures.…

cross-modal alignmentRetrievalvalid

TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization

2024-12-30 · Chia-Yu Hung, Navonil Majumder, Zhifeng Kong, Ambuj Mehrish 외

We introduce TangoFlux, an efficient Text-to-Audio (TTA) generative model with 515M parameters, capable of generating up to 30 seconds of 44.1kHz audio in just 3.7 seconds on a single A40 GPU. A key challenge in aligning…

Audio GenerationGPU

Tango: Taming Visual Signals for Efficient Video Large Language Models

2026-04-10 · Shukang Yin, Sirui Zhao, Hanchao Wang, Baozhi Jia 외 arxiv

Token pruning has emerged as a mainstream approach for developing efficient Video Large Language Models (Video LLMs). This work revisits and advances the two predominant token-pruning paradigms: attention-based selection…

Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

2024-04-15 · Navonil Majumder, Chia-Yu Hung, Deepanway Ghosal, Wei-Ning Hsu 외

Generative multimodal content is increasingly prevalent in much of the content creation arena, as it has the potential to allow artists and media personnel to create pre-production mockups by quickly bringing their ideas…

Audio Generation

TANGO: Clustering with Typicality-Aware Nonlocal Mode-Seeking and Graph-Cut Optimization

2024-08-19 · Haowen Ma, Zhiguo Long, Hua Meng

Density-based clustering methods by mode-seeking usually achieve clustering by using local density estimation to mine structural information, such as local dependencies from lower density points to higher neighbors. Howe…

ClusteringDensity Estimation