paper-with-me

Papers

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech

2026-05-14 · Bin Kang, Shaoguo Wen, Yang Fan, Shunlong Wu, Junjie Wang, Yulin Li, Junzhi Zhao, Junle Wang, Zhuotao Tian arxiv

While existing text-to-speech (TTS) models exhibit high expressiveness, fine-grained control over composite instructions remains challenging due to the structural mismatch between discrete textual intents and continuous acoustic realizations. Inspired by human cognitive decoupling, we introduce AgentSteerTTS, a multi-agent closed-loop framework designed for intent-faithful expressive control of composite instructions. First, in our framework, an adversarial disentanglement agent mitigates speaker-emotion leakage by learning separable identity and emotion-prosody subspaces with leakage-suppressing regularization. Next, a Dual-Stream Anchoring Controller grounds abstract intents using a large-scale acoustic prototype library: a Retrieval Agent selects expressive anchors, while a Synthesis Agent fuses them into continuous control vectors via gated attention. Finally, a Fast-Slow Feedback Agent refines output intensity through latent gradient correction and resolves semantic-acoustic mismatches using high-level perceptual critique. Experiments on a composite-instruction benchmark and public test sets show that AgentSteerTTS yields consistent and significant improvements to the baselines, demonstrating the effectiveness of the proposed method.

📄 PDF Abstract BibTeX arXiv:2605.17583

Code (0)

등록된 구현이 없습니다.

Tasks

Continuous Control

Similar Papers 제목 키워드 기반

MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving

2026-05-13 · Rajeev Yasarla, Deepti Hegde, Hsin-Pai Cheng, Shizhong Han 외 arxiv

Vision-language-action (VLA) models are effective as end-to-end motion planners, but can be brittle when evaluated in closed-loop settings due to being trained under traditional imitation learning framework. Existing clo…

Reinforcement LearningAutonomous Driving

SoK: Measuring What Matters for Closed-Loop Security Agents

2025-10-02 · Mudita Khurana, Raunak Jain arxiv

Cybersecurity is a relentless arms race, with AI driven offensive systems evolving faster than traditional defenses can adapt. Research and tooling remain fragmented across isolated defensive functions, creating blind sp…

On Learning Closed-Loop Probabilistic Multi-Agent Simulator

2025-08-01 · Juanwu Lu, Rohit Gupta, Ahmadreza Moradipari, Kyungtae Han 외 arxiv

The rapid iteration of autonomous vehicle (AV) deployments leads to increasing needs for building realistic and scalable multi-agent traffic simulators for efficient evaluation. Recent advances in this area focus on clos…

Trajectory PredictionBayesian Inference

Revisit Mixture Models for Multi-Agent Simulation: Experimental Study within a Unified Framework

2025-01-28 · Longzhong Lin, Xuewu Lin, Kechun Xu, Haojian Lu 외

Simulation plays a crucial role in assessing autonomous driving systems, where the generation of realistic multi-agent behaviors is a key aspect. In multi-agent simulation, the primary challenges include behavioral multi…

Autonomous Driving

FICO: Finite-Horizon Closed-Loop Factorization for Unified Multi-Agent Path Finding

2025-11-17 · Jiarui Li, Alessandro Zanardi, Federico Pecora, Runyu Zhang 외 arxiv

Multi-Agent Path Finding is a fundamental problem in robotics and AI, yet most existing formulations treat planning and execution separately and address variants of the problem in an ad hoc manner. This paper presents a …