paper-with-me

Papers

UniRL-Zero: Reinforcement Learning on Unified Models with Joint Language Model and Diffusion Model Experts

2025-10-20 · Fu-Yun Wang, Han Zhang, Michael Gharbi, Hongsheng Li, Taesung Park arxiv

We present UniRL-Zero, a unified reinforcement learning (RL) framework that boosts, multimodal language model understanding and reasoning, diffusion model multimedia generation, and their beneficial interaction capabilities within a unified model. Our work defines six scenarios for unified model reinforcement learning, providing systematic baselines for reinforcement learning of unified understanding and generation model. Our code is available at https://github.com/G-U-N/UniRL.

📄 PDF Abstract BibTeX arXiv:2510.17937

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning

2025-05-29 · Weijia Mao, Zhenheng Yang, Mike Zheng Shou

Unified multimodal large language models such as Show-o and Janus have achieved strong performance across both generation and understanding tasks. However, these models typically rely on large-scale datasets and require …

RobotDancing: Residual-Action Reinforcement Learning Enables Robust Long-Horizon Humanoid Motion Tracking

2025-09-25 · Zhenguo Sun, Yibo Peng, Yuan Meng, Xukun Li 외 arxiv

Long-horizon, high-dynamic motion tracking on humanoids remains brittle because absolute joint commands cannot compensate model-plant mismatch, leading to error accumulation. We propose RobotDancing, a simple, scalable f…

Reinforcement Learning

OpenTable-R1: A Reinforcement Learning Augmented Tool Agent for Open-Domain Table Question Answering

2025-07-02 · Zipeng Qiu

Open-domain table question answering traditionally relies on a two-stage pipeline: static table retrieval followed by a closed-domain answer. In contrast, we propose an end-to-end agentic framework that embeds multi-turn…

Language ModelingLanguage ModellingLarge Language ModelQuestion Answering+1

Learning Language Specific Sub-network for Multilingual Machine Translation

2021-05-19 · ACL 2021 5 · Zehui Lin, Liwei Wu, Mingxuan Wang, Lei LI

Multilingual neural machine translation aims at learning a single translation model for multiple languages. These jointly trained models often suffer from performance degradation on rich-resource language pairs. We attri…

AttributeMachine TranslationTranslation

A Unified Framework for Zero-Shot Reinforcement Learning

2025-10-23 · Jacopo Di Ventura, Jan Felix Kleuker, Aske Plaat, Thomas Moerland arxiv

Zero-shot reinforcement learning (RL) has emerged as a setting for developing general agents, capable of solving downstream tasks without additional training or planning at test-time. While conventional RL optimizes poli…

Reinforcement Learning