paper-with-me

Instruction Following

2개 벤치마크 · 논문 1,603편 · 이 태스크의 논문 보기 →

Benchmarks

IFBench

결과 450개

IFEval

결과 5개

Most implemented

Visual Instruction Tuning

2023-04-17 · 구현 13개

Papers

SenseNova-U1.5: Towards Native Unified Visual Intelligence

2026-09-10 · Haiwen Diao, Jiahao Wang, Chenjing Ding, Hanming Deng 외 hf

We launch SenseNova-U1.5, an 8B-MoT native unified multimodal model that understands, reasons about, and generates visual content within an encoder-free and VAE-free architecture. We strengthen its visual interface throu…

Reinforcement LearningInstruction FollowingImage Editing

Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation

2026-09-10 · Jintao Zhang, Kai Jiang, Jintao Chen, Xu Wang 외 hf

We present Vidu S2, which comprises Vidu S2-Avatar, a real-time interactive digital-character model, and Vidu S2-Editing, a real-time video editing model. Moreover, we explore the feasibility of real-time spatial video g…

Instruction FollowingVideo Generation

Building Multilingual Bridges: Data Mixing as the Pillar of Generalization for In-Language Reasoning

2026-09-09 · Mehrnaz Mofakhami, Ananya Sahu, Alejandro R. Salamanca, Daniel D'souza 외 arxiv

Reasoning language models have made substantial advances on a variety of complex tasks, yet their capabilities remain overwhelmingly English-centric: models primarily reason in English regardless of the language they are…

Instruction Following

NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness

2026-09-08 · NeoHorse Team, Guoliang Cao, Guohao Dai, Tianyu Guo 외 hf

Recursive self-improvement (RSI) requires a concrete mechanism through which an AI system observes its capabilities and converts that evidence into the next round of learning. We present NeoHorse-1, a family of agent-nat…

Instruction Following

MobileVLA-R1 2.0: RL-Enhanced Reasoning for Mobile Robot Control

2026-09-05 · Ting Huang, Yue Huang, Zeyu Zhang, Shuicheng Yan 외 hf

Grounding natural-language instructions into reliable and executable actions remains a fundamental challenge for vision-language-action (VLA) systems on mobile robots, due to the persistent gap between high-level semanti…

Reinforcement LearningInstruction FollowingMultimodal ReasoningDecision Making

EuroAlpaca: Task-Preserving Localisation of Instruction Data for European Languages

2026-09-04 · Aleix Sant, Jordi Luque, Carlos Escolano arxiv

Machine translation (MT) offers a scalable way to extend English instruction-tuning data to multiple languages, but it can distort task-critical constraints and required outputs, creating corrupted training examples and …

Instruction FollowingMachine Translation

전체 1,603편 보기 →