paper-with-me

홈 › Papers

LsrIF: Enhancing Logic-Structured Instruction Following of Large Language Models

2026-01-10 · Qingyu Ren, Qianyu He, Jingwen Chang, Geng Zhang, Jiajie Zhu, Xingzhou Chen, Zhuofei Shi, Jiaqing Liang, Yanghua Xiao, Han Xia, Zeye Sun, Fei Yu arxiv

Instruction following is critical for large language models, yet real-world instructions often involve multiple constraints with logical structures, such as parallel composition, sequential dependencies, and conditional branching. Existing methods typically construct data by simply combining constraints and aggregate rewards by averaging individual constraint scores during training, overlooking logical dependencies and introducing noisy signals. We propose LsrIF, a training framework for logic-structured instruction following. LsrIF constructs data by organizing atomic constraints into parallel, sequential, conditional, and nested structures, and applies structure-aware reward aggregation aligned with their execution semantics: averaging rewards for parallel constraints, decaying later rewards after early failures in sequential structures, and rewarding only active branches in conditional structures. Experiments show that LsrIF improves instruction following in both in-domain and out-of-domain settings while also benefiting logic reasoning. Further analysis indicates that logic-structured training increases attention to constraint-related tokens and logical connectors, suggesting improved modeling of instruction logic. We will release our data and code for future research.

📄 PDF Abstract BibTeX arXiv:2601.06431

Code (0)

등록된 구현이 없습니다.

Tasks

Instruction Following

Similar Papers 제목 키워드 기반

GraphIF: Enhancing Multi-Turn Instruction Following for Large Language Models with Relation Graph Prompt

2025-11-13 · Zhenhe Li, Can Lin, Ling Zheng, Wen-Da Wei 외 arxiv

Multi-turn instruction following is essential for building intelligent conversational systems that can consistently adhere to instructions across dialogue turns. However, existing approaches to enhancing multi-turn instr…

Instruction FollowingRelation ExtractionResponse Generation

StyleAdaptedLM: Enhancing Instruction Following Models with Efficient Stylistic Transfer

2025-07-24 · Pritika Ramu, Apoorv Saxena, Meghanath M Y, Varsha Sankar 외 arxiv

Adapting LLMs to specific stylistic characteristics, like brand voice or authorial tones, is crucial for enterprise communication but challenging to achieve from corpora which lacks instruction-response formatting withou…

Instruction Following

ImpRIF: Stronger Implicit Reasoning Leads to Better Complex Instruction Following

2026-02-04 · Yuancheng Yang, Lin Yang, Xu Wang, Chao Tong 외 arxiv

As applications of large language models (LLMs) become increasingly complex, the demand for robust complex instruction following capabilities is growing accordingly. We argue that a thorough understanding of the instruct…

Reinforcement LearningInstruction Following

Zero-Shot Instruction Following in RL via Structured LTL Representations

2026-02-15 · Mathias Jackermeier, Mattia Giuri, Jacques Cloete, Alessandro Abate arxiv

We study instruction following in multi-task reinforcement learning, where an agent must zero-shot execute novel tasks not seen during training. In this setting, linear temporal logic (LTL) has recently been adopted as a…

Reinforcement LearningInstruction Following

Zero-Shot Instruction Following in RL via Structured LTL Representations

2025-12-02 · Mattia Giuri, Mathias Jackermeier, Alessandro Abate arxiv

Linear temporal logic (LTL) is a compelling framework for specifying complex, structured tasks for reinforcement learning (RL) agents. Recent work has shown that interpreting LTL instructions as finite automata, which ca…

Reinforcement LearningInstruction FollowingGraph Neural Network