paper-with-me

홈 › Papers

ReaSon: Reinforced Causal Search with Information Bottleneck for Video Understanding

2025-11-16 · Yuan Zhou, Litao Hua, Shilong Jin, Wentao Huang, Haoran Duan arxiv

Keyframe selection has become essential for video understanding with vision-language models (VLMs) due to limited input tokens and the temporal sparsity of relevant information across video frames. Video understanding often relies on effective keyframes that are not only informative but also causally decisive. To this end, we propose Reinforced Causal Search with Information Bottleneck (ReaSon), a framework that formulates keyframe selection as an optimization problem with the help of a novel Causal Information Bottleneck (CIB), which explicitly defines keyframes as those satisfying both predictive sufficiency and causal necessity. Specifically, ReaSon employs a learnable policy network to select keyframes from a visually relevant pool of candidate frames to capture predictive sufficiency, and then assesses causal necessity via counterfactual interventions. Finally, a composite reward aligned with the CIB principle is designed to guide the selection policy through reinforcement learning. Extensive experiments on NExT-QA, EgoSchema, and Video-MME demonstrate that ReaSon consistently outperforms existing state-of-the-art methods under limited-frame settings, validating its effectiveness and generalization ability.

📄 PDF Abstract BibTeX arXiv:2511.12530

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Reinforced Context Order Recovery for Adaptive Reasoning and Planning

2025-08-18 · Long Ma, Fangwei Zhong, Yizhou Wang arxiv

Modern causal language models, followed by rapid developments in discrete diffusion models, can now produce a wide variety of interesting and useful content. However, these families of models are predominantly trained to…

The Causal Information Bottleneck and Optimal Causal Variable Abstractions

2024-10-01 · Francisco N. F. Q. Simoes, Mehdi Dastani, Thijs van Ommen

To effectively study complex causal systems, it is often useful to construct abstractions of parts of the system by discarding irrelevant details while preserving key features. The Information Bottleneck (IB) method is a…

Representation Learning

DIO: Refining Mutual Information and Causal Chain to Enhance Machine Abstract Reasoning Ability

2025-08-21 · Ruizhuo Song, Beiming Yuan arxiv

Despite deep learning's broad success, its abstract-reasoning bottleneck persists. We tackle Raven's Progressive Matrices (RPM), the benchmark for pattern, reasoning and problem-solving intelligence. We model the full ca…

CausalReasoningBenchmark: A Real-World Benchmark for Disentangled Evaluation of Causal Identification and Estimation

2026-02-24 · Ayush Sawarni, Jiyuan Tan, Vasilis Syrgkanis arxiv

Many benchmarks for automated causal inference evaluate a system's performance based on a single numerical output, such as an Average Treatment Effect (ATE). This approach conflates two distinct steps in causal analysis:…

Causal Inference

Causal-AgentIR: Self-Evolving Causal Memory for Adaptive Image Restoration Agents

2026-07-23 · Hu Gao, Yulong Chen, Lizhuang Ma arxiv

Image restoration agents have recently emerged as a flexible paradigm for handling diverse and unpredictable degradations in real-world scenarios. Existing agents typically formulate restoration as a tool-using process, …

Image Restoration