paper-with-me

홈 › Papers

Do-Undo Bench: Reversibility for Action Understanding in Image Generation

2025-12-15 · Shweta Mahajan, Shreya Kadambi, Hoang Le, Rajeev Yasarla, Apratim Bhattacharyya, Munawar Hayat, Fatih Porikli arxiv

We introduce the Do-Undo task and benchmark to address a critical gap in vision-language models: understanding and generating plausible scene transformations driven by real-world actions. Unlike prior work that relies on prompt-based image generation and editing to perform action-conditioned image manipulation, our training hypothesis requires models to simulate the outcome of a real-world action and then reverse it to the original state. This forward-reverse requirement tests genuine cause-and-effect understanding rather than stylistic or semantic edits. We curate a high-quality benchmark of reversible actions from real-world scenarios to enable robust action grounding. Our experiments reveal that current models struggle with action reversibility, highlighting the need to evaluate action understanding. Do-Undo provides an intuitive testbed for evaluating and advancing action-aware generation in multimodal systems that must reason about real-world dynamics.

📄 PDF Abstract BibTeX arXiv:2512.13609

Code (0)

등록된 구현이 없습니다.

Tasks

Action UnderstandingImage ManipulationImage Generation

Similar Papers 제목 키워드 기반

Learning to Undo: Rollback-Augmented Reinforcement Learning with Reversibility Signals

2025-10-16 · Andrejs Sorstkins, Omer Tariq, Muhammad Bilal arxiv

This paper proposes a reversible learning framework to improve the robustness and efficiency of value based Reinforcement Learning agents, addressing vulnerability to value overestimation and instability in partially irr…

Reinforcement LearningDecision Making

Vision Language Models Know Law of Conservation without Understanding More-or-Less

2024-10-01 · Dezhi Luo, Haiyun Lyu, Qingying Gao, Haoran Sun 외

Understanding law of conservation is a critical milestone in human cognitive development considered to be supported by the apprehension of quantitative concepts and the reversibility of operations. To assess whether this…

On the reversibility of adversarial attacks

2022-06-01 · Chau Yi Li, Ricardo Sánchez-Matilla, Ali Shahin Shamsabadi, Riccardo Mazzon 외

Adversarial attacks modify images with perturbations that change the prediction of classifiers. These modified images, known as adversarial examples, expose the vulnerabilities of deep neural network classifiers. In this…

Adversarial Attack

Thermodynamic Irreversibility of Training Algorithms

2026-05-21 · Liu Ziyin, Yuanjie Ren, Adam Levine, Isaac Chuang arxiv

The training algorithms for AI systems all introduce far-from-equilibrium dynamical processes, and understanding the irreversibility of these algorithms is a fundamental step towards understanding the learning dynamics o…

Information decomposition reveals hidden high-order contributions to temporal irreversibility

2023-08-10 · Andrea I Luppi, Fernando E. Rosas, Gustavo Deco, Morten L. Kringelbach 외

Temporal irreversibility, often referred to as the arrow of time, is a fundamental concept in statistical mechanics. Markers of irreversibility also provide a powerful characterisation of information processing in biolog…

Time Series