paper-with-me

Papers

PhysMirror: Physics-Aware Mirror Object Generation

2026-07-03 · Xuan-Bach Mai, Duy-Phuc Nguyen, Quoc-Van Le, Tam V. Nguyen, Thanh-Toan Do, Huu Le, Duong-Van Nguyen, Minh-Triet Tran, Trung-Nghia Le arxiv

Synthesizing physically accurate mirror reflections remains a fundamental challenge for modern text-to-image diffusion models, which are increasingly critical for generating synthetic training data for embodied AI and robotic perception. These models typically struggle with strict geometric constraints, leading to hallucinations that degrade the utility of the synthetic data. To address this, we introduce a novel, end-to-end physics-aware generation framework namely PhysMirror that natively enforces projective geometry through explicit 3D spatial priors. Our method automatically lifts prompted objects into 3D meshes and constructs a lightweight, mathematically exact mirror scene within a simulated environment. By rendering this explicit 3D scene, we extract precise 2D conditioning elements, such as depth maps and segmentation maps, that serve as robust guiding signals for downstream diffusion models, guiding them to generate images with physically correct mirror reflections. Moreover, we introduce Mirror Consistency Score (MCS), reference-free, fully automated metric that quantifies physical correctness using dense feature matching and vanishing point convergence. Experimental results on our newly constructed MirrOB dataset demonstrate that our approach outperforms state-of-the-art baselines in reflection accuracy and physical realism, while maintaining strong text-to-image semantic alignment, providing a reliable pipeline for embodied AI data generation. The source code is released at https://duyphuc0701.github.io/PhysMirror.

📄 PDF Abstract BibTeX arXiv:2607.03470

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Goal Force: Teaching Video Models To Accomplish Physics-Conditioned Goals

2026-01-09 · Nate Gillman, Yinghua Zhou, Zitian Tang, Evan Luo 외 arxiv

Recent advancements in video generation have enabled the development of ``world models'' capable of simulating potential futures for robotics and planning. However, specifying precise goals for these models remains a cha…

Zero-shot GeneralizationVideo Generation

Reflecting Reality: Enabling Diffusion Models to Produce Faithful Mirror Reflections

2024-09-23 · Ankit Dhiman, Manan Shah, Rishubh Parihar, Yash Bhalgat 외

We tackle the problem of generating highly realistic and plausible mirror reflections using diffusion-based generative models. We formulate this problem as an image inpainting task, allowing for more user control over th…

Image Inpainting

Symmetry-Aware Transformer-based Mirror Detection

2022-07-13 · Tianyu Huang, Bowen Dong, Jiaying Lin, Xiaohui Liu 외

Mirror detection aims to identify the mirror regions in the given input image. Existing works mainly focus on integrating the semantic features and structural features to mine specific relations between mirror and non-mi…

DecoderMirror Detection

PAVAS: Physics-Aware Video-to-Audio Synthesis

2025-12-09 · Oh Hyun-Bin, Yuhta Takida, Toshimitsu Uesaka, Tae-Hyun Oh 외 arxiv

Recent advances in Video-to-Audio (V2A) generation have achieved impressive perceptual quality and temporal synchronization, yet most models remain appearance-driven, capturing visual-acoustic correlations without consid…

3D Reconstruction

MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation

2026-08-07 · Youjun Zhao, Alex Warren, Gary K. L. Tam, Rynson W. H. Lau hf

Recent advances in video diffusion models (VDMs) have enabled high-fidelity video synthesis. However, generating mirror reflections remains challenging because the content within a mirror must remain consistent with the …

Video Inpainting