paper-with-me

Papers

MRD: Using Physically Based Differentiable Rendering to Probe Vision Models for 3D Scene Understanding

2025-12-13 · Benjamin Beilharz, Thomas S. A. Wallis arxiv

While deep learning methods have achieved impressive success in many vision benchmarks, it remains difficult to understand and explain the representations and decisions of these models. Though vision models are typically trained on 2D inputs, they are often assumed to develop an implicit representation of the underlying 3D scene (for example, showing tolerance to partial occlusion, or the ability to reason about relative depth). Here, we introduce MRD (metamers rendered differentiably), an approach that uses physically based differentiable rendering to probe vision models' implicit understanding of generative 3D scene properties, by finding 3D scene parameters that are physically different but produce the same model activation (i.e. are model metamers). Unlike previous pixel-based methods for evaluating model representations, these reconstruction results are always grounded in physical scene descriptions. This means we can, for example, probe a model's sensitivity to object shape while holding material and lighting constant. As a proof-of-principle, we assess multiple models in their ability to recover scene parameters of geometry (shape) and bidirectional reflectance distribution function (material). The results show high similarity in model activation between target and optimized scenes, with varying visual results. Qualitatively, these reconstructions help investigate the physical scene attributes to which models are sensitive or invariant. MRD holds promise for advancing our understanding of both computer and human vision by enabling analysis of how physical scene parameters drive changes in model responses.

📄 PDF Abstract BibTeX arXiv:2512.12307

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Understanding

Similar Papers 제목 키워드 기반

Accidental Light Probes

2023-01-12 · CVPR 2023 1 · Hong-Xing Yu, Samir Agarwala, Charles Herrmann, Richard Szeliski 외

Recovering lighting in a scene from a single image is a fundamental problem in computer vision. While a mirror ball light probe can capture omnidirectional lighting, light probes are generally unavailable in everyday ima…

Lighting Estimation

Beyond Pixels: A Differentiable Pipeline for Probing Neuronal Selectivity in 3D

2025-10-15 · Pavithra Elumalai, Mohammad Bashiri, Goirik Chakrabarty, Suhas Shrinivasan 외 arxiv

Visual perception relies on inference of 3D scene properties such as shape, pose, and lighting. To understand how visual sensory neurons enable robust perception, it is crucial to characterize their selectivity to such p…

Efficient Multi-View Inverse Rendering Using a Hybrid Differentiable Rendering Method

2023-08-19 · Xiangyang Zhu, Yiling Pan, Bailin Deng, Bin Wang

Recovering the shape and appearance of real-world objects from natural 2D images is a long-standing and challenging inverse rendering problem. In this paper, we introduce a novel hybrid differentiable rendering method to…

3D geometryInverse Rendering

RenderBender: A Survey on Adversarial Attacks Using Differentiable Rendering

2024-11-14 · Matthew Hull, Haoran Wang, Matthew Lau, Alec Helbling 외

Differentiable rendering techniques like Gaussian Splatting and Neural Radiance Fields have become powerful tools for generating high-fidelity models of 3D objects and scenes. Their ability to produce both physically pla…

Depth EstimationImage ClassificationInverse Renderingobject-detection+3

Materialist: Physically Based Editing Using Single-Image Inverse Rendering

2025-01-07 · Lezhong Wang, Duc Minh Tran, Ruiqi Cui, Thomson TG 외

Achieving physically consistent image editing remains a significant challenge in computer vision. Existing image editing methods typically rely on neural networks, which struggle to accurately handle shadows and refracti…

Inverse Rendering