paper-with-me

Papers

PhyS-EdiT: Physics-aware Semantic Image Editing with Text Description

2025-01-01 · CVPR 2025 1 · Ziqi Cai, Shuchen Weng, Yifei Xia, Boxin Shi

Achieving joint control over material properties, lighting, and high-level semantics in images is essential for applications in digital media, advertising, and interactive design. Existing methods often isolate these properties, lacking a cohesive approach to manipulating materials, lighting, and semantics simultaneously. We introduce PhyS-EdiT, a novel diffusion-based model that enables precise control over four critical material properties: roughness, metallicity, albedo, and transparency while integrating lighting and semantic adjustments within a single framework. To facilitate this disentangled control, we present PR-TIPS, a large and diverse synthetic dataset designed to improve the disentanglement of material and lighting effects. PhyS-EdiT incorporates a dual-network architecture and robust training strategies to balance low-level physical realism with high-level semantic coherence, supporting localized and continuous property adjustments. Extensive experiments demonstrate the superiority of PhyS-EdiT in editing both synthetic and real-world images, achieving state-of-the-art performance on material, lighting, and semantic editing tasks.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Disentanglement

Similar Papers 제목 키워드 기반

From Statics to Dynamics: Physics-Aware Image Editing with Latent Transition Priors

2026-02-25 · Liangbing Zhao, Le Zhuo, Sayak Paul, Hongsheng Li 외 arxiv

Instruction-based image editing has achieved remarkable success in semantic alignment, yet state-of-the-art models frequently fail to render physically plausible results when editing involves complex causal dynamics, suc…

Image Editing

Physics-Aware 3D Gaussian Editing for Driving Scene Generation

2026-05-25 · Feng Zhou, Jian Zhang, Yuhang Sun, He Wang 외 arxiv

3D Gaussian Splatting (3DGS) has shown great potential in autonomous driving simulation and data generation, enabling photorealistic reconstruction and flexible scene manipulation. However, existing 3DGS scene editing me…

Autonomous DrivingScene Generation

PhyEditBench: A Real-World Multi-Stage Benchmark for Physics-Aware Image Editing

2026-06-25 · Shengbin Guo, Shaokang He, Chaoyue Meng, Shengpeng Xiao 외 arxiv

While instruction-based image editing, enabled by multi-modal generative models, has advanced significantly, existing benchmarks lack a comprehensive evaluation of physics-based reasoning, a critical capability for handl…

Video GenerationImage Editing

Occlusion-Aware Physics-Semantic Keyframe Selection for Robust Video Editing

2026-05-22 · Lin Liu, Zhihan Xiao, Haohang Xu, Rong Cong 외 arxiv

Video editing has recently achieved remarkable progress with diffusion-based generative models, enabling diverse object-level manipulations from natural language instructions. However, existing methods often struggle und…

Beyond Rigid: Benchmarking Non-Rigid Video Editing

2026-01-26 · Bingzheng Qu, Xuefeng Bai, Kehai Chen, Min Zhang arxiv

As video generation models are increasingly expected to manipulate physical dynamics, there is a growing need to move evaluation beyond appearance fidelity and semantic alignment. Non-rigid video editing offers a uniquel…

Instruction FollowingVideo Generation