paper-with-me

홈 › Papers

PhyCustom: Towards Realistic Physical Customization in Text-to-Image Generation

2025-12-01 · Fan Wu, Cheng Chen, Zhoujie Fu, Jiacheng Wei, Yi Xu, Deheng Ye, Guosheng Lin arxiv

Recent diffusion-based text-to-image customization methods have achieved significant success in understanding concrete concepts to control generation processes, such as styles and shapes. However, few efforts dive into the realistic yet challenging customization of physical concepts. The core limitation of current methods arises from the absence of explicitly introducing physical knowledge during training. Even when physics-related words appear in the input text prompts, our experiments consistently demonstrate that these methods fail to accurately reflect the corresponding physical properties in the generated results. In this paper, we propose PhyCustom, a fine-tuning framework comprising two novel regularization losses to activate diffusion model to perform physical customization. Specifically, the proposed isometric loss aims at activating diffusion models to learn physical concepts while decouple loss helps to eliminate the mixture learning of independent concepts. Experiments are conducted on a diverse dataset and our benchmark results demonstrate that PhyCustom outperforms previous state-of-the-art and popular methods in terms of physical customization quantitatively and qualitatively.

📄 PDF Abstract BibTeX arXiv:2512.02794

Code (0)

등록된 구현이 없습니다.

Tasks

Text-to-Image Generation

Similar Papers 제목 키워드 기반

ThreeDWorld: A Platform for Interactive Multi-Modal Physical Simulation

2020-07-09 · Chuang Gan, Jeremy Schwartz, Seth Alter, Damian Mrowca 외

We introduce ThreeDWorld (TDW), a platform for interactive multi-modal physical simulation. TDW enables simulation of high-fidelity sensory data and physical interactions between mobile agents and objects in rich 3D envi…

Scene Understanding

Zero-Shot Personalization of Objects via Textual Inversion

2026-03-24 · Aniket Roy, Maitreya Suin, Rama Chellappa arxiv

Recent advances in text-to-image diffusion models have substantially improved the quality of image customization, enabling the synthesis of highly realistic images. Despite this progress, achieving fast and efficient per…

Personalized Image Generation

BlendScape: Enabling End-User Customization of Video-Conferencing Environments through Generative AI

2024-03-20 · Shwetha Rajaram, Nels Numan, Balasaravanan Thoravi Kumaravel, Nicolai Marquardt 외

Today's video-conferencing tools support a rich range of professional and social activities, but their generic meeting environments cannot be dynamically adapted to align with distributed collaborators' needs. To enable …

Image Generationmultimodal interaction

RealisID: Scale-Robust and Fine-Controllable Identity Customization via Local and Global Complementation

2024-12-22 · Zhaoyang Sun, Fei Du, Weihua Chen, Fan Wang 외

Recently, the success of text-to-image synthesis has greatly advanced the development of identity customization techniques, whose main goal is to produce realistic identity-specific photographs based on text prompts and …

Image Generation

F-Bench: Rethinking Human Preference Evaluation Metrics for Benchmarking Face Generation, Customization, and Restoration

2024-12-17 · Lu Liu, Huiyu Duan, Qiang Hu, Liu Yang 외

Artificial intelligence generative models exhibit remarkable capabilities in content creation, particularly in face image generation, customization, and restoration. However, current AI-generated faces (AIGFs) often fall…

BenchmarkingFace GenerationImage GenerationImage Quality Assessment