paper-with-me

Papers

Unified framework for diffusion generative models in SO(3): applications in computer vision and astrophysics

2023-12-18 · Yesukhei Jagvaral, Francois Lanusse, Rachel Mandelbaum

Diffusion-based generative models represent the current state-of-the-art for image generation. However, standard diffusion models are based on Euclidean geometry and do not translate directly to manifold-valued data. In this work, we develop extensions of both score-based generative models (SGMs) and Denoising Diffusion Probabilistic Models (DDPMs) to the Lie group of 3D rotations, SO(3). SO(3) is of particular interest in many disciplines such as robotics, biochemistry and astronomy/cosmology science. Contrary to more general Riemannian manifolds, SO(3) admits a tractable solution to heat diffusion, and allows us to implement efficient training of diffusion models. We apply both SO(3) DDPMs and SGMs to synthetic densities on SO(3) and demonstrate state-of-the-art results. Additionally, we demonstrate the practicality of our model on pose estimation tasks and in predicting correlated galaxy orientations for astrophysics/cosmology.

📄 PDF Abstract BibTeX arXiv:2312.11707

Code (0)

등록된 구현이 없습니다.

Tasks

AstronomyDenoisingImage GenerationPose Estimation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

VillanDiffusion: A Unified Backdoor Attack Framework for Diffusion Models

2023-06-12 · NeurIPS 2023 11 · Sheng-Yen Chou, Pin-Yu Chen, Tsung-Yi Ho

Diffusion Models (DMs) are state-of-the-art generative models that learn a reversible corruption process from iterative noise addition and denoising. They are the backbone of many generative AI applications, such as text…

Backdoor AttackDenoising

MarkDiffusion: An Open-Source Toolkit for Generative Watermarking of Latent Diffusion Models

2025-09-11 · Leyi Pan, Sheng Guan, Zheyu Fu, Luyang Si 외 arxiv

We introduce MarkDiffusion, an open-source Python toolkit for generative watermarking of latent diffusion models. It comprises three key components: a unified implementation framework for streamlined watermarking algorit…

LaVin-DiT: Large Vision Diffusion Transformer

2024-11-18 · CVPR 2025 1 · Zhaoqing Wang, Xiaobo Xia, Runnan Chen, Dongdong Yu 외

This paper presents the Large Vision Diffusion Transformer (LaVin-DiT), a scalable and unified foundation model designed to tackle over 20 computer vision tasks in a generative framework. Unlike existing large vision mod…

In-Context Learning

MMGen: Unified Multi-modal Image Generation and Understanding in One Go

2025-03-26 · Jiepeng Wang, Zhaoqing Wang, Hao Pan, YuAn Liu 외

A unified diffusion framework for multi-modal generation and understanding has the transformative potential to achieve seamless and controllable image diffusion and other cross-modal tasks. In this paper, we introduce MM…

Image Generation

UniGEM: A Unified Approach to Generation and Property Prediction for Molecules

2024-10-14 · Shikun Feng, Yuyan Ni, Yan Lu, Zhi-Ming Ma 외

Molecular generation and molecular property prediction are both crucial for drug discovery, but they are often developed independently. Inspired by recent studies, which demonstrate that diffusion model, a prominent gene…

Drug DiscoveryMolecular Property PredictionMulti-Task LearningPrediction+1