paper-with-me

Papers

StochSync: Stochastic Diffusion Synchronization for Image Generation in Arbitrary Spaces

2025-01-26 · Kyeongmin Yeo, Jaihoon Kim, Minhyuk Sung

We propose a zero-shot method for generating images in arbitrary spaces (e.g., a sphere for 360{\deg} panoramas and a mesh surface for texture) using a pretrained image diffusion model. The zero-shot generation of various visual content using a pretrained image diffusion model has been explored mainly in two directions. First, Diffusion Synchronization-performing reverse diffusion processes jointly across different projected spaces while synchronizing them in the target space-generates high-quality outputs when enough conditioning is provided, but it struggles in its absence. Second, Score Distillation Sampling-gradually updating the target space data through gradient descent-results in better coherence but often lacks detail. In this paper, we reveal for the first time the interconnection between these two methods while highlighting their differences. To this end, we propose StochSync, a novel approach that combines the strengths of both, enabling effective performance with weak conditioning. Our experiments demonstrate that StochSync provides the best performance in 360{\deg} panorama generation (where image conditioning is not given), outperforming previous finetuning-based methods, and also delivers comparable results in 3D mesh texturing (where depth conditioning is provided) with previous methods.

📄 PDF Abstract BibTeX arXiv:2501.15445

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Kuramoto Orientation Diffusion Models

2025-09-18 · Yue Song, T. Anderson Keller, Sevan Brodjian, Takeru Miyato 외 arxiv

Orientation-rich images, such as fingerprints and textures, often exhibit coherent angular directional patterns that are challenging to model using standard generative approaches based on isotropic Euclidean diffusion. M…

Image Generation

Variational Test-time Optimization for Diffusion Synchronization

2026-06-14 · Hyunsoo Lee, Farrin Marouf Sofian, Kushagra Pandey, Stephan Mandt arxiv

Collaborative generation, which coordinates multiple diffusion trajectories to extend the capabilities of pretrained priors, has emerged as a powerful paradigm for extending the applicability of diffusion models. Among e…

JavisDiT: Joint Audio-Video Diffusion Transformer with Hierarchical Spatio-Temporal Prior Synchronization

2025-03-30 · Kai Liu, Wei Li, Lai Chen, Shengqiong Wu 외

This paper introduces JavisDiT, a novel Joint Audio-Video Diffusion Transformer designed for synchronized audio-video generation (JAVG). Built upon the powerful Diffusion Transformer (DiT) architecture, JavisDiT is able …

Video Generation

Paris: A Decentralized Trained Open-Weight Diffusion Model

2025-10-03 · Zhiying Jiang, Raihan Seraj, Marcos Villagra, Bidhan Roy arxiv

We present Paris, the first publicly released diffusion model pre-trained entirely through decentralized computation. Paris demonstrates that high-quality text-to-image generation can be achieved without centrally coordi…

Text-to-Image Generation

DreamCube: 3D Panorama Generation via Multi-plane Synchronization

2025-06-20 · Yukun Huang, Yanning Zhou, Jianan Wang, Kaiyi Huang 외

3D panorama synthesis is a promising yet challenging task that demands high-quality and diverse visual appearance and geometry of the generated omnidirectional content. Existing methods leverage rich image priors from pr…

Depth EstimationImage GenerationScene Generation