paper-with-me

홈 › Papers

Masked Conditioning for Deep Generative Models

2025-05-22 · Phillip Mueller, Jannik Wiese, Sebastian Mueller, Lars Mikelsons

Datasets in engineering domains are often small, sparsely labeled, and contain numerical as well as categorical conditions. Additionally. computational resources are typically limited in practical applications which hinders the adoption of generative models for engineering tasks. We introduce a novel masked-conditioning approach, that enables generative models to work with sparse, mixed-type data. We mask conditions during training to simulate sparse conditions at inference time. For this purpose, we explore the use of various sparsity schedules that show different strengths and weaknesses. In addition, we introduce a flexible embedding that deals with categorical as well as numerical conditions. We integrate our method into an efficient variational autoencoder as well as a latent diffusion model and demonstrate the applicability of our approach on two engineering-related datasets of 2D point clouds and images. Finally, we show that small models trained on limited data can be coupled with large pretrained foundation models to improve generation quality while retaining the controllability induced by our conditioning scheme.

📄 PDF Abstract BibTeX arXiv:2505.16725

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Latent Diffusion Model Diffusion models applied to latent spaces, which are normally built with (Variational) Autoencoders.

Similar Papers 제목 키워드 기반

MaskSketch: Unpaired Structure-guided Masked Image Generation

2023-02-10 · CVPR 2023 1 · Dina Bashkirova, Jose Lezama, Kihyuk Sohn, Kate Saenko 외

Recent conditional image generation methods produce images of remarkable diversity, fidelity and realism. However, the majority of these methods allow conditioning only on labels or text prompts, which limits their level…

Conditional Image GenerationDiversityImage GenerationImage-to-Image Translation+2

Simple Self-Conditioning Adaptation for Masked Diffusion Models

2026-04-28 · Michael Cardei, Huu Binh Ta, Ferdinando Fioretto arxiv

Masked diffusion models (MDMs) generate discrete sequences by iterative denoising under an absorbing masking process. In standard masked diffusion, if a token remains masked after a reverse update, the model discards its…

CM3: A Causal Masked Multimodal Model of the Internet

2022-01-19 · Armen Aghajanyan, Bernie Huang, Candace Ross, Vladimir Karpukhin 외

We introduce CM3, a family of causally masked generative models trained over a large corpus of structured multi-modal documents that can contain both text and image tokens. Our new causally masked approach generates toke…

ArticlesEntity DisambiguationEntity Linking

Masked Generative Video-to-Audio Transformers with Enhanced Synchronicity

2024-07-15 · Santiago Pascual, Chunghsin Yeh, Ioannis Tsiamas, Joan Serrà

Video-to-audio (V2A) generation leverages visual-only video features to render plausible sounds that match the scene. Importantly, the generated sound onsets should match the visual actions that are aligned with them, ot…

Video-to-Sound Generation

What Matters in Virtual Try-Off? Dual-UNet Diffusion Model For Garment Reconstruction

2026-04-09 · Loc-Phat Truong, Meysam Madadi, Sergio Escalera arxiv

Virtual Try-On (VTON) has seen rapid advancements, providing a strong foundation for generative fashion tasks. However, the inverse problem, Virtual Try-Off (VTOFF)-aimed at reconstructing the canonical garment from a dr…

Garment ReconstructionVirtual Try-OffVirtual Try-on