paper-with-me

홈 › Papers

Understanding the Limitations of Diffusion Concept Algebra Through Food

2024-06-05 · E. Zhixuan Zeng, Yuhao Chen, Alexander Wong

Image generation techniques, particularly latent diffusion models, have exploded in popularity in recent years. Many techniques have been developed to manipulate and clarify the semantic concepts these large-scale models learn, offering crucial insights into biases and concept relationships. However, these techniques are often only validated in conventional realms of human or animal faces and artistic style transitions. The food domain offers unique challenges through complex compositions and regional biases, which can shed light on the limitations and opportunities within existing methods. Through the lens of food imagery, we analyze both qualitative and quantitative patterns within a concept traversal technique. We reveal measurable insights into the model's ability to capture and represent the nuances of culinary diversity, while also identifying areas where the model's biases and limitations emerge.

📄 PDF Abstract BibTeX arXiv:2406.03582

Code (0)

등록된 구현이 없습니다.

Tasks

DiversityImage Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Concept Algebra for (Score-Based) Text-Controlled Generative Models

2023-02-07 · NeurIPS 2023 11 · ZiHao Wang, Lin Gui, Jeffrey Negrea, Victor Veitch

This paper concerns the structure of learned representations in text-guided generative models, focusing on score-based models. A key property of such models is that they can compose disparate concepts in a `disentangled'…

Hybrid Diffusion Policies with Projective Geometric Algebra for Efficient Robot Manipulation Learning

2025-07-08 · Xiatao Sun, Yuxuan Wang, Shuo Yang, Yinxing Chen 외 arxiv

Diffusion policies are a powerful paradigm for robot learning, but their training is often inefficient. A key reason is that networks must relearn fundamental spatial concepts, such as translations and rotations, from sc…

Robot Manipulation

LLM-based Cognitive Models of Students with Misconceptions

2024-10-16 · Shashank Sonkar, Xinghe Chen, Naiming Liu, Richard G. Baraniuk 외

Accurately modeling student cognition is crucial for developing effective AI-driven educational technologies. A key challenge is creating realistic student models that satisfy two essential properties: (1) accurately rep…

Misconceptions

NoReGeo: Non-Reasoning Geometry Benchmark

2026-01-15 · Irina Abdullaeva, Anton Vasiliuk, Elizaveta Goncharova, Temurbek Rahmatullaev 외 arxiv

We present NoReGeo, a novel benchmark designed to evaluate the intrinsic geometric understanding of large language models (LLMs) without relying on reasoning or algebraic computation. Unlike existing benchmarks that prim…

Binary Classification

Emergence and Evolution of Interpretable Concepts in Diffusion Models

2025-04-21 · Berk Tınaz, Zalan Fabian, Mahdi Soltanolkotabi

Diffusion models have become the go-to method for text-to-image generation, producing high-quality images from noise through a process called reverse diffusion. Understanding the dynamics of the reverse diffusion process…

Image GenerationText to Image GenerationText-to-Image Generation