paper-with-me

홈 › Papers

Image2Lego: Customized LEGO Set Generation from Images

2021-08-19 · Kyle Lennon, Katharina Fransen, Alexander O'Brien, Yumeng Cao, Matthew Beveridge, Yamin Arefeen, Nikhil Singh, Iddo Drori

Although LEGO sets have entertained generations of children and adults, the challenge of designing customized builds matching the complexity of real-world or imagined scenes remains too great for the average enthusiast. In order to make this feat possible, we implement a system that generates a LEGO brick model from 2D images. We design a novel solution to this problem that uses an octree-structured autoencoder trained on 3D voxelized models to obtain a feasible latent representation for model reconstruction, and a separate network trained to predict this latent representation from 2D images. LEGO models are obtained by algorithmic conversion of the 3D voxelized model to bricks. We demonstrate first-of-its-kind conversion of photographs to 3D LEGO models. An octree architecture enables the flexibility to produce multiple resolutions to best fit a user's creative vision or design needs. In order to demonstrate the broad applicability of our system, we generate step-by-step building instructions and animations for LEGO models of objects and human faces. Finally, we test these automatically generated LEGO sets by constructing physical builds using real LEGO bricks.

📄 PDF Abstract BibTeX arXiv:2108.08477

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Lego: Learning to Disentangle and Invert Personalized Concepts Beyond Object Appearance in Text-to-Image Diffusion Models

2023-11-23 · Saman Motamed, Danda Pani Paudel, Luc van Gool

Text-to-Image (T2I) models excel at synthesizing concepts such as nouns, appearances, and styles. To enable customized content creation based on a few example images of a concept, methods such as Textual Inversion and Dr…

Language ModellingLarge Language ModelQuestion AnsweringVisual Question Answering+1

Learning Stackable and Skippable LEGO Bricks for Efficient, Reconfigurable, and Variable-Resolution Diffusion Modeling

2023-10-10 · Huangjie Zheng, Zhendong Wang, Jianbo Yuan, Guanghan Ning 외

Diffusion models excel at generating photo-realistic images but come with significant computational costs in both training and sampling. While various techniques address these computational challenges, a less-explored is…

Image Generation

AttentionLego: An Open-Source Building Block For Spatially-Scalable Large Language Model Accelerator With Processing-In-Memory Technology

2024-01-21 · Rongqing Cong, Wenyang He, Mingxuan Li, Bangning Luo 외

Large language models (LLMs) with Transformer architectures have become phenomenal in natural language processing, multimodal generative artificial intelligence, and agent-oriented artificial intelligence. The self-atten…

Language ModelingLanguage ModellingLarge Language Model

LEGO: LoRA-Enabled Generator-Oriented Framework for Synthetic Image Detection

2026-05-06 · Yutong Xiao, Ran Ran, Jiwei Wei, Shuchang Zhou 외 arxiv

The rapid advancement of generative technologies has made synthetic images nearly indistinguishable from real ones, thereby creating an urgent need for robust detectors to counter misinformation. However, existing method…

UnrealEgo: A New Dataset for Robust Egocentric 3D Human Motion Capture

2022-08-02 · Hiroyasu Akada, Jian Wang, Soshi Shimada, Masaki Takahashi 외

We present UnrealEgo, i.e., a new large-scale naturalistic dataset for egocentric 3D human pose estimation. UnrealEgo is based on an advanced concept of eyeglasses equipped with two fisheye cameras that can be used in un…

3D Human Pose EstimationEgocentric Pose EstimationKeypoint EstimationPose Estimation