paper-with-me

Papers

Hitem3D 2.0: Multi-View Guided Native 3D Texture Generation

2026-04-10 · Huiang He, Shengchu Zhao, Jianwen Huang, Jie Li, Jiaqi Wu, Hu Zhang, Pei Tang, Heliang Zheng, Yukun Li, Rongfei Jia arxiv

Although recent advances have improved the quality of 3D texture generation, existing methods still struggle with incomplete texture coverage, cross-view inconsistency, and misalignment between geometry and texture. To address these limitations, we propose Hitem3D 2.0, a multi-view guided native 3D texture generation framework that enhances texture quality through the integration of 2D multi-view generation priors and native 3D texture representations. Hitem3D 2.0 comprises two key components: a multi-view synthesis framework and a native 3D texture generation model. The multi-view generation is built upon a pre-trained image editing backbone and incorporates plug-and-play modules that explicitly promote geometric alignment, cross-view consistency, and illumination uniformity, thereby enabling the synthesis of high-fidelity multi-view images. Conditioned on the generated views and 3D geometry, the native 3D texture generation model projects multi-view textures onto 3D surfaces while plausibly completing textures in unseen regions. Through the integration of multi-view consistency constraints with native 3D texture modeling, Hitem3D 2.0 significantly improves texture completeness, cross-view coherence, and geometric alignment. Experimental results demonstrate that Hitem3D 2.0 outperforms existing methods in terms of texture detail, fidelity, consistency, coherence, and alignment.

📄 PDF Abstract BibTeX arXiv:2604.09231

Code (0)

등록된 구현이 없습니다.

Tasks

Image Editing

Similar Papers 제목 키워드 기반

Unveiling the Cognitive Compass: Theory-of-Mind-Guided Multimodal Emotion Reasoning

2026-02-01 · Meng Luo, Bobo Li, Shanqing Xu, Shize Zhang 외 arxiv

Despite rapid progress in multimodal large language models (MLLMs), their capability for deep emotional understanding remains limited. We argue that genuine affective intelligence requires explicit modeling of Theory of …

Reinforcement Learning

WhiteMatter: All-to-All Cross-Layer Connections via KV Mixing

2026-08-19 · Wenbo Zhang, Xiang Ren arxiv

In a Transformer, each layer attends to past tokens only through KV produced at its own depth, despite the presence of deeper representations during autoregressive decoding. Feedback architectures allow shallow consumer …

GarmentDreamer: 3DGS Guided Garment Synthesis with Diverse Geometry and Texture Details

2024-05-20 · Boqian Li, Xuan Li, Ying Jiang, Tianyi Xie 외

Traditional 3D garment creation is labor-intensive, involving sketching, modeling, UV mapping, and texturing, which are time-consuming and costly. Recent advances in diffusion-based generative models have enabled new pos…

3D Generation3D Geometry Prediction3DGSGarment Reconstruction+2

TexGen: Text-Guided 3D Texture Generation with Multi-view Sampling and Resampling

2024-08-02 · Dong Huo, Zixin Guo, Xinxin Zuo, Zhihao Shi 외

Given a 3D mesh, we aim to synthesize 3D textures that correspond to arbitrary textual descriptions. Current methods for generating and assembling textures from sampled views often result in prominent seams or excessive …

DenoisingTexture Synthesis

Multi-Scale Geometric Consistency Guided Multi-View Stereo

2019-04-17 · CVPR 2019 6 · Qingshan Xu, Wenbing Tao

In this paper, we propose an efficient multi-scale geometric consistency guided multi-view stereo method for accurate and complete depth map estimation. We first present our basic multi-view stereo method with Adaptive C…

Depth EstimationMulti-View 3D ReconstructionPoint Clouds