paper-with-me

Papers

RenderMe-360: A Large Digital Asset Library and Benchmarks Towards High-fidelity Head Avatars

2023-05-22 · NeurIPS 2023 11 · Dongwei Pan, Long Zhuo, Jingtan Piao, Huiwen Luo, Wei Cheng, Yuxin Wang, Siming Fan, Shengqi Liu, Lei Yang, Bo Dai, Ziwei Liu, Chen Change Loy, Chen Qian, Wayne Wu, Dahua Lin, Kwan-Yee Lin

Synthesizing high-fidelity head avatars is a central problem for computer vision and graphics. While head avatar synthesis algorithms have advanced rapidly, the best ones still face great obstacles in real-world scenarios. One of the vital causes is inadequate datasets -- 1) current public datasets can only support researchers to explore high-fidelity head avatars in one or two task directions; 2) these datasets usually contain digital head assets with limited data volume, and narrow distribution over different attributes. In this paper, we present RenderMe-360, a comprehensive 4D human head dataset to drive advance in head avatar research. It contains massive data assets, with 243+ million complete head frames, and over 800k video sequences from 500 different identities captured by synchronized multi-view cameras at 30 FPS. It is a large-scale digital library for head avatars with three key attributes: 1) High Fidelity: all subjects are captured by 60 synchronized, high-resolution 2K cameras in 360 degrees. 2) High Diversity: The collected subjects vary from different ages, eras, ethnicities, and cultures, providing abundant materials with distinctive styles in appearance and geometry. Moreover, each subject is asked to perform various motions, such as expressions and head rotations, which further extend the richness of assets. 3) Rich Annotations: we provide annotations with different granularities: cameras' parameters, matting, scan, 2D/3D facial landmarks, FLAME fitting, and text description. Based on the dataset, we build a comprehensive benchmark for head avatar research, with 16 state-of-the-art methods performed on five main tasks: novel view synthesis, novel expression synthesis, hair rendering, hair editing, and talking head generation. Our experiments uncover the strengths and weaknesses of current methods. RenderMe-360 opens the door for future exploration in head avatars.

📄 PDF Abstract BibTeX arXiv:2305.13353

Code (1)

renderme-360/renderme-360 공식 구현 pytorch

Tasks

2kImage MattingNovel View SynthesisTalking Head Generation

Methods 이 논문이 사용한 방법론

Library 설명 없음

Similar Papers 제목 키워드 기반

From Physics-Based Models to Predictive Digital Twins via Interpretable Machine Learning

2020-04-23 · Michael G. Kapteyn, Karen E. Willcox

This work develops a methodology for creating a data-driven digital twin from a library of physics-based models representing various asset states. The digital twin is updated using interpretable machine learning. Specifi…

BIG-bench Machine LearningInterpretable Machine Learning

RenderMem: Rendering as Spatial Memory Retrieval

2026-03-15 · JooHyun Park, HyeongYeop Kang arxiv

Embodied reasoning is inherently viewpoint-dependent: what is visible, occluded, or reachable depends critically on where the agent stands. However, existing spatial memory systems for embodied agents typically store eit…

Spatial Reasoning

Turning Adaptation into Assets: Cross-Domain Bridging for Online Vision-Language Navigation

2026-05-22 · Zixuan Hu, Xuantuo Huang, Yancheng Li, Yichun Hu 외 arxiv

Navigating under non-stationary environment shifts poses a critical challenge for a Vision-and-Language Navigation (VLN) agent deployed in the wild. Yet, existing Test-Time Adaptation (TTA) methods for VLN largely treat …

Vision-Language NavigationTest-time Adaptation

Imaginarium: Vision-guided High-Quality 3D Scene Layout Generation

2025-10-17 · Xiaoming Zhu, Xu Huang, Qinghongbing Xie, Zhi Deng 외 arxiv

Generating artistic and coherent 3D scene layouts is crucial in digital content creation. Traditional optimization-based methods are often constrained by cumbersome manual rules, while deep generative models face challen…

Image Generation

A Library Perspective on Supervised Text Processing in Digital Libraries: An Investigation in the Biomedical Domain

2024-11-06 · Hermann Kroll, Pascal Sackhoff, Bill Matthias Thang, Maha Ksouri 외

Digital libraries that maintain extensive textual collections may want to further enrich their content for certain downstream applications, e.g., building knowledge graphs, semantic enrichment of documents, or implementi…

Knowledge GraphsRelation Extractiontext-classificationText Classification