paper-with-me

Papers

ICo3D: An Interactive Conversational 3D Virtual Human

2026-01-19 · Richard Shaw, Youngkyoon Jang, Athanasios Papaioannou, Arthur Moreau, Helisa Dhamo, Zhensong Zhang, Eduardo Pérez-Pellitero arxiv

This work presents Interactive Conversational 3D Virtual Human (ICo3D), a method for generating an interactive, conversational, and photorealistic 3D human avatar. Based on multi-view captures of a subject, we create an animatable 3D face model and a dynamic 3D body model, both rendered by splatting Gaussian primitives. Once merged together, they represent a lifelike virtual human avatar suitable for real-time user interactions. We equip our avatar with an LLM for conversational ability. During conversation, the audio speech of the avatar is used as a driving signal to animate the face model, enabling precise synchronization. We describe improvements to our dynamic Gaussian models that enhance photorealism: SWinGS++ for body reconstruction and HeadGaS++ for face reconstruction, and provide as well a solution to merge the separate face and body models without artifacts. We also present a demo of the complete system, showcasing several use cases of real-time conversation with the 3D avatar. Our approach offers a fully integrated virtual avatar experience, supporting both oral and written form interactions in immersive environments. ICo3D is applicable to a wide range of fields, including gaming, virtual assistance, and personalized education, among others. Project page: https://ico3d.github.io/

📄 PDF Abstract BibTeX arXiv:2601.13148

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels

2026-04-11 · Yuzhe Weng, Haotian Wang, Xinyi Yu, Xiaoyan Wu 외 arxiv

Audio-driven human video generation has achieved remarkable success in monologue scenarios, largely driven by advancements in powerful video generation foundation models. Moving beyond monologues, authentic human communi…

Physical IntuitionVideo Generation

The influence of persona and conversational task on social interactions with a LLM-controlled embodied conversational agent

2024-11-08 · Leon O. H. Kroczek, Alexander May, Selina Hettenkofer, Andreas Ruider 외

Large Language Models (LLMs) have demonstrated remarkable capabilities in conversational tasks. Embodying an LLM as a virtual human allows users to engage in face-to-face social interactions in Virtual Reality. However, …

Situated and Interactive Multimodal Conversations

2020-06-02 · COLING 2020 8 · Seungwhan Moon, Satwik Kottur, Paul A. Crook, Ankita De 외

Next generation virtual assistants are envisioned to handle multimodal inputs (e.g., vision, memories of previous interactions, in addition to the user's utterances), and perform multimodal actions (e.g., displaying a ro…

Response Generation

9th Workshop on Sign Language Translation and Avatar Technologies (SLTAT 2025)

2025-08-11 · Fabrizio Nunnari, Cristina Luna Jiménez, Rosalee Wolfe, John C. McDonald 외 arxiv

The Sign Language Translation and Avatar Technology (SLTAT) workshops continue a series of gatherings to share recent advances in improving deaf / human communication through non-invasive means. This 2025 edition, the 9t…

Sign Language TranslationSign Language Recognition

Conversational Human Audio-visual Talking Dialogue Generation

2026-07-02 · Junhao Song, Lluis Guasch, Xilin He, Zhongyu Yang 외 arxiv

Large-scale dyadic interactive audio-visual dialogue (DIAD) datasets provide fundamental data resources for developing humanoid interactive virtual agents and digital humans. However, collecting such data is time-consumi…

Dialogue Generation