paper-with-me

홈 › Papers

RITA: A Real-time Interactive Talking Avatars Framework

2024-06-18 · Wuxinlin Cheng, Cheng Wan, Yupeng Cao, Sihan Chen

RITA presents a high-quality real-time interactive framework built upon generative models, designed with practical applications in mind. Our framework enables the transformation of user-uploaded photos into digital avatars that can engage in real-time dialogue interactions. By leveraging the latest advancements in generative modeling, we have developed a versatile platform that not only enhances the user experience through dynamic conversational avatars but also opens new avenues for applications in virtual reality, online education, and interactive gaming. This work showcases the potential of integrating computer vision and natural language processing technologies to create immersive and interactive digital personas, pushing the boundaries of how we interact with digital content.

📄 PDF Abstract BibTeX arXiv:2406.13093

Code (1)

MindSpore-scientific/code-14/tree/main/RITA mindspore

Similar Papers 제목 키워드 기반

GaussianHeadTalk: Wobble-Free 3D Talking Heads with Audio Driven Gaussian Splatting

2025-12-11 · Madhav Agarwal, Mingtian Zhang, Laura Sevilla-Lara, Steven McDonagh arxiv

Speech-driven talking heads have recently emerged and enable interactive avatars. However, real-world applications are limited, as current methods achieve high visual fidelity but slow or fast yet temporally unstable. Di…

Image Generation

DEGAS: Detailed Expressions on Full-Body Gaussian Avatars

2024-08-20 · Zhijing Shao, Duotun Wang, Qing-Yao Tian, Yao-Dong Yang 외

Although neural rendering has made significant advances in creating lifelike, animatable full-body and head avatars, incorporating detailed expressions into full-body avatars remains largely unexplored. We present DEGAS,…

3DGSNeural Rendering

Comparative Analysis of Audio Feature Extraction for Real-Time Talking Portrait Synthesis

2024-11-20 · Pegah Salehi, Sajad Amouei Sheshkal, Vajira Thambawita, Sushant Gautam 외

This paper examines the integration of real-time talking-head generation for interviewer training, focusing on overcoming challenges in Audio Feature Extraction (AFE), which often introduces latency and limits responsive…

Talking Head Generation

StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars

2025-12-26 · Zhiyao Sun, Ziqiao Peng, Yifeng Ma, Yi Chen 외 arxiv

Real-time, streaming interactive avatars represent a critical yet challenging goal in digital human research. Although diffusion-based human avatar generation methods achieve remarkable success, their non-causal architec…

Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation

2026-01-02 · Taekyung Ki, Sangwon Jang, Jaehyeong Jo, Jaehong Yoon 외 arxiv

Talking head generation creates lifelike avatars from static portraits for virtual communication and content creation. However, current models do not yet convey the feeling of truly interactive communication, often gener…

Talking Head Generation