paper-with-me

홈 › Papers

HDTR-Net: A Real-Time High-Definition Teeth Restoration Network for Arbitrary Talking Face Generation Methods

2023-09-14 · Yongyuan Li, Xiuyuan Qin, Chao Liang, Mingqiang Wei

Talking Face Generation (TFG) aims to reconstruct facial movements to achieve high natural lip movements from audio and facial features that are under potential connections. Existing TFG methods have made significant advancements to produce natural and realistic images. However, most work rarely takes visual quality into consideration. It is challenging to ensure lip synchronization while avoiding visual quality degradation in cross-modal generation methods. To address this issue, we propose a universal High-Definition Teeth Restoration Network, dubbed HDTR-Net, for arbitrary TFG methods. HDTR-Net can enhance teeth regions at an extremely fast speed while maintaining synchronization, and temporal consistency. In particular, we propose a Fine-Grained Feature Fusion (FGFF) module to effectively capture fine texture feature information around teeth and surrounding regions, and use these features to fine-grain the feature map to enhance the clarity of teeth. Extensive experiments show that our method can be adapted to arbitrary TFG methods without suffering from lip synchronization and frame coherence. Another advantage of HDTR-Net is its real-time generation ability. Also under the condition of high-definition restoration of talking face video synthesis, its inference speed is $300\%$ faster than the current state-of-the-art face restoration based on super-resolution.

📄 PDF Abstract BibTeX arXiv:2309.07495

Code (1)

yylgoodlucky/hdtr 공식 구현 pytorch

Tasks

Face GenerationSuper-ResolutionTalking Face Generation

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

BiHDTrans: binary hyperdimensional transformer for efficient multivariate time series classification

2025-09-29 · Jingtao Zhang, Yi Liu, Qi Shen, Changhong Wang arxiv

The proliferation of Internet-of-Things (IoT) devices has led to an unprecedented volume of multivariate time series (MTS) data, requiring efficient and accurate processing for timely decision-making in resource-constrai…

Time Series Classification

Processing and Segmentation of Human Teeth from 2D Images using Weakly Supervised Learning

2023-11-13 · Tomáš Kunzo, Viktor Kocur, Lukáš Gajdošech, Martin Madaras

Teeth segmentation is an essential task in dental image analysis for accurate diagnosis and treatment planning. While supervised deep learning methods can be utilized for teeth segmentation, they often require extensive …

Keypoint DetectionSegmentationWeakly-supervised Learning

TeethGenerator: A two-stage framework for paired pre- and post-orthodontic 3D dental data generation

2025-07-07 · Changsong Lei, Yaqian Liang, Shaofeng Wang, Jiajia Dai 외 arxiv

Digital orthodontics represents a prominent and critical application of computer vision technology in the medical field. So far, the labor-intensive process of collecting clinical data, particularly in acquiring paired 3…

3DTeethSAM: Taming SAM2 for 3D Teeth Segmentation

2025-12-12 · Zhiguo Lu, Jianwen Lou, Mingjun Ma, Hairong Jin 외 arxiv

3D teeth segmentation, involving the localization of tooth instances and their semantic categorization in 3D dental models, is a critical yet challenging task in digital dentistry due to the complexity of real-world dent…

Video Segmentation

Accurate 3D Prediction of Missing Teeth in Diverse Patterns for Precise Dental Implant Planning

2023-07-16 · Lei Ma, Peng Xue, Yuning Gu, Yue Zhao 외

In recent years, the demand for dental implants has surged, driven by their high success rates and esthetic advantages. However, accurate prediction of missing teeth for precise digital implant planning remains a challen…

Prediction