paper-with-me

VOCASET

홈페이지 · 논문 60편

VOCASET is a 4D face dataset with about 29 minutes of 4D scans captured at 60 fps and synchronized audio. The dataset has 12 subjects and 480 sequences of about 3-4 seconds each with sentences chosen from an array of standard protocols that maximize phonetic diversity. Source: timzhang642

3DSpeech

벤치마크

3D Face Animation on VOCASET 결과 14개