[INTERSPEECH'24] Official repository for "MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset"
☆195Nov 5, 2024Updated last year
Alternatives and similar repositories for MultiTalk
Users that are interested in MultiTalk are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [INTERSPEECH'24] Official repository for "Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert"☆19Jun 25, 2025Updated last year
- DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer☆166Mar 31, 2024Updated 2 years ago
- [CVPR'25] Official repository for "Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Eva…☆49Jan 7, 2026Updated 6 months ago
- [CVPR 2024] FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models☆240Mar 17, 2024Updated 2 years ago
- Official code for ICCV 2023 paper: "Efficient Emotional Adaptation for Audio-Driven Talking-Head Generation".☆300Mar 4, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A novel apporach for personalized speech-driven 3D facial animation☆58Apr 26, 2024Updated 2 years ago
- DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models☆355Mar 11, 2025Updated last year
- ☆182Feb 15, 2024Updated 2 years ago
- This is the official source for our ACM MM 2023 paper "SelfTalk: A Self-Supervised Commutative Training Diagram to Comprehend 3D Talking …☆143Dec 5, 2023Updated 2 years ago
- ☆103Nov 26, 2025Updated 7 months ago
- [BMVC'25] Official repository for "Learning Correlation-aware Aleatoric Uncertainty for 3D Hand Pose Estimation"☆23Dec 8, 2025Updated 7 months ago
- This is the official source for our ICCV 2023 paper "EmoTalk: Speech-Driven Emotional Disentanglement for 3D Face Animation"☆420Feb 23, 2024Updated 2 years ago
- Official Pytorch Implementation of SPECTRE: Visual Speech-Aware Perceptual 3D Facial Expression Reconstruction from Videos☆301Mar 24, 2025Updated last year
- ☆183Jul 12, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆39Jan 30, 2026Updated 5 months ago
- LinguaLinker: Audio-Driven Portraits Animation with Implicit Facial Control Enhancement☆73Jul 29, 2024Updated last year
- ☆58Jul 9, 2025Updated last year
- [ICASSP 2024] Adaptive Super Resolution For One-Shot Talking-Head Generation☆182Mar 26, 2024Updated 2 years ago
- ARTalk generates realistic 3D head motions (lip sync, blinking, expressions, head poses) from audio in ⚡ real-time ⚡.☆136May 19, 2026Updated 2 months ago
- ☆239Sep 5, 2024Updated last year
- Official Pytorch Implementation of SMIRK: 3D Facial Expressions through Analysis-by-Neural-Synthesis (CVPR 2024)☆385May 31, 2024Updated 2 years ago
- [ECCV 2024 Oral] EDTalk - Official PyTorch Implementation☆468Sep 29, 2025Updated 9 months ago
- Official code release of "DEEPTalk: Dynamic Emotion Embedding for Probabilistic Speech-Driven 3D Face Animation" [AAAI2025]☆65Feb 13, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 🔥🔥🔥 Set the world of 3D faces on fire with INFERNO 🔥🔥🔥☆321Mar 1, 2026Updated 4 months ago
- [RA-L'24, IROS'24] Official PyTorch Implementation of "Uni-DVPS: Unified Model for Depth-Aware Video Panoptic Segmentation"☆13Oct 11, 2024Updated last year
- One-shot Audio-driven 3D Talking Head Synthesis via Generative Prior, CVPRW 2024☆65Oct 24, 2024Updated last year
- [NeurIPS'25] Automated Model Discovery via Multi-modal & Multi-step Pipeline☆22Dec 10, 2025Updated 7 months ago
- [NAACL'24] Repository for "SMILE: Multimodal Dataset for Understanding Laughter in Video with Language Models"☆15Jun 18, 2024Updated 2 years ago
- [IEEE TMM] Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation☆62May 8, 2026Updated 2 months ago
- [AAAI'24] Official PyTorch implementation of the paper "FPRF: Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radianc…☆18Nov 29, 2024Updated last year
- Pytorch official implementation for our paper "HyperLips: Hyper Control Lips with High Resolution Decoder for Talking Face Generation".☆212Mar 9, 2024Updated 2 years ago
- PyTorch implementation of "StyleSync: High-Fidelity Generalized and Personalized Lip Sync in Style-based Generator"☆215Aug 8, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- LIA-X: Interpretable Latent Portrait Animator☆105Sep 17, 2025Updated 10 months ago
- SAiD: Blendshape-based Audio-Driven Speech Animation with Diffusion☆135Jan 25, 2024Updated 2 years ago
- Official repository of Siggraph Asia 2025 paper "LSF-Animation: Label-Free Speech-Driven Facial Animation via Implicit Feature Representa…☆26Dec 24, 2025Updated 6 months ago
- ICCV 2025 ACTalker: an end-to-end video diffusion framework for talking head synthesis that supports both single and multi-signal control…☆461Aug 20, 2025Updated 11 months ago
- [ACM MM 2025] Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis☆841Nov 12, 2025Updated 8 months ago
- DICE-Talk is a diffusion-based emotional talking head generation method that can generate vivid and diverse emotions for speaking portrai…☆305Aug 7, 2025Updated 11 months ago
- [ECCV'24] TalkingGaussian: Structure-Persistent 3D Talking Head Synthesis via Gaussian Splatting☆385Mar 15, 2025Updated last year