Audio-Visual Lip Synthesis via Intermediate Landmark Representation
β19May 16, 2023Updated 3 years ago
Alternatives and similar repositories for lip-synthesis
Users that are interested in lip-synthesis are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SyncTalkFace: Talking Face Generation for Precise Lip-syncing via Audio-Lip Memoryβ33Nov 3, 2022Updated 3 years ago
- Vid Driven Portrait Animation π€’π·β18Jul 7, 2024Updated 2 years ago
- Aligns faces to the canonical face in both videos and imagesβ17Apr 11, 2022Updated 4 years ago
- KAN-based Fusion of Dual Domain for Audio-Driven Landmarks Generation of the model can help you generate an sequence of facial lanmarks fβ¦β32Oct 28, 2025Updated 9 months ago
- β27Jun 19, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- β13Feb 19, 2022Updated 4 years ago
- Spliting the ASR probability distribution results into the chinese pinyin, so as to extract more effective feature for the chinese speechβ¦β21Mar 16, 2023Updated 3 years ago
- PersonaTalk Hackβ16Jan 10, 2025Updated last year
- Bunch of custom nodes; Segformer - allows you to segment images by specifying the segment in an arrayβ32Mar 28, 2025Updated last year
- Audio Entailment: Deductive Reasoning for Audio Understandingβ17Dec 10, 2024Updated last year
- BlendShapeMaker python3.6β45Jun 17, 2021Updated 5 years ago
- Project page for "Improving Few-shot Learning for Talking Face System with TTS Data Augmentation" for ICASSP2023β86Oct 10, 2023Updated 2 years ago
- β95Jun 23, 2021Updated 5 years ago
- β39Nov 10, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Styleβ15Aug 18, 2025Updated 11 months ago
- Official project repo for paper "Speech Driven Video Editing via an Audio-Conditioned Diffusion Model"β228Jun 30, 2023Updated 3 years ago
- Recognize speech from an audio file and convert it into animation FBXβ24Mar 7, 2022Updated 4 years ago
- optimized wav2lipβ18Jan 6, 2024Updated 2 years ago
- The pytorch implementation of our WACV23 paper "Cross-identity Video Motion Retargeting with Joint Transformation and Synthesis".β148Sep 8, 2023Updated 2 years ago
- Preprocessing Scipts for Talking Face Generationβ97Jan 21, 2025Updated last year
- This is a project about talking faces. We use 576X576 sized facial images for training, which can generate 2k, 4k, 6k, and 8k digital humβ¦β56Mar 18, 2024Updated 2 years ago
- PyTorch implementation of "StyleSync: High-Fidelity Generalized and Personalized Lip Sync in Style-based Generator"β215Aug 8, 2023Updated 2 years ago
- code for paper "Audio2Head: Audio-driven One-shot Talking-head Generation with Natural Head Motion" in the conference of IJCAI 2021β353Feb 15, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A pre-trained face parser based on SegNeXtβ51May 16, 2023Updated 3 years ago
- Cyberdolphin Suite of ComfyUI nodes for wiring up OpenAI and compatible LLM APIs.β15Jul 31, 2024Updated last year
- Grounded-SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and β¦β12Sep 5, 2023Updated 2 years ago
- Towards Efficient Audio-Visual Learners via Empowering Pre-trained Vision Transformers with Cross-Modal Adaptationβ15Apr 13, 2024Updated 2 years ago
- SAiD: Blendshape-based Audio-Driven Speech Animation with Diffusionβ135Jan 25, 2024Updated 2 years ago
- Automatic Facial Retargetingβ62Oct 22, 2020Updated 5 years ago
- Something about Talking Head Generationβ31Sep 5, 2023Updated 2 years ago
- Official Implementation of "MoDiTalker: Motion-Disentangled Diffusion Model for High-Fidelity Talking Head Generation" (AAAI 2025)β175Jan 14, 2025Updated last year
- One-shot Audio-driven 3D Talking Head Synthesis via Generative Prior, CVPRW 2024β65Oct 24, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official code of our ICCV2023 work: Implicit Identity Representation Conditioned Memory Compensation Network for Talking Head video Gβ¦β252Oct 5, 2023Updated 2 years ago
- Talking Face Generation systemβ17Oct 16, 2023Updated 2 years ago
- [ICLR 2026] Official implementation of "Enhancing Multi-Image Understanding Through Delimiter Token Scaling"β16Jul 10, 2026Updated 2 weeks ago
- [CVPR 2023] MetaPortrait: Identity-Preserving Talking Head Generation with Fast Personalized Adaptationβ542May 21, 2023Updated 3 years ago
- β60Jun 23, 2023Updated 3 years ago
- [ECCV 2024 Oral] EDTalk - Official PyTorch Implementationβ468Sep 29, 2025Updated 10 months ago
- Class-agnostic Object Detection and Instance Segmentation using Mask R-CNNβ17Nov 5, 2021Updated 4 years ago