☆350Jan 2, 2026Updated 8 months ago
Alternatives and similar repositories for LiveTalk
Users that are interested in LiveTalk are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2026 Spotlight] Implementation of "Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length"☆2,439Aug 24, 2026Updated last month
- [CVPR 2026] Official Pytorch implementation of Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation☆353Jun 3, 2026Updated 3 months ago
- [CVPR2026]We present FlashPortrait, an end-to-end video diffusion transformer capable of synthesizing ID-preserving, infinite-length vide…☆481Feb 21, 2026Updated 7 months ago
- SoulX-FlashTalk is the first 14B model to achieve sub-second start-up latency (0.87s) while maintaining a real-time throughput of 32 FPS …☆1,525Jul 30, 2026Updated last month
- BeHonest: Benchmarking Honesty in Large Language Models☆36Aug 15, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A real-time streaming conversational video system that transforms text interactions into continuous, high-fidelity video responses using …☆343Dec 15, 2025Updated 9 months ago
- SoulX-FlashHead: A unified 1.3B-parameter framework designed for high-fidelity, infinite-length, and real-time streaming portrait video g…☆1,084May 28, 2026Updated 3 months ago
- ACM MM | IMTalker: Efficient Audio-driven Talking Face Generation with Implicit Motion Transfer☆197Dec 23, 2025Updated 9 months ago
- [ACM MM 2025] Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis☆894Nov 12, 2025Updated 10 months ago
- DICE-Talk is a diffusion-based emotional talking head generation method that can generate vivid and diverse emotions for speaking portrai…☆307Aug 7, 2025Updated last year
- ☆1,862Aug 6, 2025Updated last year
- [CVPR 2026 Highlight] Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation☆363Dec 15, 2025Updated 9 months ago
- Long Video Gen Infrastructure☆2,640Sep 7, 2026Updated 2 weeks ago
- ☆33Jan 30, 2026Updated 7 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- A foundation model that generates synchronized video and audio in a single model☆1,117Updated this week
- [NeurIPS'25 Spotlight] Official implementation of "JavisGPT: A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation"☆76Feb 26, 2026Updated 6 months ago
- We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven av…☆1,259Jan 20, 2026Updated 8 months ago
- [SIGGRAPH 2025] LAM: Large Avatar Model for One-shot Animatable Gaussian Head☆1,066Jun 10, 2026Updated 3 months ago
- ☆42Feb 7, 2026Updated 7 months ago
- Helios: Real Real-Time Long Video Generation Model☆2,158Aug 24, 2026Updated last month
- [CVPR 2026] PersonaLive! : Expressive Portrait Image Animation for Live Streaming☆3,827Aug 28, 2026Updated 3 weeks ago
- ☆2,117Apr 11, 2026Updated 5 months ago
- [ECCV 2026] ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling☆181Sep 16, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- LLIA - Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models☆152Jun 11, 2025Updated last year
- [AAAI 2026] EchoMimicV3: 1.3B Parameters are All You Need for Unified Multi-Modal and Multi-Task Human Animation☆1,063Mar 18, 2026Updated 6 months ago
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactiv…☆974Updated this week
- Official Pytorch implementation of AvatarForcing: One-Step Streaming Talking Avatars via Local-Future Sliding-Window Denoising☆81May 9, 2026Updated 4 months ago
- A 2D customized lip-sync model for high-fidelity real-time driving.☆139Jun 26, 2025Updated last year
- [ECCV 2026 Oral] Official implementation of "OmniForcing: Unleashing Real-time Joint Audio-Visual Generation"[arXiv:2603.11647]. OmniForc…☆195Jul 23, 2026Updated 2 months ago
- ☆32Mar 15, 2026Updated 6 months ago
- AnyTalker: Scaling Multi-person Talking Video Generation with Interactivity Refinement☆326Apr 15, 2026Updated 5 months ago
- ☆2,175Dec 16, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICLR2026] The open-source code for FlowCache, including accelerated implementations of the MAGI-1 and Skyreels-V2.☆34Apr 24, 2026Updated 5 months ago
- The official code of Yume☆684Jan 14, 2026Updated 8 months ago
- The official SpeakerVid-5M data curation code.☆88Jul 23, 2025Updated last year
- TurboDiffusion: 100–200× Acceleration for Video Diffusion Models☆3,820Aug 27, 2026Updated 3 weeks ago
- SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representations (CVPR 2026 Findings)☆1,053May 6, 2026Updated 4 months ago
- A unified inference and post-training framework for accelerated video generation.☆4,496Updated this week
- ☆364Feb 9, 2026Updated 7 months ago