AVTR-1: Avatars that listen back
☆387May 26, 2026Updated last month
Alternatives and similar repositories for avtr-1
Users that are interested in avtr-1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACM MM 2025] Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis☆842Nov 12, 2025Updated 8 months ago
- Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆214Jun 30, 2026Updated 3 weeks ago
- [CVPR 2026] Official Pytorch implementation of Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation☆333Jun 3, 2026Updated last month
- The official code implementation of the paper “Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration“.☆107May 24, 2026Updated 2 months ago
- ACM MM | IMTalker: Efficient Audio-driven Talking Face Generation with Implicit Motion Transfer☆186Dec 23, 2025Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ECCV 2026 Oral] Implementation of "Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length"☆2,272Updated this week
- SoulX-FlashHead: A unified 1.3B-parameter framework designed for high-fidelity, infinite-length, and real-time streaming portrait video g…☆920May 28, 2026Updated last month
- Real-time stream editing pipeline powered by the FLUX.2-klein-4B model, optimized for consumer GPUs☆420Jun 13, 2026Updated last month
- ☆5,292May 27, 2026Updated last month
- [ICML 2026] ByteDance's All-in-One Video Generation Model for Human-Object Interaction Video Generation☆459May 19, 2026Updated 2 months ago
- [CVPR 2026] PersonaLive! : Expressive Portrait Image Animation for Live Streaming☆3,429May 15, 2026Updated 2 months ago
- TalkingMachines☆178Aug 2, 2025Updated 11 months ago
- [CVPR2026]We present FlashPortrait, an end-to-end video diffusion transformer capable of synthesizing ID-preserving, infinite-length vide…☆479Feb 21, 2026Updated 5 months ago
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆347Jun 24, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆328Jan 2, 2026Updated 6 months ago
- A real-time streaming conversational video system that transforms text interactions into continuous, high-fidelity video responses using …☆334Dec 15, 2025Updated 7 months ago
- Take control back!☆18Jun 22, 2022Updated 4 years ago
- ☆22Aug 21, 2025Updated 11 months ago
- Zero-shot expressive voice cloning and speech generation. Generate anything from short clips to full-length audiobooks with realistic emo…☆536Jul 7, 2026Updated 2 weeks ago
- Lean neural real-time acoustic echo cancellation with soft delay estimation - GGML and PyTorch inference☆182Updated this week
- Gaussian Wardrobe code release☆27Jul 5, 2026Updated 2 weeks ago
- StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars☆18Mar 31, 2026Updated 3 months ago
- Code Release for "OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data"☆76Jun 12, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- super expressive prompting model based on ltx2.3☆469May 23, 2026Updated 2 months ago
- Open-source AI that watches and hears your screen and reacts live as any personality you design — for faceless content, autonomous AI str…☆22Jul 7, 2026Updated 2 weeks ago
- MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fi…☆3,893Jun 22, 2026Updated last month
- Legible, Scalable, Reproducible Foundation Models with Named Tensors and Jax☆16Jun 16, 2024Updated 2 years ago
- LiveTalk is a unified, high-performance talking head generation system that combines the power of LivePortrait and MuseTalk open-source r…☆44Jan 15, 2026Updated 6 months ago
- [AAAI 2026] EchoMimicV3: 1.3B Parameters are All You Need for Unified Multi-Modal and Multi-Task Human Animation☆992Mar 18, 2026Updated 4 months ago
- FineMotion: A Dataset and Benchmark with both Spatial and Temporal Annotation for Fine-grained Motion Generation and Editing☆16Mar 4, 2025Updated last year
- Official Implementation of SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning☆998Jul 16, 2026Updated last week
- DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation☆162Jun 26, 2026Updated 3 weeks ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆24Jan 31, 2024Updated 2 years ago
- Long Video Gen Infrastructure☆2,491Jul 15, 2026Updated last week
- Ideogram 4: Open image model at the forefront of design☆2,583Jun 30, 2026Updated 3 weeks ago
- Official Implementation of CoInteract: Spatially-Structured Co-Generation for Interactive Human-Object Video Synthesis☆163May 7, 2026Updated 2 months ago
- [CVPR-2025] The official code of HunyuanPortrait: Implicit Condition Control for Enhanced Portrait Animation☆345Dec 7, 2025Updated 7 months ago
- ICCV 2025 ACTalker: an end-to-end video diffusion framework for talking head synthesis that supports both single and multi-signal control…☆461Aug 20, 2025Updated 11 months ago
- [CVPR 2026 Poster] One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer☆490Apr 19, 2026Updated 3 months ago