AVTR-1: Avatars that listen back
☆456May 26, 2026Updated 3 months ago
Alternatives and similar repositories for avtr-1
Users that are interested in avtr-1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACM MM 2025] Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis☆877Nov 12, 2025Updated 9 months ago
- Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆228Jun 30, 2026Updated 2 months ago
- [CVPR 2026] Official Pytorch implementation of Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation☆349Jun 3, 2026Updated 3 months ago
- The official code implementation of the paper “Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration“.☆112May 24, 2026Updated 3 months ago
- ACM MM | IMTalker: Efficient Audio-driven Talking Face Generation with Implicit Motion Transfer☆193Dec 23, 2025Updated 8 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [ECCV 2026 Spotlight] Implementation of "Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length"☆2,403Aug 24, 2026Updated last week
- SoulX-FlashHead: A unified 1.3B-parameter framework designed for high-fidelity, infinite-length, and real-time streaming portrait video g…☆1,037May 28, 2026Updated 3 months ago
- Real-time stream editing pipeline powered by the FLUX.2-klein-4B model, optimized for consumer GPUs☆443Jun 13, 2026Updated 2 months ago
- [ICML 2026] ByteDance's All-in-One Video Generation Model for Human-Object Interaction Video Generation☆469Jul 29, 2026Updated last month
- ☆7,752May 27, 2026Updated 3 months ago
- [CVPR 2026] PersonaLive! : Expressive Portrait Image Animation for Live Streaming☆3,626Updated this week
- TalkingMachines☆178Aug 2, 2025Updated last year
- [CVPR2026]We present FlashPortrait, an end-to-end video diffusion transformer capable of synthesizing ID-preserving, infinite-length vide…☆481Feb 21, 2026Updated 6 months ago
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆351Jun 24, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Connect to Gemini powered Pipecat bots using the Voice UI Kit and screen shares☆18Jan 27, 2026Updated 7 months ago
- ☆346Jan 2, 2026Updated 8 months ago
- A real-time streaming conversational video system that transforms text interactions into continuous, high-fidelity video responses using …☆341Dec 15, 2025Updated 8 months ago
- Take control back!☆18Jun 22, 2022Updated 4 years ago
- ☆22Aug 21, 2025Updated last year
- Zero-shot expressive voice cloning and speech generation. Generate anything from short clips to full-length audiobooks with realistic emo…☆551Jul 7, 2026Updated last month
- Lean neural real-time acoustic echo cancellation with soft delay estimation - GGML and PyTorch inference☆209Jul 20, 2026Updated last month
- Gaussian Wardrobe code release☆33Jul 5, 2026Updated last month
- StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars☆21Mar 31, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code Release for "OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data"☆81Aug 4, 2026Updated last month
- super expressive prompting model based on ltx2.3☆484May 23, 2026Updated 3 months ago
- Open-source AI that watches and hears your screen and reacts live as any personality you design — for faceless content, autonomous AI str…☆36Jul 7, 2026Updated last month
- An open-source model family for long-form speech, dialogue synthesis, voice design, sound effects, and real-time streaming TTS☆4,061Updated this week
- Legible, Scalable, Reproducible Foundation Models with Named Tensors and Jax☆16Jun 16, 2024Updated 2 years ago
- LiveTalk is a unified, high-performance talking head generation system that combines the power of LivePortrait and MuseTalk open-source r…☆46Updated this week
- [AAAI 2026] EchoMimicV3: 1.3B Parameters are All You Need for Unified Multi-Modal and Multi-Task Human Animation☆1,037Mar 18, 2026Updated 5 months ago
- FineMotion: A Dataset and Benchmark with both Spatial and Temporal Annotation for Fine-grained Motion Generation and Editing☆17Mar 4, 2025Updated last year
- Official Implementation of SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning☆1,166Aug 24, 2026Updated last week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation☆168Jun 26, 2026Updated 2 months ago
- ☆24Jan 31, 2024Updated 2 years ago
- Long Video Gen Infrastructure☆2,593Aug 7, 2026Updated 3 weeks ago
- Ideogram 4: Open image model at the forefront of design☆2,800Jun 30, 2026Updated 2 months ago
- Official Implementation of CoInteract: Spatially-Structured Co-Generation for Interactive Human-Object Video Synthesis☆167May 7, 2026Updated 3 months ago
- [CVPR-2025] The official code of HunyuanPortrait: Implicit Condition Control for Enhanced Portrait Animation☆345Dec 7, 2025Updated 8 months ago
- ICCV 2025 ACTalker: an end-to-end video diffusion framework for talking head synthesis that supports both single and multi-signal control…☆463Aug 20, 2025Updated last year