AVTR-1: Avatars that listen back
☆482May 26, 2026Updated 3 months ago
Alternatives and similar repositories for avtr-1
Users that are interested in avtr-1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACM MM 2025] Ditto: Motion-Space Diffusion for Controllable Realtime Talking Head Synthesis☆894Nov 12, 2025Updated 10 months ago
- Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆225Jun 30, 2026Updated 2 months ago
- [CVPR 2026] Official Pytorch implementation of Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation☆353Jun 3, 2026Updated 3 months ago
- The official code implementation of the paper “Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration“.☆114May 24, 2026Updated 4 months ago
- ACM MM | IMTalker: Efficient Audio-driven Talking Face Generation with Implicit Motion Transfer☆197Dec 23, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ECCV 2026 Spotlight] Implementation of "Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length"☆2,439Aug 24, 2026Updated last month
- SoulX-FlashHead: A unified 1.3B-parameter framework designed for high-fidelity, infinite-length, and real-time streaming portrait video g…☆1,084May 28, 2026Updated 3 months ago
- Real-time stream editing pipeline powered by the FLUX.2-klein-4B model, optimized for consumer GPUs☆451Jun 13, 2026Updated 3 months ago
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆359Jun 24, 2026Updated 3 months ago
- [ICML 2026] ByteDance's All-in-One Video Generation Model for Human-Object Interaction Video Generation☆472Jul 29, 2026Updated last month
- ☆8,392May 27, 2026Updated 3 months ago
- [CVPR 2026] PersonaLive! : Expressive Portrait Image Animation for Live Streaming☆3,827Aug 28, 2026Updated 3 weeks ago
- TalkingMachines☆178Aug 2, 2025Updated last year
- [CVPR2026]We present FlashPortrait, an end-to-end video diffusion transformer capable of synthesizing ID-preserving, infinite-length vide…☆481Feb 21, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Connect to Gemini powered Pipecat bots using the Voice UI Kit and screen shares☆18Jan 27, 2026Updated 7 months ago
- ☆350Jan 2, 2026Updated 8 months ago
- A real-time streaming conversational video system that transforms text interactions into continuous, high-fidelity video responses using …☆343Dec 15, 2025Updated 9 months ago
- Take control back!☆18Jun 22, 2022Updated 4 years ago
- ☆22Aug 21, 2025Updated last year
- Zero-shot expressive voice cloning and speech generation. Generate anything from short clips to full-length audiobooks with realistic emo…☆551Jul 7, 2026Updated 2 months ago
- Lean neural real-time acoustic echo cancellation with soft delay estimation - GGML and PyTorch inference☆219Jul 20, 2026Updated 2 months ago
- Gaussian Wardrobe code release☆33Jul 5, 2026Updated 2 months ago
- StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars☆21Mar 31, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- An open-source model family for long-form speech, dialogue synthesis, voice design, sound effects, and real-time streaming TTS☆4,136Sep 6, 2026Updated 2 weeks ago
- Code Release for "OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data"☆81Aug 4, 2026Updated last month
- super expressive prompting model based on ltx2.3☆495May 23, 2026Updated 4 months ago
- Open-source AI that watches and hears your screen and reacts live as any personality you design — for faceless content, autonomous AI str…☆39Jul 7, 2026Updated 2 months ago
- Official Implementation of SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning☆1,221Aug 24, 2026Updated last month
- Legible, Scalable, Reproducible Foundation Models with Named Tensors and Jax☆16Jun 16, 2024Updated 2 years ago
- [AAAI 2026] EchoMimicV3: 1.3B Parameters are All You Need for Unified Multi-Modal and Multi-Task Human Animation☆1,063Mar 18, 2026Updated 6 months ago
- FineMotion: A Dataset and Benchmark with both Spatial and Temporal Annotation for Fine-grained Motion Generation and Editing☆17Mar 4, 2025Updated last year
- Long Video Gen Infrastructure☆2,640Sep 7, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆24Jan 31, 2024Updated 2 years ago
- DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation☆168Jun 26, 2026Updated 2 months ago
- Ideogram 4: Open image model at the forefront of design☆2,848Jun 30, 2026Updated 2 months ago
- Official Implementation of CoInteract: Spatially-Structured Co-Generation for Interactive Human-Object Video Synthesis☆170May 7, 2026Updated 4 months ago
- [CVPR-2025] The official code of HunyuanPortrait: Implicit Condition Control for Enhanced Portrait Animation☆351Dec 7, 2025Updated 9 months ago
- ICCV 2025 ACTalker: an end-to-end video diffusion framework for talking head synthesis that supports both single and multi-signal control…☆464Aug 20, 2025Updated last year
- [CVPR 2026 Poster] One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer☆498Apr 19, 2026Updated 5 months ago