NeuralAudio python SDK
☆22Mar 5, 2025Updated last year
Alternatives and similar repositories for neuralaudio-python
Users that are interested in neuralaudio-python are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Seamless Voice Interactions with LLMs☆12Oct 28, 2023Updated 2 years ago
- Easy local FLUX.1 Inference☆11Aug 29, 2024Updated last year
- ☆18Jan 17, 2025Updated last year
- TTS with RVC-module to generate .wav audios☆42Sep 17, 2025Updated 10 months ago
- Evaluation of Sentence Representations in Polish☆22Dec 29, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A Modality-agnostic Multi-task Foundation Model for Human Brain Imaging☆20Jun 23, 2026Updated last month
- Simple AI chat bubble for your website: Wordpress, React, HTML, Shopify. Answer questions about a website's content using RAG, streaming,…☆23Mar 24, 2025Updated last year
- [NeurIPS 2024] Low rank memory efficient optimizer without SVD☆33Jul 1, 2025Updated last year
- Live text-to-speech web app using GCP text-to-speech API☆33Jan 27, 2025Updated last year
- Talking AI Avatar in Realtime☆24Mar 30, 2024Updated 2 years ago
- FastAPI Implementation of Orpheus TTS streaming Chatbot☆30Jun 19, 2025Updated last year
- ☆34Apr 30, 2025Updated last year
- Expressive Gaussian Human Avatars from Monocular RGB Video (NeurIPS 2024)☆60May 28, 2025Updated last year
- Fine-tunes a student LLM using teacher feedback for improved reasoning and answer quality. Implements GRPO with teacher-provided evaluati…☆54May 7, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [AAAI 2025] VQTalker: Towards Multilingual Talking Avatars through Facial Motion Tokenization☆53Dec 16, 2024Updated last year
- ☆68Aug 11, 2025Updated last year
- This extension enhances the capabilities of textgen-webui by integrating advanced vision models, allowing users to have contextualized co…☆58Oct 22, 2024Updated last year
- TTS pipeline that uses RVC to enhance audio quality and cloning☆150Jan 25, 2024Updated 2 years ago
- [WACV 2025] - EmoVOCA: Speech-Driven Emotional 3D Talking Heads☆48Jun 27, 2025Updated last year
- Diffusion_TTS extension for booga☆71Sep 6, 2025Updated 11 months ago
- Triton with Windows support☆226Aug 3, 2026Updated last week
- This is official implementation of the paper: "iHuman: Instant Animatable Digital Humans From Monocular Videos" [ECCV 2024]☆88Dec 29, 2024Updated last year
- [CVPR 2024] IntrinsicAvatar: Physically Based Inverse Rendering of Dynamic Humans from Monocular Videos via Explicit Ray Tracing☆103Sep 16, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Git Re-Basin: Merging Models modulo Permutation Symmetries in PyTorch☆78Feb 9, 2023Updated 3 years ago
- MPMAvatar: Learning 3D Gaussian Avatars with Accurate and Robust Physics-Based Dynamics (NeurIPS 2025)☆87Dec 30, 2025Updated 7 months ago
- [TMLR 2025] When Attention Collapses: How Degenerate Layers in LLMs Enable Smaller, Stronger Models☆126Mar 6, 2026Updated 5 months ago
- Lightweight Gradio based WebUI for orpheusTTS - WSL / Linux [CUDA]☆109Nov 19, 2025Updated 8 months ago
- Code Repository for MeshAvatar: Learning High-quality Triangular Human Avatars from Multi-view Videos (ECCV 2024)☆129Nov 10, 2024Updated last year
- [NeurIPS '25] Create 3DGS heads from latent vectors☆111Apr 30, 2026Updated 3 months ago
- Accurate and general beat tracker☆353May 28, 2026Updated 2 months ago
- An efficent implementation of the method proposed in "The Era of 1-bit LLMs"☆155Oct 15, 2024Updated last year
- GAN-based Mel-Spectrogram Inversion Network for Text-to-Speech Synthesis☆1,040Aug 28, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆240Sep 5, 2024Updated last year
- [CVPR 2024] The official repo for FlashAvatar☆246Apr 12, 2024Updated 2 years ago
- 🤢 LipSick: Fast, High Quality, Low Resource Lipsync Tool 🤮☆225Jul 16, 2024Updated 2 years ago
- [CVPR 2026] CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction☆207May 16, 2026Updated 2 months ago
- Memoir+ a persona memory extension for Text Gen Web UI.☆225Feb 5, 2026Updated 6 months ago
- [IJCV 2025] Unlock Pose Diversity: Accurate and Efficient Implicit Keypoint-based Spatiotemporal Diffusion for Audio-driven Talking Portr…☆308Feb 6, 2026Updated 6 months ago
- A highly compressive and high-quality neural audio codec for speech models.☆270Jan 23, 2026Updated 6 months ago