Implementation of BEST-RQ - a model for self-supervised learning of speech signals using a random projection quantizer, in Pytorch.
☆137Sep 25, 2023Updated 2 years ago
Alternatives and similar repositories for best-rq-pytorch
Users that are interested in best-rq-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of the paper "Self-supervised Learning with Random-projection Quantizer for Speech Recognition" in Pytorch.☆97May 25, 2023Updated 3 years ago
- Sequence alignement methods with helpers for PyTorch.☆24Nov 30, 2022Updated 3 years ago
- E2E TTS using Conditional Flow Matching (Experimental*)☆71Nov 10, 2023Updated 2 years ago
- ☆39Oct 1, 2023Updated 2 years ago
- VAE modified from Descript Audio Codec, which replaces the RVQ with VAE☆92Apr 2, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official repository of the paper "MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization".☆380Aug 4, 2025Updated last year
- ACM MM 2023 CoMoSpeech: One-Step Speech and Singing Voice Synthesis via Consistency Model☆214Apr 26, 2024Updated 2 years ago
- Models and code for RepCodec: A Speech Representation Codec for Speech Tokenization☆196Jul 12, 2024Updated 2 years ago
- AAAI 2025: Codec Does Matter: Exploring the Semantic Shortcoming of Codec for Audio Language Model☆316Oct 12, 2025Updated 11 months ago
- [ICASSP 2024] This is the official code for "VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching"☆374Sep 3, 2024Updated 2 years ago
- [AAAI 2024] Code for CTX-vec2wav in UniCATS☆130Jun 11, 2024Updated 2 years ago
- Implementation of Spear-TTS - multi-speaker text-to-speech attention network, in Pytorch☆278Oct 30, 2023Updated 2 years ago
- Ultra-low bitrate neural audio codec (0.31~1.40 kbps) with a better semantic in the latent space.☆256Mar 7, 2025Updated last year
- The open source code for SimpleSpeech series☆145Oct 8, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Singing Voice Synthesis based on VITS, different from VISinger☆197Nov 13, 2023Updated 2 years ago