SCOREQ: Speech COntrastive REgression for Quality Assessment (NeurIPS 2024)
☆120Aug 1, 2025Updated last year
Alternatives and similar repositories for scoreq
Users that are interested in scoreq are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of DNSMOS Pro (accepted at INTERSPEECH 2024).☆108Jun 8, 2025Updated last year
- Versatile Evaluation of Speech and Audio☆437Sep 20, 2026Updated last week
- NOMAD: Non-Matching Audio Distance (ICASSP 2024)☆32Jun 17, 2025Updated last year
- Expressive Anechoic Recordings of Speech (EARS)☆228Jun 25, 2024Updated 2 years ago
- A single-layer, streaming codec model providing SOTA audio quality and discrete tokens designed for superior downstream modelability.☆129Jun 4, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- This is the repository for the work "BridgeVoC: Revitalizing Neural Vocoder from a Restoration Perspective".☆67Nov 5, 2025Updated 10 months ago
- ☆95Jun 9, 2024Updated 2 years ago
- Speech Human Evaluation Estimation Toolkit (SHEET)☆142Mar 31, 2026Updated 5 months ago
- This is the official implementation of the SEMamba paper. (Accepted to IEEE SLT 2024)☆276Dec 12, 2025Updated 9 months ago
- A Neural Audio Codec (NAC) for Universal Audio☆47May 30, 2025Updated last year
- An neural full-band audio codec for general audio sampled at 48 kHz with 7.5 kps or 4.5 kbps.☆216Jul 26, 2026Updated 2 months ago
- ☆60Dec 2, 2024Updated last year
- Official data preparation and metric evaluation scripts for the Interspeech 2025 URGENT challenge.☆88May 21, 2025Updated last year
- Official implementation of the paper "BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec"☆221Sep 19, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [EMNLP 2024] ESC: Efficient Speech Coding with Cross-Scale Residual Vector Quantized Transformers☆127Mar 20, 2025Updated last year
- Reference-aware automatic speech evaluation toolkit☆190Dec 5, 2024Updated last year
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆19Aug 1, 2025Updated last year
- Room impulse response simulation for various array architectures using Monte-Carlo simulation and quaternions (Python)☆18Feb 25, 2026Updated 7 months ago
- Generating non-stationary multi-sensor signals under a spatial coherence constraint (Python)☆34Apr 12, 2026Updated 5 months ago
- The open source code for SimpleSpeech series☆147Oct 8, 2024Updated last year
- [InterSpeech 24] FreeV: Free Lunch For Vocoders Through Pseudo Inversed Mel Filter☆99Jul 4, 2024Updated 2 years ago
- Score-based Generative Models (Diffusion Models) for Speech Enhancement and Dereverberation☆772May 12, 2026Updated 4 months ago
- UTokyo-SaruLab MOS Prediction System☆371Apr 2, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Generation scripts for EARS-WHAM and EARS-Reverb☆48Jul 4, 2025Updated last year
- PyTorch implementation of Miipher-2 [2025] which is a speech restoration model by Google DeepMind☆71Sep 22, 2025Updated last year
- Variable Bitrate Residual Vector Quantization for Audio Coding☆56May 1, 2025Updated last year
- Official PyTorch implementation of "Paralinguistics-Aware Speech-Empowered LLMs for Natural Conversation" (NeurIPS 2024)☆95Dec 3, 2024Updated last year
- PASE: Phonologically Anchored Speech Enhancer☆99Jul 15, 2026Updated 2 months ago
- Unified automatic quality assessment for speech, music, and sound.☆762Jun 5, 2025Updated last year
- Official implementation of the paper "Laughter Synthesis using Pseudo Phonetic Tokens with a Large-scale In-the-wild Laughter Corpus" acc…☆77Jul 16, 2023Updated 3 years ago
- Official implementation: "AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation"☆20Oct 9, 2025Updated 11 months ago
- PAM is a no-reference audio quality metric for audio generation tasks☆81Jul 19, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Open implementation of UNIVERSE and UNIVERSE++ diffusion-based speech enhancement models.☆121Aug 29, 2024Updated 2 years ago
- Emotion Rendering for Conversational Speech Synthesis with Heterogeneous Graph-Based Context Modeling (Accepted by AAAI'2024)☆61Jun 20, 2024Updated 2 years ago
- ACM MM 2024 FlashSpeech: Efficient Zero-Shot Speech Synthesis☆159Sep 20, 2024Updated 2 years ago
- ☆174Sep 9, 2026Updated 2 weeks ago
- ☆17Mar 30, 2023Updated 3 years ago
- Audio Codec Speech processing Universal PERformance Benchmark☆311Jul 4, 2026Updated 2 months ago
- Official repository of Fast-ULCNet.☆44Jun 17, 2026Updated 3 months ago