☆66Sep 22, 2026Updated this week
Alternatives and similar repositories for mseb
Users that are interested in mseb are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official repository for U-SAM (Interspeech 2025)☆28Jun 3, 2025Updated last year
- Tools for visualizing neural nets☆22Jul 29, 2025Updated last year
- Code for T5lephone: Bridging Speech and Text Self-supervised Models for Spoken Language Understanding via Phoneme level T5☆19Nov 29, 2022Updated 3 years ago
- A comprehensive framework to test audio comprehension of Large Audio Language Models.☆71Jun 9, 2026Updated 3 months ago
- Don't just regulate gradients like in Muon, regulate the weights too☆32Jul 30, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆22Apr 9, 2026Updated 5 months ago
- Deep Convolutional TTS pytorch implementation☆27Jul 2, 2019Updated 7 years ago
- [ICLR 2025] Enhancing Self-Supervised Models with Audio Mixtures for Polyphonic Soundscapes☆81Oct 8, 2025Updated 11 months ago
- A toolkit for Spoken Language Understanding Evaluation (SLUE) benchmark. Refer paper https://arxiv.org/abs/2111.10367 for more details. O…☆66Feb 26, 2024Updated 2 years ago
- Calculates and compares perceptual sound texture statistics☆16Mar 16, 2021Updated 5 years ago
- A temporal module for PyTorch-ComplexTensor☆44Jun 28, 2024Updated 2 years ago
- Open-source models of NH and HI auditory processing☆11Jun 20, 2024Updated 2 years ago
- ☆24Jun 30, 2023Updated 3 years ago
- Textless (ASR-transcript free) Spoken Question Answering. The official release of NMSQA dataset and the implementation of "DUAL: Textless…☆35Aug 10, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2026] StableToken: A state-of-the-art noise-robust semantic speech tokenizer featuring Voting-LFQ for resilient SpeechLLMs.☆34Feb 27, 2026Updated 6 months ago
- Audio Entailment: Deductive Reasoning for Audio Understanding☆17Dec 10, 2024Updated last year
- An original package of the dynamic compressive gammachirp filterbank (dcGC-FB)☆14Aug 20, 2026Updated last month
- Python wrapper around the BEZ2018 auditory nerve model☆12Aug 24, 2026Updated 3 weeks ago
- Official Code Repo for "Learning on the Manifold: Unlocking Standard Diffusion Transformers with Representation Encoders" (ECCV 2026))☆22Jul 15, 2026Updated 2 months ago
- Xmixers: A collection of SOTA efficient token/channel mixers☆29Sep 4, 2025Updated last year
- ☆33Nov 4, 2025Updated 10 months ago
- A lightweight library for normalizing speech transcripts before computing WER☆29Updated this week
- Repository for speech paper reading☆33Aug 19, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A toolkit for processing speech data and creating speech datasets☆221Mar 29, 2026Updated 5 months ago
- These are Jupyter Notebooks to help guide people to learn how to use Praat-Parselmouth☆43Sep 29, 2021Updated 4 years ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated last year
- This is the official codebase for WavJEPA. Time-domain audio foundation model for holistic downstream tasks. "Self-supervised learning fr…☆36Sep 11, 2026Updated last week
- AudioBench: A Universal Benchmark for Audio Large Language Models☆329May 29, 2026Updated 3 months ago
- ☆43Nov 18, 2025Updated 10 months ago
- 書籍「意思決定分析と予測の活用 基礎理論からPython実装まで」のサンプルコードとデータ☆15Feb 18, 2021Updated 5 years ago
- Supervised Vector-Quantized Autoencoder☆20Jul 6, 2021Updated 5 years ago
- Overlap-Add filter in Python (using numpy)☆14Mar 11, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MiMo-Audio: Audio Language Models are Few-Shot Learners☆1,081Jun 17, 2026Updated 3 months ago
- A 6-million Audio-Caption Paired Dataset Built with a LLMs and ALMs-based Automatic Pipeline☆209Dec 13, 2024Updated last year
- ☆14Mar 16, 2019Updated 7 years ago
- Audio-JEPA is an adaptation of the Joint-Embedding Predictive Architecture (JEPA) for self-supervised audio representation learning. Buil…☆68Jul 16, 2026Updated 2 months ago
- OpenFLAM: Framewise Language Audio Model☆115Jun 4, 2026Updated 3 months ago
- [ASRU 2023] Code of paper SALT: Distinguishable Speaker Anonymization Through Latent Space Transformation☆23Aug 13, 2024Updated 2 years ago
- Official Implementation of LauraTSE: Target Speaker Extraction using Auto-Regressive Decoder-Only Language Models.☆40Nov 9, 2025Updated 10 months ago