☆65Aug 4, 2026Updated last week
Alternatives and similar repositories for mseb
Users that are interested in mseb are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official repository for U-SAM (Interspeech 2025)☆28Jun 3, 2025Updated last year
- Code for T5lephone: Bridging Speech and Text Self-supervised Models for Spoken Language Understanding via Phoneme level T5☆19Nov 29, 2022Updated 3 years ago
- A comprehensive framework to test audio comprehension of Large Audio Language Models.☆69Jun 9, 2026Updated 2 months ago
- ☆19Apr 9, 2026Updated 4 months ago
- A toolkit for Spoken Language Understanding Evaluation (SLUE) benchmark. Refer paper https://arxiv.org/abs/2111.10367 for more details. O…☆65Feb 26, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Calculates and compares perceptual sound texture statistics☆16Mar 16, 2021Updated 5 years ago
- A temporal module for PyTorch-ComplexTensor☆44Jun 28, 2024Updated 2 years ago
- ☆32Nov 10, 2025Updated 9 months ago
- ☆51Feb 20, 2026Updated 5 months ago
- ☆24Jun 30, 2023Updated 3 years ago
- [ICLR 2026] StableToken: A state-of-the-art noise-robust semantic speech tokenizer featuring Voting-LFQ for resilient SpeechLLMs.☆33Feb 27, 2026Updated 5 months ago
- Audio Entailment: Deductive Reasoning for Audio Understanding☆17Dec 10, 2024Updated last year
- An original package of the dynamic compressive gammachirp filterbank (dcGC-FB)☆14Jul 7, 2026Updated last month
- Python wrapper around the BEZ2018 auditory nerve model☆12Mar 10, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Xmixers: A collection of SOTA efficient token/channel mixers☆29Sep 4, 2025Updated 11 months ago
- ☆33Nov 4, 2025Updated 9 months ago
- A lightweight library for normalizing speech transcripts before computing WER☆28Jul 14, 2026Updated 3 weeks ago
- Repository for speech paper reading☆33Aug 19, 2021Updated 4 years ago
- A toolkit for processing speech data and creating speech datasets☆212Mar 29, 2026Updated 4 months ago
- Tooling for agile modeling on large machine perception embedding databases.☆120Updated this week
- Repository of the IJCV'26 & WACV'24 paper☆35Apr 27, 2026Updated 3 months ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated 11 months ago
- This is the official codebase for WavJEPA. Time-domain audio foundation model for holistic downstream tasks. "Self-supervised learning fr…☆34Feb 28, 2026Updated 5 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- AudioBench: A Universal Benchmark for Audio Large Language Models☆323May 29, 2026Updated 2 months ago
- 書籍「意思決定分析と予測の活用 基礎理論からPython実装まで」のサンプルコードとデータ☆15Feb 18, 2021Updated 5 years ago
- MiMo-Audio: Audio Language Models are Few-Shot Learners☆1,074Jun 17, 2026Updated last month
- A 6-million Audio-Caption Paired Dataset Built with a LLMs and ALMs-based Automatic Pipeline☆208Dec 13, 2024Updated last year
- All information and news with respect to Falcon-H1 series☆122Oct 9, 2025Updated 10 months ago
- Audio-JEPA is an adaptation of the Joint-Embedding Predictive Architecture (JEPA) for self-supervised audio representation learning. Buil…☆66Jul 16, 2026Updated 3 weeks ago
- ☆21Mar 25, 2025Updated last year
- Effortless Push-to-Talk Transcription, Anywhere.☆28Jul 25, 2026Updated 2 weeks ago
- Official Implementation of LauraTSE: Target Speaker Extraction using Auto-Regressive Decoder-Only Language Models.☆37Nov 9, 2025Updated 9 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ASRU 2025] Omni-R1: Do You Really Need Audio to Fine-Tune Your Audio LLM?☆47Nov 21, 2025Updated 8 months ago
- Fast algorithm for determined blind source separation with update of demixing filters with joint adjustment of the remaining sources.☆36Mar 22, 2021Updated 5 years ago
- Accompanying code for our paper "Optimizing Short-Time Fourier Transform Parameters via Gradient Descent"☆33Oct 30, 2020Updated 5 years ago
- my ghostty config☆20Feb 25, 2026Updated 5 months ago
- Audio-Visual Lip Synthesis via Intermediate Landmark Representation☆19May 16, 2023Updated 3 years ago
- Towards Efficient Audio-Visual Learners via Empowering Pre-trained Vision Transformers with Cross-Modal Adaptation☆15Apr 13, 2024Updated 2 years ago
- A liberal plugin hosting library (VST3/AU/LV2/CLAP/AAP/WebCLAP) with fully-featured MIDI 2.0 (UMP + MIDI-CI) frontend where applicable. C…☆30Updated this week