The official implementation of the method discussed in the paper Improving Spoken Language Identification with Map-Mix(work accepted at ICASSP-2023)
☆18Feb 17, 2023Updated 3 years ago
Alternatives and similar repositories for Map-Mix
Users that are interested in Map-Mix are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2022] "Losses Can Be Blessings: Routing Self-Supervised Speech Representations Towards Efficient Multilingual and Multitask Spee…☆17Sep 19, 2023Updated 2 years ago
- PHO-LID: A Unified Model to Incorporate Acoustic-Phonetic and Phonotactic Information for Language Identification☆21Aug 24, 2023Updated 3 years ago
- Dataset Release for Phone Number Entity capture task☆14Sep 2, 2022Updated 4 years ago
- Dataset Release for Intent Classification from Speech☆48Feb 23, 2025Updated last year
- An awesome spoken LID repository. (Working in progress☆110Apr 22, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code repository for the paper "Improving End-to-End SLU performance with Prosodic Attention and Distillation" accepted at Interspeech 202…☆27May 17, 2023Updated 3 years ago
- Code for ICASSP 2024 Paper: RECAP: Retrieval-Augmented Audio Captioning☆16Jun 23, 2024Updated 2 years ago
- ☆18Mar 13, 2024Updated 2 years ago
- Official implement of "Dual-stream Time-Delay Neural Network with Dynamic Global Filter for Speaker Verification" in PyTorch☆41Aug 31, 2023Updated 3 years ago
- Skit's tech website☆11Jul 1, 2024Updated 2 years ago
- A time delay estimation method for event-based time-series data. Time delay estimation is also known as the correction of time offsets an…☆16Dec 3, 2025Updated 8 months ago
- Source code for "Inside Cricket: A fifth umpire' view of your favorite sport"☆12Apr 15, 2018Updated 8 years ago
- Create flowcharts in elm☆13Apr 19, 2021Updated 5 years ago
- This repository contains all the code necessary for running the multilingual distilwhisper from Ferraz et al. 2024 IEEE ICASSP paper.☆34Apr 22, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 🗣️ Convert between phonetic alphabets☆11Feb 7, 2022Updated 4 years ago
- Vim plugin to get scores and commentary of live cricket matches☆12Jun 13, 2019Updated 7 years ago
- An MRCP server load balancer using OpenSIPS☆19Jun 4, 2020Updated 6 years ago
- Spoken Language Identification on Common Voice and AudioSet using Deep Learning☆42Feb 4, 2026Updated 6 months ago
- golang vad (voice activity detection) library based on webrtc☆13Dec 13, 2021Updated 4 years ago
- This repository contains a short introduction on the topic of audio and speech processing -- from basics to applications.☆19Dec 20, 2023Updated 2 years ago
- A hackable Emacs based data-tagging framework☆21Jul 28, 2019Updated 7 years ago
- Speaker diarization with GMM-UBM and MAP Adaptation☆30Sep 13, 2018Updated 7 years ago
- ☆11Sep 4, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Vim plugin to fuzzy search tabs opened in all the browser windows and switch.☆19Feb 5, 2020Updated 6 years ago
- Repository for "Training Audio Captioning Models without Audio"☆10Sep 26, 2023Updated 2 years ago
- Run commands on remote hosts, inspecting key indicators to manage infrastructure☆15Jan 29, 2026Updated 7 months ago
- This repo contains the code for "Voice Disorder Analysis: A Transformer-based Approach", accepted at Interspeech 2024☆15Jun 11, 2024Updated 2 years ago
- ☆12Jun 14, 2024Updated 2 years ago
- ☆14Jan 17, 2023Updated 3 years ago
- Models and codes for INTERSPEECH 2023 paper DistilXLSR: A Light Weight Cross-Lingual Speech Representation Model☆13Mar 30, 2025Updated last year
- Learning Domain-Invariant Transformation for Speaker Verification.☆11Jun 13, 2023Updated 3 years ago
- The official pytorch implemention of the Intespeech 2024 paper "Reshape Dimensions Network for Speaker Recognition"☆212Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Rainbow Keywords - Official PyTorch Implementation☆14Jun 27, 2024Updated 2 years ago
- Job descriptions for Tech roles at Skit☆14Aug 29, 2024Updated 2 years ago
- Spoken Language Identification from Short Utterances☆13Jul 6, 2022Updated 4 years ago
- Zafar's Audio Functions in Python for audio signal analysis: STFT, inverse STFT, mel filterbank, mel spectrogram, MFCC, CQT kernel, CQT s…☆59Aug 8, 2025Updated last year
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignment☆14Feb 5, 2025Updated last year
- Research code for "Towards multi-task learning of speech and speaker recognition" at https://arxiv.org/pdf/2302.12773.pdf☆12Dec 2, 2024Updated last year
- Aim to implement a classifier which classifies an audio sample into speech or music.☆10Sep 17, 2019Updated 6 years ago