Source code of paper <End-to-End Language Diarization for Bilingual Code-switching Speech>
☆19Jan 23, 2022Updated 4 years ago
Alternatives and similar repositories for E2E-language-diarization
Users that are interested in E2E-language-diarization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Adnabod lleferydd Cymraeg i'r Gymraeg gyda HuggingFace // Speech Recognition for Welsh with HuggingFace☆13Nov 29, 2022Updated 3 years ago
- Official implementation of the paper "Distilling a Pretrained Language Model to a Multilingual ASR Model" (Interspeech 2022)☆12Mar 12, 2024Updated 2 years ago
- steps to perform text-based speaker diarization with kaldi toolkit☆12Nov 2, 2018Updated 7 years ago
- Syllable Segmentation and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Model☆35Aug 27, 2023Updated 3 years ago
- Estimating the Age, Height, and Gender of a speaker with their speech signal.☆15Sep 19, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆16Aug 1, 2025Updated last year
- ☆14Jun 12, 2015Updated 11 years ago
- ☆13Mar 25, 2021Updated 5 years ago
- Word Error Rate Estimation☆17Aug 25, 2020Updated 6 years ago
- ☆46Feb 16, 2023Updated 3 years ago
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- Accompanying code for paper "Attention-Based Contextual Language Model Adaptation for Speech Recognition", submitted to ACL 2021.☆14Jul 25, 2023Updated 3 years ago
- ☆12Aug 9, 2021Updated 5 years ago
- a standalone pitch extractor☆13Oct 19, 2017Updated 8 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆23Jun 25, 2026Updated 3 months ago
- Auto-KWS 2021 Challenge 1st place solution.☆11Jul 20, 2021Updated 5 years ago
- PHO-LID: A Unified Model to Incorporate Acoustic-Phonetic and Phonotactic Information for Language Identification☆21Aug 24, 2023Updated 3 years ago
- Source code and speech samples for the DSU-AVO paper accepted to INTERSPEECH 2023☆12May 13, 2024Updated 2 years ago
- Pytorch implementation of 'Improving Self-supervised Lightweight Model Learning via Hard-aware Metric Distillation. In ECCV 2022'☆11Mar 22, 2023Updated 3 years ago
- Dataset Catalogue Homepage for Indonesian Languages☆12Feb 19, 2024Updated 2 years ago
- ☆14Feb 9, 2023Updated 3 years ago
- American Sign Language Recognizer using Various Structures of CNN☆10Apr 26, 2020Updated 6 years ago
- Trained a CNN to understand American Sign Language and convert gestures into text☆14Oct 8, 2019Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Built a GUI application using Tkinter that helps Doctors to prepare prescriptions more efficiently. This uses speech to text conversion a…☆17Mar 26, 2021Updated 5 years ago
- Forced alignment decoder for Whisper.☆16Mar 13, 2024Updated 2 years ago
- Worked on hand gestures and recognizing the words and combine them to form a meaningful sentences. I have used two machine learning model…☆16Oct 17, 2020Updated 5 years ago
- End-to-end MOdeling of ASR (Automatic Speech Recognition)☆33Feb 16, 2023Updated 3 years ago
- Conformer: Convolution-augmented Transformer for Speech Recognition☆15Sep 4, 2025Updated last year
- Conformer RNN-Transducer☆14May 25, 2022Updated 4 years ago
- 🎹 pyannote + 🗒 notebook = pyannotebook☆27Jun 12, 2023Updated 3 years ago
- ☆16May 15, 2019Updated 7 years ago
- ☆10Sep 19, 2018Updated 8 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Interpretable unified language safety checking with large language models☆32Apr 15, 2023Updated 3 years ago
- 🎯 Speech Recognition Challenge by Speech Lab - IIT Madras☆10Nov 5, 2020Updated 5 years ago
- Repository containing experimentation platform on how to train, infer on wav2vec2 models.☆90Sep 22, 2022Updated 4 years ago
- ☆18Mar 13, 2024Updated 2 years ago
- ☆35Updated this week
- EfficientNet-Absolute Zero for Continuous Speech Keyword Spotting☆23Jun 16, 2022Updated 4 years ago
- An attempt to reproduce CALM (Continuous Audio Language Models) using DACVAE as the audio VAE.☆19Feb 20, 2026Updated 7 months ago