A curated list of awesome Speaker Diarization papers, libraries, datasets, and other resources.
☆17Dec 15, 2019Updated 6 years ago
Alternatives and similar repositories for awesome-diarization
Users that are interested in awesome-diarization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Probabilistic Spherical Discriminant Analysis☆13Oct 29, 2022Updated 3 years ago
- Code for the Paper Speech Recognition and Multi-Speaker Diarization of Long Conversations☆39Jun 12, 2023Updated 3 years ago
- Original implementation of the pooling method introduced in "Speaker embeddings by modeling channel-wise correlations"☆11Sep 20, 2021Updated 4 years ago
- Schema2QA Question Answering Dataset☆19Aug 22, 2022Updated 4 years ago
- FEERCI: A Package for Fast non-parametric confidence intervals for Equal Error Rates☆12Mar 13, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆59Mar 28, 2025Updated last year
- Data Dialogue enables natural language querying of databases by integrating LLMs with SQL databases.☆15May 3, 2025Updated last year
- A simple pyaudio microphone interface☆11Jul 27, 2018Updated 8 years ago
- Official Implementation of Integrating Physics-Informed Vectors for Improved Wind Speed Forecasting with Neural Networks☆13Jul 2, 2026Updated 2 months ago
- Software for Decoding of High Order Ambisonics to Irregular Layouts☆13Mar 20, 2014Updated 12 years ago
- A Deep Convolutional Neural Network (DCNN) designed for the task of localizing human speech to 168 location classes using binaural microp…☆10Dec 16, 2017Updated 8 years ago
- A repo dedicated to different approaches in building a Persian Generative Chatbot.☆12Sep 7, 2022Updated 4 years ago
- Read and write HTK and HTS files from python.☆20Mar 17, 2015Updated 11 years ago
- 🔊 A comprehensive list of open-source datasets for voice and sound computing (95+ datasets).☆2,225Jun 6, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Aug 13, 2023Updated 3 years ago
- ☆160Jan 9, 2023Updated 3 years ago
- ☆10Apr 7, 2022Updated 4 years ago
- Radam+lookahead implemented by tensorflow☆11Oct 14, 2019Updated 6 years ago
- Simulation environment for sweep-based room impulse response measurements (student project)☆11Jun 10, 2017Updated 9 years ago
- Keywords and phrases that can be used for identifying mental-health-related conversation on Twitter☆12Jun 18, 2020Updated 6 years ago
- Dual fisheye video stitching in Python3, forked from : https://github.com/cynricfu/dual-fisheye-video-stitching☆13Dec 20, 2018Updated 7 years ago
- list of related work on AI DJ research☆15Apr 4, 2020Updated 6 years ago
- HRTF data preparation for machine learning by finding common measurement angles☆12May 14, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Recognizing common speech commands using Keras and Tensorflow.☆10Dec 17, 2018Updated 7 years ago
- This repository contains Google Collaboratory Notebooks for Deep Learning and Computer Vision Projects☆13Sep 29, 2021Updated 4 years ago
- Identifying individual speakers in an audio stream based on the unique characteristics found in individual voices using Python☆18Jun 18, 2023Updated 3 years ago
- Python Wrapper of visqol☆11Dec 23, 2024Updated last year
- An empathetic counselling chatbot. Retrieval-based, uses finetuned LMs for emotion identification and to boost empathy, novelty and fluen…☆18Jun 8, 2023Updated 3 years ago
- Chatbot: https://github.com/ChrisRahme/fyp-chatbot☆10Jun 22, 2021Updated 5 years ago
- Repo for the FB AI Speech team.☆27Aug 24, 2021Updated 5 years ago
- ARMAN: Pre-training with Semantically Selecting and Reordering of Sentences for Persian Abstractive Summarization☆11Oct 3, 2021Updated 4 years ago
- ☆13Dec 19, 2018Updated 7 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- MelGAN and Tacotron 2 in PyTorch☆11Oct 22, 2019Updated 6 years ago
- a conversion of Dadegan corpus (first Persian dependency corpus) to the universal dependency version☆15May 6, 2026Updated 4 months ago
- Python library for encoding and decoding APRS packets supporting RX/TX via APRS-IS or KISS☆13Jun 13, 2022Updated 4 years ago
- Official repo of ICASSP 2022 paper - Don't Separate, Learn to Remix: End-to-End Neural Remixing with Joint Optimization☆20Jan 7, 2025Updated last year
- This is a project of speech emotion recognition using KERAS based Semi-Generative Adversarial Networks.☆11May 17, 2018Updated 8 years ago
- Domoticz Plugin for controlling the ESP Milight Hib☆10Sep 8, 2021Updated 5 years ago
- The repository is created to support a Capstone project on the topic of "Study and Implementation of Sound Source Localization Techniques…☆13Apr 27, 2021Updated 5 years ago