The project is associated with the recently-launched ICASSP 2022 Multi-channel Multi-party Meeting Transcription Challenge (M2MeT) to provide participants with baseline systems for speech recognition and speaker diarization in conference scenario.
☆143Jun 10, 2022Updated 4 years ago
Alternatives and similar repositories for AliMeeting
Users that are interested in AliMeeting are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- End-to-End Neural Diarization☆436Aug 30, 2021Updated 4 years ago
- MagicData-RAMC Dataset and Baseline☆64Sep 13, 2022Updated 3 years ago
- ☆40Mar 30, 2021Updated 5 years ago
- Variational Bayes HMM over x-vectors diarization☆288Jan 15, 2024Updated 2 years ago
- ☆140Jul 21, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Python package for combining diarization system outputs.☆94Aug 3, 2026Updated 2 weeks ago
- A simple package for Guided source separation (GSS)☆134May 20, 2024Updated 2 years ago
- Code for the ICASSP-2021 paper: Continuous Speech Separation with Conformer.☆120Mar 18, 2023Updated 3 years ago
- Tools for Speech Enhancement integrated with Kaldi☆433Jul 6, 2023Updated 3 years ago
- ☆33Sep 14, 2022Updated 3 years ago
- A PyTorch implementation of End-to-End Neural Diarization☆110Jun 19, 2023Updated 3 years ago
- The baseline system for the ICASSP2024 ICMC-ASR Challenge.☆57Dec 6, 2023Updated 2 years ago
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- This repo is for the SPL paper "Auto-Tuning Spectral Clustering for Speaker Diarization Using Normalized Maximum Eigengap"☆125Apr 8, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This repository contains a set of codes to run (i.e., train, perform inference with, evaluate) a diarization method called EEND-vector-cl…☆81Oct 18, 2022Updated 3 years ago
- Production first, nn-based on-device signal processing toolkit.☆63May 30, 2023Updated 3 years ago
- An Open Source Tools for Speaker Recognition☆636Aug 5, 2024Updated 2 years ago
- Diarization scoring tools.☆272Apr 8, 2026Updated 4 months ago
- An unofficial implementation of the Personal VAD speaker-conditioned voice activity detection method. Bachelor's thesis project.☆90Sep 22, 2022Updated 3 years ago
- Cross-Speaker Encoding Network for Multi-talker Speech Recognition☆12Mar 14, 2025Updated last year
- An open source dataset for source separation☆501Feb 9, 2024Updated 2 years ago
- Python library for Room Impulse Response (RIR) simulation with GPU acceleration☆611Jul 18, 2025Updated last year
- ☆15Sep 6, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- target speaker extraction and verification for multi-talker speech☆211Jan 24, 2021Updated 5 years ago
- ☆33Jun 26, 2023Updated 3 years ago
- ☆95Apr 24, 2025Updated last year
- ☆33Mar 11, 2022Updated 4 years ago
- Clustering-based methods for overlapping diarization☆84Jan 12, 2024Updated 2 years ago
- SMS-WSJ: Spatialized Multi-Speaker Wall Street Journal database for multi-channel source separation and recognition☆131Jun 7, 2024Updated 2 years ago
- ☆149Oct 25, 2021Updated 4 years ago
- ☆42Oct 14, 2022Updated 3 years ago
- Speech enhancement system for the CHiME-5 dinner party scenario☆111Feb 6, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Speaker-aware CTC (SACTC) for multi-talker overlapped speech recognition.☆22May 26, 2025Updated last year
- This repository contains the baseline system for CHiME-8 MMCSG challenge focusing on transcribing both sides of a conversation where one …☆41Mar 13, 2024Updated 2 years ago
- Libri-CSS: dataset and evaluation pipeline☆159Jan 18, 2023Updated 3 years ago
- SpEx+(tied) source code☆96Jul 6, 2023Updated 3 years ago
- E2E system with LF-MMI; word N-gram for Mandarin☆167Apr 29, 2022Updated 4 years ago
- ☆55Jan 15, 2021Updated 5 years ago