The project is associated with the recently-launched ICASSP 2022 Multi-channel Multi-party Meeting Transcription Challenge (M2MeT) to provide participants with baseline systems for speech recognition and speaker diarization in conference scenario.
☆142Jun 10, 2022Updated 4 years ago
Alternatives and similar repositories for AliMeeting
Users that are interested in AliMeeting are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- End-to-End Neural Diarization☆435Aug 30, 2021Updated 4 years ago
- MagicData-RAMC Dataset and Baseline☆64Sep 13, 2022Updated 3 years ago
- ☆38Mar 30, 2021Updated 5 years ago
- Variational Bayes HMM over x-vectors diarization☆287Jan 15, 2024Updated 2 years ago
- Python package for combining diarization system outputs.☆94Oct 12, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆140Jul 21, 2021Updated 5 years ago
- A simple package for Guided source separation (GSS)☆134May 20, 2024Updated 2 years ago
- Code for the ICASSP-2021 paper: Continuous Speech Separation with Conformer.☆120Mar 18, 2023Updated 3 years ago
- Tools for Speech Enhancement integrated with Kaldi☆432Jul 6, 2023Updated 3 years ago
- ☆32Sep 14, 2022Updated 3 years ago
- A PyTorch implementation of End-to-End Neural Diarization☆110Jun 19, 2023Updated 3 years ago
- The baseline system for the ICASSP2024 ICMC-ASR Challenge.☆57Dec 6, 2023Updated 2 years ago
- PyTorch implementation of TinyWASE described in our paper "Compressing Speaker Extraction Model with Ultra-low Precision Quantization and…☆11Jun 28, 2021Updated 5 years ago
- This repo is for the SPL paper "Auto-Tuning Spectral Clustering for Speaker Diarization Using Normalized Maximum Eigengap"☆125Apr 8, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repository contains a set of codes to run (i.e., train, perform inference with, evaluate) a diarization method called EEND-vector-cl…☆81Oct 18, 2022Updated 3 years ago
- Production first, nn-based on-device signal processing toolkit.☆63May 30, 2023Updated 3 years ago
- An Open Source Tools for Speaker Recognition☆638Aug 5, 2024Updated last year
- Diarization scoring tools.☆268Apr 8, 2026Updated 3 months ago
- An unofficial implementation of the Personal VAD speaker-conditioned voice activity detection method. Bachelor's thesis project.☆90Sep 22, 2022Updated 3 years ago
- Cross-Speaker Encoding Network for Multi-talker Speech Recognition☆12Mar 14, 2025Updated last year
- An open source dataset for source separation☆502Feb 9, 2024Updated 2 years ago
- Python library for Room Impulse Response (RIR) simulation with GPU acceleration☆607Jul 18, 2025Updated last year
- ☆15Sep 6, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- target speaker extraction and verification for multi-talker speech☆211Jan 24, 2021Updated 5 years ago
- ☆33Jun 26, 2023Updated 3 years ago
- ☆95Apr 24, 2025Updated last year
- ☆33Mar 11, 2022Updated 4 years ago
- Clustering-based methods for overlapping diarization☆84Jan 12, 2024Updated 2 years ago
- SMS-WSJ: Spatialized Multi-Speaker Wall Street Journal database for multi-channel source separation and recognition☆131Jun 7, 2024Updated 2 years ago
- ☆145Oct 25, 2021Updated 4 years ago
- ☆42Oct 14, 2022Updated 3 years ago
- Speech enhancement system for the CHiME-5 dinner party scenario☆111Feb 6, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Speaker-aware CTC (SACTC) for multi-talker overlapped speech recognition.☆22May 26, 2025Updated last year
- This repository contains the baseline system for CHiME-8 MMCSG challenge focusing on transcribing both sides of a conversation where one …☆41Mar 13, 2024Updated 2 years ago
- Libri-CSS: dataset and evaluation pipeline☆157Jan 18, 2023Updated 3 years ago
- SpEx+(tied) source code☆96Jul 6, 2023Updated 3 years ago
- E2E system with LF-MMI; word N-gram for Mandarin☆167Apr 29, 2022Updated 4 years ago
- ☆55Jan 15, 2021Updated 5 years ago
- ☆59Mar 28, 2025Updated last year