πΉ pyannote + π notebook = pyannotebook
β27Jun 12, 2023Updated 3 years ago
Alternatives and similar repositories for pyannotebook
Users that are interested in pyannotebook are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Automatically setup the AISHELL-4 and MSDWild dataset for usage with pyannote-database (and pyannote-audio)β15Oct 22, 2025Updated 9 months ago
- C++ version of pyannote audio overlapped speech detection pipelineβ13Feb 14, 2024Updated 2 years ago
- Official repository for the "Powerset multi-class cross entropy loss for neural speaker diarization" paper published in Interspeech 2023.β96Oct 18, 2023Updated 2 years ago
- Learnable STRF, from Riad et al. 2021 JASAβ13Aug 21, 2021Updated 4 years ago
- β14Jun 12, 2015Updated 11 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- FEERCI: A Package for Fast non-parametric confidence intervals for Equal Error Ratesβ12Mar 13, 2024Updated 2 years ago
- Companion repository for the paper "A Comparison of Metric Learning Loss Functions for End-to-End Speaker Verification" published at SLSPβ¦β61Oct 7, 2020Updated 5 years ago
- Implementation of vocoders empowered with pytorch lightningβ18Jan 27, 2024Updated 2 years ago
- β12Nov 7, 2024Updated last year
- Multipurpose Multi Speaker Mixture Signal Generatorβ46Feb 6, 2025Updated last year
- Frequency-Dependent Adaptive Filtering Double Talk Detector.β13Mar 26, 2020Updated 6 years ago
- PyTorch implementation of PLDA as described in https://ravisoji.com/assets/papers/ioffe2006probabilistic.pdfβ15Oct 16, 2020Updated 5 years ago
- [NeurIPS 2022] "Losses Can Be Blessings: Routing Self-Supervised Speech Representations Towards Efficient Multilingual and Multitask Speeβ¦β17Sep 19, 2023Updated 2 years ago
- Automatic speech annotator processing speech with voice activaty detection, overlapping speech detection, speaker diarization and automatβ¦β33Jun 14, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- DiariZen Explained: A Tutorial for the Open Source State-of-the-Art Speaker Diarization Pipeline.β22Apr 24, 2026Updated 3 months ago
- β18Oct 24, 2025Updated 9 months ago
- Provide Gradio custom components to make the diarization-based audio labeling process easier and faster.β71Apr 22, 2026Updated 3 months ago
- Application for viewing Rich Transcription Time Marked (RTTM) files in an interactive wayβ48Apr 19, 2023Updated 3 years ago
- Simple Python package for fast DER computationβ35Jun 29, 2023Updated 3 years ago
- β12Mar 11, 2025Updated last year
- Companion repo for the paper "PixIT: Joint Training of Speaker Diarization and Speech Separation from Real-world Multi-speaker Recordingsβ¦β105Jan 10, 2025Updated last year
- β13Mar 23, 2026Updated 4 months ago
- β22Jun 30, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- PyTorch implementation of "Nextformer: A ConvNeXt Augmented Conformer For End-To-End Speech Recognition"β10Dec 15, 2022Updated 3 years ago
- Clustering-based methods for overlapping diarizationβ84Jan 12, 2024Updated 2 years ago
- Discriminative Training of VBx Diarizationβ28Sep 23, 2024Updated last year
- Tunable pipelinesβ41Sep 9, 2025Updated 10 months ago
- Dataset Catalogue Homepage for Indonesian Languagesβ12Feb 19, 2024Updated 2 years ago
- β73Feb 15, 2021Updated 5 years ago
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequβ¦β31Sep 20, 2025Updated 10 months ago
- β327Jun 14, 2024Updated 2 years ago
- Command line reference manager with a single source of truth: the .bib file. Inspired by beets.β34Jun 16, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- DUSTED: Spoken-Term Discovery using Discrete Speech Unitsβ17Oct 2, 2024Updated last year
- Clean and modernized implementation of FastSpeech2/LightSpeech using IPAβ18Aug 16, 2024Updated last year
- β13Sep 12, 2024Updated last year
- A general purpose task-agnostic speech augmentation policyβ16Mar 13, 2026Updated 4 months ago
- eSNN - Learning similarity measure from dataβ12Nov 28, 2019Updated 6 years ago
- A mini, simple, and fast end-to-end automatic speech recognition toolkit.β53Dec 6, 2022Updated 3 years ago
- Official implementation of "Wave-Trainer-Fit: Neural Vocoder with Trainable Prior and Fixed-Point Iteration towards High-Quality Speech Gβ¦β16Feb 6, 2026Updated 5 months ago