Segment speech sequences based on speaker transitions, using ML and DSP.
☆17Jul 30, 2018Updated 8 years ago
Alternatives and similar repositories for Speaker-recognition
Users that are interested in Speaker-recognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Neural Turing machine for source separation in Tensorflow☆18Aug 16, 2017Updated 9 years ago
- Speaker Diarization is the problem of separating speakers in an audio. There could be any number of speakers and final result should stat…☆64Jan 8, 2021Updated 5 years ago
- Deep Discriminative Embeddings for Duration Robust Speaker Verification☆19Dec 16, 2019Updated 6 years ago
- Code for https://arxiv.org/abs/1712.00254☆18Dec 6, 2017Updated 8 years ago
- Losses and decoders for end-to-end ASR and OCR☆34Oct 30, 2020Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Deep neural models for core NLP tasks☆13Nov 9, 2017Updated 8 years ago
- Speaker diarization scripts, based on AaltoASR☆191Jan 3, 2019Updated 7 years ago
- Denoising autoencoders for speaker identification on MCE 2018 challenge☆12Nov 8, 2018Updated 7 years ago
- Single-channel blind source separation☆48Feb 5, 2018Updated 8 years ago
- Deep Neural Network for Speaker Count Estimation☆157Sep 5, 2020Updated 5 years ago
- Experiments for paper untitlted☆14Jul 25, 2020Updated 6 years ago
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 7 years ago
- Remove noise from sound clips by use of supervised training and an ideal ratio mask.☆14Apr 2, 2019Updated 7 years ago
- Visual Relocalization on the COLMAP reconstruction model☆14Nov 21, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆14Sep 21, 2022Updated 3 years ago
- A module for normalising text.☆10Nov 6, 2019Updated 6 years ago
- 以音素建模构建NN-CTC声学模型☆16May 14, 2019Updated 7 years ago
- Perform exploration, navigation and coverage path planning covering a room with UV energy with the Turtlebot3☆14Jul 31, 2022Updated 4 years ago
- ☆12Nov 9, 2018Updated 7 years ago
- Sequence.js Theme - A minimalist theme for showcasing products☆10Aug 21, 2015Updated 10 years ago
- ☆18Oct 14, 2022Updated 3 years ago
- STT Service based on Kaldi ASR☆15Aug 17, 2018Updated 8 years ago
- AsoSoft Speech Corpus can be used for spoken language processing tasks in Central Kurdish such as speech recognition, speaker recognition…☆10Mar 8, 2022Updated 4 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- 用多层BLSTM模型同时进行中文分词和标点符号预测☆18Nov 8, 2024Updated last year
- An eXample Programming Language☆11Dec 20, 2018Updated 7 years ago
- A pattern recognition project to estimate age using KNN, SVM and CNN☆10Jun 11, 2019Updated 7 years ago
- Speaker Diarization library in Python. Performs VAD, Segmentation, Linear Clustering, Hierarchical Clustering☆15Jul 28, 2017Updated 9 years ago
- Using Deep Learning for singing voice separation - Project for the course DT2119 Speech and Speaker Recognition offered by KTH in 2018☆15Jun 16, 2018Updated 8 years ago
- This is a online social robot navigation framework that implements several techniques for that matter, like the social relevance validity…☆12May 18, 2026Updated 3 months ago
- A zero-shot relation extractor, easily downloadable from the HuggingFace repo.☆12Aug 13, 2021Updated 5 years ago
- Language and Speech Technology for Central Kurdish Varieties (LREC-COLING 2024)☆14Nov 29, 2024Updated last year
- ☆13Nov 26, 2019Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Experiments with generating GPT-2 fanfiction on specified topics.☆11Jun 2, 2019Updated 7 years ago
- End-to-end speech recognition using TensorFlow☆48Apr 2, 2018Updated 8 years ago
- ROS2 driver pkg for wheeltec N100 IMU module.