基于Kersa实现的声纹识别模型
☆156Sep 19, 2024Updated last year
Alternatives and similar repositories for VoiceprintRecognition-Keras
Users that are interested in VoiceprintRecognition-Keras are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 本项目使用了EcapaTdnn、ResNetSE、ERes2Net、CAM++等多种先进的声纹识别模型,同时本项目也支持了MelSpectrogram、Spectrogram、MFCC、Fbank等多种数据预处理方法☆318Dec 17, 2025Updated 7 months ago
- 说话人识别(声纹识别)算法的Python实现。包括GMM(已完成)、GMM-UBM、ivector、基于深度学习的声纹识别(self-attention已完成)。☆108Feb 21, 2023Updated 3 years ago
- This project uses a variety of advanced voiceprint recognition models such as EcapaTdnn, ResNetSE, ERes2Net, CAM++, etc. It is not exclud…☆1,307Dec 17, 2025Updated 7 months ago
- 声纹识别(Voiceprint Recognition, VPR),也称为说话人识别(Speaker Recognition),有两类,即说话人辨认(Speaker Identification)和说话人确认(Speaker Verification)☆58Mar 31, 2020Updated 6 years ago
- AliosThings 嵌入式声纹识别项目 https://github.com/alibaba/AliOS-Things/issues/976☆100Sep 4, 2019Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 基于语音的语种识别☆30Jul 23, 2023Updated 3 years ago
- Chinese voice corpus. 中文语音语料,语音更加清晰自然,包含8个开源数据集,3200个说话人,900小时语音,1300万字。☆751Jun 12, 2020Updated 6 years ago
- Causal Speech Enhancement Based on a Two-Branch Nested U-Net Architecture Using Self-Supervised Speech Embeddings☆21Jun 6, 2025Updated last year
- 监控视角下车辆检测和测速☆22Apr 30, 2022Updated 4 years ago
- Implementation of "FastSpeech: Fast, Robust and Controllable Text to Speech"☆64Jul 6, 2023Updated 3 years ago
- Kaldi based speaker verification☆47Jan 26, 2018Updated 8 years ago
- Vue移动商城项目,练习Vue时的demo。☆10Jan 6, 2023Updated 3 years ago
- A ResNet Speaker Recognition&Verification Demo☆27Oct 19, 2021Updated 4 years ago
- 基于 rasa 1.x 版本搭建的中文天气 查询 demo | A simple & micro Chinese Weatherbot based on rasa framework☆12Aug 14, 2019Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- wenet_LLM_from_ASLP☆15Nov 26, 2024Updated last year
- speech enhancement algorithms for microphone arrays☆15May 12, 2020Updated 6 years ago
- ☆12Jun 14, 2024Updated 2 years ago
- 说话人特征(声纹)提取工具,基于VGG-SR预训练模型。☆40Mar 7, 2020Updated 6 years ago
- [CVPR 2024] Not All Prompts Are Secure: A Switchable Backdoor Attack Against Pre-trained Vision Transfomers☆16Oct 24, 2024Updated last year
- Speaker recognition ,Voiceprint recognition☆53Feb 6, 2020Updated 6 years ago
- 一个简单的音频降噪工具,提高web UI界面和api接口☆46Nov 21, 2024Updated last year
- ☆13Oct 27, 2021Updated 4 years ago
- Age/Gender detection in Tensorflow☆13Mar 3, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 数字水印project,基于DWT的图像水印项目☆11Jun 22, 2023Updated 3 years ago
- ☆13Sep 25, 2024Updated last year
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- 基于dVector的说话人识别keras☆89Nov 30, 2020Updated 5 years ago
- (WIP)long form speech generatoins☆30Apr 2, 2025Updated last year
- PiRelay-V2 is developed by SB Components with the potential to control 4 appliances and loads up to 20V AC/ 7 A, 30V DC/ 10A to provide a…☆11Feb 9, 2022Updated 4 years ago
- 基于Pytorch实现的语音情感识别☆272Dec 17, 2025Updated 7 months ago
- 集美大学人工智能期末作业,实现通过声纹识别人物☆27Jun 9, 2019Updated 7 years ago
- ☆28Apr 24, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of Hybrid CTC/Attention Architecture for End-to-End Speech Recognition in pure python and PyTorch☆26Jul 25, 2024Updated 2 years ago
- An implement of GlowTTS model. Several modes are added: speaker embedding, prosody encoder(GST), and gradient reversal.☆55Sep 14, 2022Updated 3 years ago
- FastApi + pywebview + Vue3 + Vite Demo☆12Jan 14, 2024Updated 2 years ago
- 基于深度学习识别THCHS30数据集☆14Oct 27, 2021Updated 4 years ago
- 深度学习是利用卷积网络的深层结构提取的信息,卷积网络目前主要用于图像识别分类技术,其实在其中间层中包含了丰富的有用信息,而这些正是风格迁移的基础。 如果研究 CNN 的各层级结构,会发现里面的每一层神经元的激活态都对应了一种特定的信息,越是底层的就越接近画面的纹理信息,如…☆10Aug 25, 2021Updated 4 years ago
- OpenSpeaker is a completely independent and open source speaker recognition project. It provides the entire process of speaker recognitio…☆68Feb 16, 2022Updated 4 years ago
- ☆68Jul 17, 2024Updated 2 years ago