☆21Dec 17, 2024Updated last year
Alternatives and similar repositories for h2s
Users that are interested in h2s are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Methods to compute various chroma audio features and audio similarity measures particularly for the task of cover song identification☆26Feb 7, 2020Updated 6 years ago
- MusicYOLO framework uses the object detection model, YOLOx, to locate notes in the spectrogram.☆18Jan 29, 2022Updated 4 years ago
- ☆28Jun 29, 2026Updated 2 months ago
- dog-can-sing-song☆159Jul 24, 2026Updated 2 months ago
- FastSAG: Towards Fast Non-Autoregressive Singing Accompaniment Generation☆30Dec 19, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- a Neural Vocoder supporting Ring Attention, Conformer and NSF.☆25Aug 1, 2025Updated last year
- Music generation☆25May 2, 2024Updated 2 years ago
- 蜜罐捕获的数据☆12May 16, 2016Updated 10 years ago
- Unofficial implementation JEN-1 Composer: A Unified Framework for High-Fidelity Multi-Track Music Generation(https://arxiv.org/abs/2310.1…☆32Jan 19, 2024Updated 2 years ago
- 夏目悠李/男声歌声データベースの最新ラベルデータ☆12Sep 2, 2020Updated 6 years ago
- this is a command line (linux, osx) rfc reader☆15Oct 12, 2014Updated 11 years ago
- Pytorch implementation for “V2C: Visual Voice Cloning”☆35Jan 28, 2023Updated 3 years ago
- r0ak ("roak") is the Ring 0 Army Knife -- A Command Line Utility To Read/Write/Execute Ring Zero on for Windows 10 Systems☆14Jan 16, 2019Updated 7 years ago
- Speech Recognition implementation using Artificial Neural Networks☆10Sep 7, 2015Updated 11 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Visualization for hidden Markov model computations☆14Dec 19, 2014Updated 11 years ago
- Expected edit distance implementation using OpenFst tools☆11May 13, 2015Updated 11 years ago
- Unsupervised speech activity detection system.☆11Jul 2, 2018Updated 8 years ago
- Wenet speech to text for react native☆10Nov 1, 2022Updated 3 years ago
- The demo page of UniAudio☆35Feb 5, 2024Updated 2 years ago
- Official implementation of Mozart's Touch: A Lightweight Multi-modal Music Generation Framework Based on Pre-Trained Large Models☆43Mar 17, 2026Updated 6 months ago
- Lie Detection by voice and heart rate☆10Dec 20, 2017Updated 8 years ago
- Automatically exported from code.google.com/p/transducersaurus☆11Apr 1, 2015Updated 11 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Multiobjective Optimization Training of PLDA for Speaker Verification☆10Jun 14, 2018Updated 8 years ago
- Open source image editor for windows 10. Can be controlled by voice commands and Cortana.☆17Nov 14, 2017Updated 8 years ago
- EditEvo is a browser-based video editor designed to provide users with editing capabilities directly within their web browser. The appli…☆11May 7, 2024Updated 2 years ago
- AnyAccomp: Generalizable accompaniment generation for vocals and solo instruments, powered by a quantized melodic bottleneck.☆41Dec 22, 2025Updated 9 months ago
- Text-Dependent Speaker Recognition System with Machine Learning Techniques☆10Dec 31, 2017Updated 8 years ago
- JavaScript libraries to interact with the Ispikit pronunciation assessment server☆11Nov 16, 2016Updated 9 years ago
- A simple pyaudio microphone interface☆11Jul 27, 2018Updated 8 years ago
- Hadoop-based tool for extraction of large scale synchronous grammars for paraphrasing and machine translation☆15Dec 2, 2016Updated 9 years ago
- This is application for dysarthria to improve their pronunciation by using deep learning☆10Dec 29, 2020Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Perform the forced decoding with target transcription☆11Sep 12, 2018Updated 8 years ago
- A python implementation of the neural network joint language model and an extension of it using global source context.☆11May 17, 2017Updated 9 years ago
- A C++ library implementing fast language models estimation using the 1-Sort algorithm.☆16May 18, 2023Updated 3 years ago
- The official implementation of the IJCAI 2024 paper "MusicMagus: Zero-Shot Text-to-Music Editing via Diffusion Models".☆49Sep 11, 2024Updated 2 years ago
- steps to perform text-based speaker diarization with kaldi toolkit☆12Nov 2, 2018Updated 7 years ago
- Implementation of joint bayesian model, written in python.☆11Aug 2, 2021Updated 5 years ago
- Demo WebApp using Kaldi DNN engine to convert speech to text☆11Jun 12, 2016Updated 10 years ago