☆21Mar 7, 2023Updated 3 years ago
Alternatives and similar repositories for whisper-flash-attention
Users that are interested in whisper-flash-attention are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Oct 17, 2021Updated 4 years ago
- Fast and differentiable hidden Markov model in C++☆19Jan 20, 2023Updated 3 years ago
- ☆19Feb 28, 2018Updated 8 years ago
- ☆15May 8, 2021Updated 5 years ago
- A BKTree written in C++☆11Jul 8, 2011Updated 15 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Adaptive Multimodal Reasoning via Reinforcement Learning☆23Jan 11, 2026Updated 6 months ago
- ☆24Jun 17, 2020Updated 6 years ago
- PyTorch implementation of "Jasper: An End-to-End Convolutional Neural Acoustic Model" (INTERSPEECH 2019)☆32Mar 4, 2021Updated 5 years ago
- Pytorch implementation of "Natural TTS Synthesis by Conditioning WaveNet on Mel Spectrogram Predictions", ICASSP, 2018.☆19Jan 21, 2021Updated 5 years ago
- c++ Kaldi IO lib (static and dynamic).☆25Nov 26, 2018Updated 7 years ago
- Code for paper "direct speech-to-image translation"☆26Jun 8, 2020Updated 6 years ago
- A simple n-gram language model.☆12Sep 11, 2018Updated 7 years ago
- Review of papers I read☆14Dec 11, 2020Updated 5 years ago
- Linear Prediction Coefficients estimation from mel-spectrogram implemented in Python based on Levinson-Durbin algorithm.☆72Mar 19, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆36Mar 14, 2025Updated last year
- Semi-supervised spoken language understanding (SLU) via self-supervised speech and language model pretraining☆12Mar 23, 2021Updated 5 years ago
- A TensorFlow implementation of light convolutional neural network (LCNN)☆12Dec 27, 2018Updated 7 years ago
- Official PyTorch implementation of "t-EER: Parameter-Free Tandem Evaluation Metric of Countermeasures and Biometric Comparators"☆14Sep 25, 2023Updated 2 years ago
- Repository for speech paper reading☆33Aug 19, 2021Updated 4 years ago
- 🚀 Implementation of easy-to-use 3D parallelism based on Huggingface Transformers & Microsoft DeepSpeed☆31Feb 5, 2022Updated 4 years ago
- simple energy vad☆19Jun 3, 2017Updated 9 years ago
- ☆13Jan 10, 2017Updated 9 years ago
- Code for "Phoneme Segmentation Using Self-Supervised Speech Models", Strgar & Harwath, Proceedings of the IEEE Spoken Language Technology…☆55Nov 4, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Collection of Common Machine Translation Tools☆11Jul 26, 2022Updated 4 years ago
- High-level API for tar-based dataset☆12Feb 3, 2024Updated 2 years ago
- ☆11Jun 4, 2021Updated 5 years ago
- This project is about performing Speaker diarization for Hindi Language.☆58Mar 21, 2021Updated 5 years ago
- Into the depths of some concepts of Artificial Intelligence and Machine Learning☆10Apr 4, 2026Updated 3 months ago
- Library about construction helper for Generative models e.g. Flow-based Model with Tensorflow 2.x.☆12Feb 16, 2023Updated 3 years ago
- Here we will track the latest Audio AI Agent, including speech, music, sound effects, etc.☆16Dec 8, 2023Updated 2 years ago
- KGML for EMNLP 2021☆10Feb 2, 2022Updated 4 years ago
- ☆13Jul 4, 2020Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆18Jan 10, 2024Updated 2 years ago
- Pulse Model vocoder☆42Dec 5, 2018Updated 7 years ago
- ONNXモデルをpyca/cryptographyを用いて暗号化/復号化するサンプル☆16Mar 19, 2022Updated 4 years ago
- Pytorch implemenation of the model proposed in the paper: Double Multi-Head Attention for Speaker Verification☆19Jul 25, 2024Updated 2 years ago
- A Benchmark Corpus for Low-Resource Cantonese Punctuation Restoration from Speech Transcripts☆15Dec 3, 2024Updated last year
- a compact audio-to-phoneme aligner for singing voice☆12Jan 17, 2024Updated 2 years ago
- Korean Speech to English Translation Corpus☆45Sep 3, 2021Updated 4 years ago