☆66Jan 2, 2023Updated 3 years ago
Alternatives and similar repositories for whisper-ios-demo
Users that are interested in whisper-ios-demo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Nov 4, 2022Updated 3 years ago
- ☆12Nov 9, 2022Updated 3 years ago
- Dataset release for Emotional TTS in Indian Accent☆42Mar 25, 2026Updated 5 months ago
- This is a Python project that uses Selenium and OpenAI to scrape data from the web, process it with GPT-3, and generate reports based on …☆12Oct 28, 2025Updated 10 months ago
- Project for HIDING SPEAKER’S SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINE☆15Nov 30, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- Wenet speech to text for react native☆10Nov 1, 2022Updated 3 years ago
- NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment☆16Apr 13, 2022Updated 4 years ago
- Speechflow for emotion recognition related information decomposition☆10Jul 27, 2021Updated 5 years ago
- S3PRL for Speech Emotion Recognition (see s3prl > downstream)☆15Feb 28, 2026Updated 6 months ago
- End-to-end MOdeling of ASR (Automatic Speech Recognition)☆33Feb 16, 2023Updated 3 years ago
- This is a mirror of https://gitlab.com/tiro-is/tiro-speech-core☆15Jun 19, 2023Updated 3 years ago
- Wav2kws is keyword spotting (KWS) based on Wav2Vec 2.0. This model shows state-of-the-art in Google Speech Commands datasets V1 and V2.☆13Jun 11, 2021Updated 5 years ago
- From a large speech audio file and its corresponding body of text, automatically chunk the audio and text into (phrase, audio_snippet) pa…☆17May 15, 2015Updated 11 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation of the Rhythm Formant Analysis methodology for identifying speech rhythms and rhythm variation in the low frequency spectr…☆17Apr 27, 2023Updated 3 years ago
- Project that uses SFSpeechRecognizer to produce real-time Closed Captioning for videos☆23Oct 2, 2019Updated 6 years ago
- Code for the winning solution in the SE&R 2022 Challenge - SER track.☆16Mar 28, 2023Updated 3 years ago
- Unsupervised Voice Activity Detection by Modeling Source and System Information using Zero Frequency Filtering☆23Oct 19, 2023Updated 2 years ago
- A list of similar sounding words to help disambiguate voice coding☆11May 20, 2020Updated 6 years ago
- ☆16Jun 13, 2022Updated 4 years ago
- Implementation of the DIVA model of speech acquisition and production using PyTorch☆23Jan 18, 2023Updated 3 years ago
- Zero-shot Audio Classification using Whisper☆79Dec 12, 2022Updated 3 years ago
- ☆12Jul 20, 2020Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆18Apr 15, 2020Updated 6 years ago
- Supervisor trees for Go☆10Nov 4, 2017Updated 8 years ago
- Non Parallel Voice Conversion based on VITS☆24Mar 31, 2023Updated 3 years ago
- The world's simplest Computer Vision API for iOS developers.☆39Dec 12, 2022Updated 3 years ago
- A simple AppKit suggestion / autocompletion popup for macOS.☆21Feb 24, 2023Updated 3 years ago
- Simplifies running UI-tests 🏃☆13Mar 2, 2025Updated last year
- 「行動データの計算論モデリング」のサポートページです。☆11Mar 1, 2021Updated 5 years ago
- A free & open tool for transcribing audio interviews with offline ASR support☆26Dec 21, 2023Updated 2 years ago
- Make videos where "ChromaKey-Green" areas get twirled☆13Jan 30, 2018Updated 8 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Self-Supervised Speech/Sound Pre-training and Representation Learning Toolkit☆13Nov 18, 2022Updated 3 years ago
- ☆32Jan 6, 2022Updated 4 years ago
- ☆10Feb 19, 2018Updated 8 years ago
- A toolset for easy formant extraction and visualization from wav files and TTS models☆33Sep 2, 2022Updated 4 years ago
- Accelerate Whisper tasks such as transcription, by multiprocesing through parallelization☆25Oct 29, 2022Updated 3 years ago
- An High-resolution implementation of HiFi-GAN Vocoder for Voice Conversion.☆32Apr 10, 2023Updated 3 years ago
- Swift library for Buffers, Arrays, Bits and Bytes.☆13Jul 23, 2022Updated 4 years ago