Demo and samples for universal speech translator
☆25Nov 15, 2022Updated 3 years ago
Alternatives and similar repositories for speech_translation
Users that are interested in speech_translation are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A library of speech gadgets.☆15Oct 15, 2022Updated 3 years ago
- Simple Telegram bot to annotate and varify automatic speech recognition datasets☆12Mar 30, 2021Updated 5 years ago
- CMU multilingual speech repository☆30Apr 15, 2022Updated 4 years ago
- ☆20Nov 22, 2020Updated 5 years ago
- DSing ASR task: Resources and Baseline for an unaccompanied singing ASR.☆19Jul 9, 2026Updated last month
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Helping build fair, safe, ethical, and RIGHT General Artificial Intelligence and helping to introduce humanity to (RG)AI through the magi…☆11Aug 6, 2020Updated 6 years ago
- Vu Tran, Gihan Jayatilaka, Ashwin Ashok and Archan Misra, 2021, April. Deeplight : Robust & Unobtrusive Real-time Screen-Camera Communica…☆14Feb 10, 2023Updated 3 years ago
- Automation Testing☆10Apr 7, 2018Updated 8 years ago
- CVSS: A Massively Multilingual Speech-to-Speech Translation Corpus☆220Aug 26, 2022Updated 3 years ago
- One-shot TTS with Improved Unseen Speaker and Style Transfer☆37Mar 2, 2022Updated 4 years ago
- [DEPRECIATED] Symbolic MIDI Music AI implementation☆21Jun 11, 2022Updated 4 years ago
- PyTorch implementation of Listen Attend and Spell Automatic Speech Recognition (ASR).☆39Jul 25, 2019Updated 7 years ago
- the Tensorflow version of multi-speaker TTS training with feedback constraint☆40Oct 12, 2020Updated 5 years ago
- Using YouTube to prepare a speech recognition dataset for any language☆10Mar 30, 2021Updated 5 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Pytorch library for factorized L0-based pruning.☆45Oct 10, 2023Updated 2 years ago
- Incorporating KenLM language model with HuggingFace implementation of Wav2Vec2CTC Model using beam search decoding☆74Oct 11, 2021Updated 4 years ago
- Deep voice 3 + WORLD vocoder.☆16Jan 7, 2020Updated 6 years ago
- Transformer-based online speech recognition system with TensorFlow 2☆26Jan 22, 2021Updated 5 years ago
- This project shows how to build a simple handwriting recognizer in Keras with the IAM dataset.☆13Aug 15, 2021Updated 4 years ago
- Alphabot: a screen-less interactive spelling primer powered by computer vision☆14Sep 11, 2018Updated 7 years ago
- An android VoIP application using native SIP API & ConnectionService (CallKit in iOS) API☆10Mar 13, 2020Updated 6 years ago
- All class material☆16Jan 8, 2018Updated 8 years ago
- Text Classification Dataset for Turkish Language☆10Nov 16, 2021Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Implements a proof-of-concept of a multi-level clustering algorithm designed to enable extremely fast approximate match search in a large…☆12Feb 24, 2013Updated 13 years ago
- Tacotron2 with BERT examples☆10Jul 8, 2019Updated 7 years ago
- A generative model that could generate photo-realistic face images from hand-sketch face images.☆15Jun 17, 2022Updated 4 years ago
- Embedding Recycling for Language models☆38Jul 11, 2023Updated 3 years ago
- [ICLR 2022] "Audio Lottery: Speech Recognition Made Ultra-Lightweight, Noise-Robust, and Transferable", by Shaojin Ding, Tianlong Chen, Z…☆32Apr 8, 2022Updated 4 years ago
- Pytorch implementation of GauGAN, from https://arxiv.org/abs/1903.07291 (Park et al. 2019)☆14Oct 7, 2020Updated 5 years ago
- Transformer-based approaches for an efficient docstrings generation on a piece of Python's code.☆17Feb 16, 2026Updated 5 months ago
- ☆29Apr 28, 2026Updated 3 months ago
- A toolkit for Spoken Language Understanding Evaluation (SLUE) benchmark. Refer paper https://arxiv.org/abs/2111.10367 for more details. O…☆65Feb 26, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Text-to-Speech tutorial at SLTU 2016☆35May 10, 2016Updated 10 years ago
- [ICASSP'22] Integer-only Zero-shot Quantization for Efficient Speech Recognition☆34Oct 11, 2021Updated 4 years ago
- This repo is containing notes and implementations for cherry-picked publications of my particular interest☆12May 14, 2020Updated 6 years ago
- Pre-trained models for Honk☆11Apr 1, 2019Updated 7 years ago
- Python library for audio augmentation☆84Jul 6, 2023Updated 3 years ago
- Github repository for ACL 2025 paper: VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models☆24Jun 16, 2025Updated last year
- pytorch implementation for MultiSpeech: Multi-Speaker Text to Speech with Transformer paper☆21Jun 23, 2022Updated 4 years ago