β50Nov 23, 2022Updated 3 years ago
Alternatives and similar repositories for ULCA-asr-dataset-corpus
Users that are interested in ULCA-asr-dataset-corpus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- English ASR Challenge organized by Speech Lab, IIT Madrasβ10Feb 3, 2021Updated 5 years ago
- π― Speech Recognition Challenge by Speech Lab - IIT Madrasβ10Nov 5, 2020Updated 5 years ago
- β11Sep 5, 2025Updated 10 months ago
- β17Apr 14, 2023Updated 3 years ago
- This repository creates speaker diarization recipes to be used within the egs folder of kaldi.β17Aug 12, 2024Updated last year
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- β18Apr 28, 2021Updated 5 years ago
- Deno Library to upload files to GCS and obtain signed urlβ11Jan 16, 2024Updated 2 years ago
- β22Sep 24, 2018Updated 7 years ago
- β14Jun 12, 2015Updated 11 years ago
- Pronunciation-assisted Subword Modelingβ31May 30, 2019Updated 7 years ago
- Syllable Segmentation and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Modelβ35Aug 27, 2023Updated 2 years ago
- π LanMIT: A Toolkit for Improving Language Models in Low-resourced Speech Recognition based on Kaldi.β22Jul 12, 2019Updated 7 years ago
- This repo contains the baseline model recipes and pre-trained model for GramVanni hindi ASR challengeβ16Mar 26, 2022Updated 4 years ago
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequβ¦β31Sep 20, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Project for HIDING SPEAKERβS SEX IN SPEECH USING ZERO-EVIDENCE SPEAKER REPRESENTATION IN AN ANALYSIS/SYNTHESIS PIPELINEβ15Nov 30, 2022Updated 3 years ago
- Dictionary of pairs of Korean word and IPA crawled from Wiktionary (Korean edition)β23Nov 12, 2025Updated 8 months ago
- β24Jun 10, 2026Updated last month
- An espeak-compatible, permissively-licensed IPA phonemizer (G2P) based on DeepPhonemizer. Usable as a drop-in replacement for espeak's GPβ¦β111Mar 15, 2026Updated 4 months ago
- A simple lightweight library for text normalization for Indian Languagesβ18Sep 30, 2025Updated 9 months ago
- A transcription text editor with respeak moduleβ14Jan 24, 2026Updated 5 months ago
- π A list of accessible speech corpora for ASR, TTS, and other Speech Technologiesβ1,397Jun 6, 2024Updated 2 years ago
- The VoxTube dataset official repositoryβ71Feb 14, 2024Updated 2 years ago
- CLASP: Contrastive Language-Speech Pretraining for Multilingual Multimodal Information Retrievalβ13Jun 27, 2025Updated last year
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Python library for converting numbers to words for all Indian Languages.β38May 23, 2025Updated last year
- β14Jul 15, 2020Updated 6 years ago
- An implementation of Tiling and Corruption (TACo) Augmentations for OCR/HTRβ15Dec 4, 2021Updated 4 years ago
- sherpa with mlxβ15Aug 2, 2025Updated 11 months ago
- Aty-TTS: Improving fairness for spoken language understanding in atypical speech with Text-to-Speechβ11May 14, 2025Updated last year
- Speech recognition module for Python, supporting several engines and APIs, online and offline.β13Mar 9, 2022Updated 4 years ago
- A high-quality, varied ~30hr voice dataset suitable for training a TTS modelβ70Jan 7, 2023Updated 3 years ago
- Java Bindings for the C++ library DeepSpeechβ10Jun 4, 2020Updated 6 years ago
- Once more Diarization: Improving meeting transcription systems through segment-level speaker reassignmentβ14Feb 5, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Pybind11 bindings for Kaldiβ15Jul 11, 2026Updated last week
- A python tool that converts Arabic diacritised text to a sequence of phonemes and creates a pronunciation dictionary. This code is based β¦β15Sep 5, 2017Updated 8 years ago
- β13May 1, 2026Updated 2 months ago
- Dataset Catalogue Homepage for Indonesian Languagesβ12Feb 19, 2024Updated 2 years ago
- β13Oct 27, 2021Updated 4 years ago
- Convert English text from written expressions into spoken formsβ32Jun 22, 2022Updated 4 years ago
- Pretraining, fine-tuning and evaluation scripts for Indic-Wav2Vec2β117Aug 28, 2025Updated 10 months ago