☆66Jan 27, 2025Updated last year
Alternatives and similar repositories for javad
Users that are interested in javad are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- text-independent speaker identification☆12Apr 9, 2018Updated 8 years ago
- AI cover model for your own voice.☆34Aug 14, 2024Updated last year
- This repository contains a fine-tuning script for the transcription task of Mistral's Voxtral model.☆28Jul 31, 2025Updated 11 months ago
- Deno Library to upload files to GCS and obtain signed url☆11Jan 16, 2024Updated 2 years ago
- High accuracy code-switching whisper / qwen3 transcription☆39Jun 17, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆13Nov 30, 2022Updated 3 years ago
- ☆12Mar 24, 2024Updated 2 years ago
- ☆22Apr 29, 2025Updated last year
- NNSE (Neural Network Speech Enhancement) is a speech-denoiser optimized to run on Ambiq's low power platform☆44Nov 13, 2025Updated 8 months ago
- Convert Numerical Representations to Korean Pronunciation☆14Apr 20, 2020Updated 6 years ago
- A python package for finding words that sound like other words. Useful for entity resolution and poetry, among other things.☆15Oct 26, 2022Updated 3 years ago
- A curated list of awesome papers on contextualizing E2E ASR outputs☆81May 10, 2023Updated 3 years ago
- Speaker Diarization with Transformers☆70Jun 8, 2025Updated last year
- Cosine Similary Search in ElasticSearch + FAISS GPU☆12Mar 24, 2022Updated 4 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Synchronize Whisper's timestamps over an existing accurate transcription☆165May 28, 2024Updated 2 years ago
- This is the public repository for SALSA-Lite features for polyphonic sound event localization and detection using microphone arrays.☆15Dec 3, 2021Updated 4 years ago
- Xposed module that hooks into various HTTP libraries to log network calls.☆13Jan 4, 2023Updated 3 years ago
- On-device iOS clothing tagger powered by MLX-Swift.☆17Mar 11, 2025Updated last year
- This repository contains code for fine-tuning the Whisper speech-to-text model.☆24Jul 9, 2026Updated 2 weeks ago
- Data preparation utility for the finetuning of OpenAI's Whisper model.☆16Jun 18, 2026Updated last month
- ☆17Aug 8, 2021Updated 4 years ago
- 🔊😊 A fastapi voice-assistant framework to quickly prototype LLM-powered voice assistants in <5 minutes.☆31Jan 15, 2024Updated 2 years ago
- Local Action, Global Impact (Selected as Top 50 in the 2022 Solution Challenge.)☆17Jan 18, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Voice activity detection and speaker gender segmentation audiovisual corpus☆16Jan 20, 2025Updated last year
- Jupyter Notebook running Mamba speech synthesis example on Determined AI. Based on https://2084.substack.com/p/2084-marcrandbot-speech-sy…☆23Feb 8, 2024Updated 2 years ago
- ☆17Nov 17, 2020Updated 5 years ago
- ☆11Jan 12, 2026Updated 6 months ago
- My system for the DCASE 2022 Task 3 Sound Event Localizaiton and Detection.☆12Nov 12, 2022Updated 3 years ago
- Use Codestral Mamba with Visual Studio Code and the Continue extension. A local LLM alternative to GitHub Copilot.☆31Jul 18, 2024Updated 2 years ago
- Scripts to generate Dash docsets☆17Jun 21, 2026Updated last month
- Hosting a local Docker registry on a USB drive☆15Aug 19, 2017Updated 8 years ago
- A Javascript Chatbot built with the Gemini AI☆10Jan 26, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- JPEG Encoder for Cortex-M☆14Sep 2, 2016Updated 9 years ago
- A open-source toolkit for single and multi-modal speaker verification from modelscope and funasr with onnx☆15Dec 16, 2023Updated 2 years ago
- A Python library for learning and verification of neural networks and other machine learning models☆14Sep 18, 2025Updated 10 months ago
- GGML implementation of BERT model with Python bindings and quantization.☆57Feb 19, 2024Updated 2 years ago
- Ecoacoustic analysis platform empowering conservationists to analyze acoustic data and to derive insights about the ecosystem at scale☆19Jul 20, 2026Updated last week
- xcapi: A Python package for downloading animal sound recordings from xeno-canto API.☆20May 29, 2026Updated 2 months ago
- a GEGL based video editor☆20Aug 14, 2017Updated 8 years ago