A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, with special optimizations for very long audio files.
☆18Dec 9, 2025Updated 8 months ago
Alternatives and similar repositories for parakeet-tdt-0.6b-v2-Batch-Transcriber
Users that are interested in parakeet-tdt-0.6b-v2-Batch-Transcriber are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pure-PyTorch Parakeet TDT inference☆52Mar 10, 2026Updated 5 months ago
- SLT 2024 Challenge: Post-ASR-Speaker-Tagging☆16Jun 16, 2024Updated 2 years ago
- ☆24Aug 1, 2026Updated 3 weeks ago
- SLISEMAP: Combining supervised dimensionality reduction with local explanations☆21Apr 24, 2025Updated last year
- ☆39Jul 21, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 🤯 A code collection for learning and exploration.☆18Mar 14, 2025Updated last year
- A demo-level low-latency, high-throughput inference engine for whisper☆20Nov 9, 2025Updated 9 months ago
- ☆19Aug 22, 2025Updated last year
- PyTorch implementation of RWKV blocks☆31Jul 22, 2025Updated last year
- Extract video from Samsung Motion Photo. Supports JPEG, HEIF/HEIC☆20Dec 11, 2025Updated 8 months ago
- ☆13Apr 16, 2025Updated last year
- AAAI-26 Oral | Implementation of "LoKI: Low-damage Knowledge Implanting of Large Language Models"☆23Apr 11, 2026Updated 4 months ago
- Ultra-Sortformer for Scalable Speaker Diarization☆28Apr 9, 2026Updated 4 months ago
- Compute WER and SER for speech recognition evaluation☆28Jun 6, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- We implemented the DEMUCS model for speech enhancement in the time-frequency domain, and additionally implemented HD-DEMUCS.☆34Nov 8, 2023Updated 2 years ago
- A repo containing download guidance and corresponding scripts of the VoxBlink dataset.☆30Apr 16, 2024Updated 2 years ago
- Finetune Nemo parakeet ASR model with new language (support 8 bit optimizer). Experimental birwkv-fastconformer TDT for long-form ASR(8.5…☆26Nov 27, 2025Updated 9 months ago
- A comprehensive manager and fixer for Gemini CLI (@google/gemini-cli) and Qwen CLI (@qwen-code/qwen-code) on Termux (Android).☆16Apr 25, 2026Updated 4 months ago
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated 11 months ago
- A unified evaluation suite for speech-to-text translation, covering SpeechLLMs, SFMs, and cascaded systems across diverse real-world spee…☆33Aug 9, 2026Updated 3 weeks ago
- ☆37Jan 6, 2026Updated 7 months ago
- Turn SMS Backup & Restore SMS XML files into HTML transcripts☆27May 17, 2019Updated 7 years ago
- An onnx-exportable implementation of iSTFT in torch☆34Feb 19, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- RWKV Batch infer backend ⚡Base on albatross https://github.com/BlinkDL/Albatross 🕊️☆37Jul 27, 2026Updated last month
- Project video on my Youtube channel about building an audio content analyzer dashboard.☆23Feb 22, 2023Updated 3 years ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆44Apr 11, 2026Updated 4 months ago
- Official Repository for "Efficient Vocal Source Separation Through Windowed RoFormer"☆46Oct 30, 2025Updated 10 months ago
- Convert your ChatGPT Message History into Data Viz and Markdown Notes☆28Oct 26, 2023Updated 2 years ago
- wav2vec2 audio classification for prosodic boundary detection and other tasks☆42Aug 11, 2023Updated 3 years ago
- Variations of L1 SNR Loss function for training audio source separation machine learning models☆45Aug 11, 2026Updated 2 weeks ago
- Softened ROSA QKV Operators for Training Next-Generation LLM Models☆39Aug 5, 2026Updated 3 weeks ago
- This repository implement a novel zero-shot TTS framework, named Flamed-TTS, focusing on the efficient generation and dynamic pacing in …☆57Aug 9, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official repository for Mamba-based Segmentation Model for Speaker Diarization☆47May 13, 2025Updated last year
- A web-based application to parse, view, and manage SMS backup files (XML format) from "SMS Backup & Restore" with advanced features like …☆30Aug 2, 2026Updated 3 weeks ago
- Implementation of 2-simplicial attention proposed by Clift et al. (2019) and the recent attempt to make practical in Fast and Simplex, Ro…☆49Sep 2, 2025Updated 11 months ago
- Export the STFT or ISTFT process in ONNX format.☆45Jun 6, 2026Updated 2 months ago
- Gemma-based Multilingual Machine Translation Models☆90Aug 14, 2026Updated 2 weeks ago
- Faster whisper Running on AMD GPUs with modified CTranslate 2 Libraries served up with Wyoming protocol☆35Aug 17, 2024Updated 2 years ago
- sms-db is a tool to build an SQLite database out of collections of SMS and MMS messages in various formats. The database can then be quer…☆46Dec 7, 2021Updated 4 years ago