Official implementation of "WhisperNER: Unified Open Named Entity and Speech Recognition"
☆199Feb 25, 2025Updated last year
Alternatives and similar repositories for whisper-ner
Users that are interested in whisper-ner are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Whisper with Medusa heads☆859Jul 2, 2026Updated last month
- Zero-shot Domain-sensitive Speech Recognition with Prompt-conditioning Fine-tuning (ASRU2023)☆26Oct 10, 2023Updated 2 years ago
- Code for ICASSP 2024 Paper: RECAP: Retrieval-Augmented Audio Captioning☆16Jun 23, 2024Updated 2 years ago
- A P2P blog and P2P Chat with no signalling server. Nothin' but RTC!☆16Nov 17, 2023Updated 2 years ago
- ☆20Jun 3, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ANDROID APP that can RECOGNIZE VLC LIVE AUDIO/VIDEO STREAMING (using free Android Developers Speech Recognition API) then TRANSLATE (usin…☆14Updated this week
- Code for the paper: How Much Context Does My Attention-Based ASR System Need?☆13Jul 29, 2026Updated last month
- A fork of Lyra (version 1) that supports a webassembly build. See https://github.com/mayitayew/soundstream-wasm for a more recent version…☆25Jul 19, 2022Updated 4 years ago
- DuckDB WebMacro: Share and Load your SQL Macros via gists☆15Jul 26, 2026Updated last month
- CTC decoder with hotwords for ASR.☆39Updated this week
- PANiC - PAraphrasing Noun-Compounds☆15Apr 6, 2018Updated 8 years ago
- Bulk unsubscribe from emails in your gmail account☆25Jun 15, 2024Updated 2 years ago
- Acoustic Neighbor Embeddings☆33Jul 13, 2025Updated last year
- Embroid: Unsupervised Prediction Smoothing Can Improve Few-Shot Classification☆11Aug 12, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Fully neural approach for text chunking☆420Oct 23, 2025Updated 10 months ago
- [Interspeech 2024] Enhancing Dysarthric Speech Recognition for Unseen Speakers via Prototype-Based Adaptation☆14Nov 28, 2024Updated last year
- Task-based Agentic Framework using StrictJSON as the core☆462Jun 8, 2026Updated 2 months ago
- Pytorch Implementation of ENet: A Deep Neural Network Architecture for Real-Time Semantic Segmentation (https://arxiv.org/abs/1606.02147)☆11Jan 24, 2020Updated 6 years ago
- An example of using pion/opus in WASM to decode audio files - https://sean-der.github.io/wasm-audio-decode/☆15Oct 21, 2022Updated 3 years ago
- Things you can do with the token embeddings of an LLM☆1,451Dec 1, 2025Updated 8 months ago
- Ultra fast and portable Parakeet implementation for on-device inference in C++ using Axiom with MPS+Unified Memory☆304May 4, 2026Updated 3 months ago
- first base model for full-duplex conversational audio☆1,799Jan 5, 2025Updated last year
- ☆37May 20, 2022Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆88Jul 31, 2025Updated last year
- INTERSPEECH 23 - Refunction Whisper to recognize new tasks with adapters!☆41Sep 11, 2023Updated 2 years ago
- Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.☆4,112Jan 8, 2025Updated last year
- Confidence Estimation for Black Box Automatic Speech Recognition Systems Using Lattice Recurrent Neural Networks https://arxiv.org/abs/19…☆14Apr 16, 2020Updated 6 years ago
- An OpenAI-compatible ASR/STT API server powered by Meta's omnilingual-asr model. Supports real-time streaming via WebSocket and batch tra…☆21Jan 2, 2026Updated 7 months ago
- [EMNLP Main '25] LiteASR: Efficient Automatic Speech Recognition with Low-Rank Approximation☆157May 18, 2025Updated last year
- The creative suite for character-driven AI experiences.☆192Sep 6, 2024Updated last year
- [ICASSP 2022] AISHELL-NER: Named Entity Recognition from Chinese Speech☆27Apr 20, 2022Updated 4 years ago
- 2023年iThome鐵人賽「AI & Data」組佳作【30天內成為NLP大師:掌握關鍵工具和技巧】完整程式碼,該文章會從零開始教你該如何微調大型語言模型☆18Nov 21, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Custom firmware for Realtek RTL8761B* adapters. From research "Reverse engineering Realtek RTL8761B* Bluetooth chips, to make better Blue…☆20May 3, 2026Updated 3 months ago
- An N-gram punctuator for Chinese and English.☆20Oct 14, 2025Updated 10 months ago
- mdast extension to parse and serialize MDX (or MDX.js)☆25Feb 10, 2024Updated 2 years ago
- Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces☆10,967Updated this week
- An electron Wrapper for Open-Interpreter for the lablab.ai hackathon☆12Oct 14, 2023Updated 2 years ago
- An Open Source text-to-speech system built by inverting Whisper.☆4,643Dec 14, 2025Updated 8 months ago
- Use DEMUCS to split songs into multiple sources☆20Apr 11, 2022Updated 4 years ago