FastAPI + MLX offline-first voice agent with <1s latency. Minimal UI
☆55Oct 21, 2025Updated 9 months ago
Alternatives and similar repositories for offline-voice-ai
Users that are interested in offline-voice-ai are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- this is an easy way to make ai podcast useing ai loccaly like ollama and the tts of piper☆16Feb 17, 2026Updated 6 months ago
- Local-first social memory search engine with browser capture, hybrid AI retrieval, and optional C++ acceleration.☆16May 13, 2026Updated 3 months ago
- A user-friendly GUI for llama.cpp — convert, quantize, and run GGUF models without touching the terminal.☆21Jun 9, 2026Updated 2 months ago
- ClaudeCode to OpenCode Migration guide☆16Jan 17, 2026Updated 7 months ago
- Bloat Free, Portable and Lightweight LLM Frontend (Single HTML file). With Lorebook, Web Search, Macro Engine etc.☆22Aug 1, 2026Updated 2 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆13Mar 10, 2025Updated last year
- scrape web content into readable markdown for llms and human readers☆10Feb 19, 2024Updated 2 years ago
- From-scratch implementation of OpenAI's GPT-OSS model in Python. No Torch, No GPUs.☆111Nov 5, 2025Updated 9 months ago
- ☆58Feb 8, 2026Updated 6 months ago
- Materialize is a program for converting images to materials for use in video games and whatnot☆10Aug 24, 2020Updated 5 years ago
- your private, personal assistant☆77Apr 16, 2026Updated 4 months ago
- Pure C wrapper library to use llama.cpp with Linux and Windows as simple as possible.☆15Jul 28, 2026Updated 3 weeks ago
- 🗣️ Real‑time, low‑latency voice, vision, and conversational‑memory AI assistant built on LiveKit and local LLMs☆112Jun 25, 2025Updated last year
- A sophisticated biologically inspired memory system for AI agents. Provides organic, high quality, persistent memory with self-maintenanc…☆72May 31, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- High-performance batched Top-K selection for CPU inference. Up to 80x faster than PyTorch, optimized for LLM sampling with AVX2 SIMD.☆18Mar 20, 2026Updated 4 months ago
- A functioning Sesame CSM project with a desktop GUI - Real-time factor: 0.6x with 4070 Ti Super - Requires only 8GB VRAM☆81May 19, 2025Updated last year
- Interactive Jacobian-Lens visualizer and live steerer for GGUF models on llama.cpp☆83Jul 12, 2026Updated last month
- OpenAI-compatible TTS API that unifies multiple backends with smart chunking for unlimited-length generation☆50Aug 3, 2026Updated 2 weeks ago
- A fast and lightweight macOS menu bar app to track GitHub releases. Features inline actions and Homebrew integration. Built entirely with…☆21Aug 5, 2026Updated last week
- Store, query, and create YAML workflow playbooks for LLM agents via MCP. STDIO or Streamable HTTP.☆31Jul 30, 2026Updated 2 weeks ago
- Dashboard v5 Coming Soon!!☆64Feb 15, 2026Updated 6 months ago
- Qwen2-VL for OCR & VQA☆19Sep 3, 2024Updated last year
- A Multi-Agentic AI Assistant/Builder☆28Jun 1, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Blazing-fast rust implementation of Sesame's Conversational Speech Model (CSM)☆86Mar 26, 2026Updated 4 months ago
- Decentralizing distribution of open-source AI models.☆21Updated this week
- Simple inbound/outbound packet sniffer☆31Oct 2, 2024Updated last year
- Ragamuffin - Chat with your document, articles or code☆13Nov 13, 2024Updated last year
- A CLI application for querying your local notes and files using AI-powered semantic search via Nia.☆15Feb 7, 2026Updated 6 months ago
- ☆14Sep 16, 2024Updated last year
- Create text chunks which end at natural stopping points without using a tokenizer☆26Nov 26, 2025Updated 8 months ago
- Anthropic's Contextual Retrieval implementation with visual chunk comparison. Preview context enrichment before/after embedding.☆30Sep 25, 2025Updated 10 months ago
- ☆25Aug 26, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Vector functions and indexing for SQLite☆10Mar 26, 2023Updated 3 years ago
- VisionEye with Ultralytics YOLOv10, achieving a new level of precision in real-time object detection and classification. With the powerfu…☆19Nov 17, 2025Updated 9 months ago
- Try Open WebUI on the Cloud!☆13Jan 3, 2025Updated last year
- AI debugger and AI coder integrated. Use AI to code and drives runtime debugger☆83Mar 17, 2026Updated 5 months ago
- Local LLM Server Manager + LlaMA.cpp + Chat☆21Updated this week
- GPU-accelerated voice assistant — local LLM, fine-tuned Whisper, Kokoro TTS, AMD ROCm☆21Apr 2, 2026Updated 4 months ago
- A production-ready multi-tenant RAG as a Service (RaaS) orchestrator☆31Nov 10, 2025Updated 9 months ago