The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
☆3,289Jul 19, 2026Updated this week
Alternatives and similar repositories for Rapid-MLX
Users that are interested in Rapid-MLX are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 3x decode TPS increase On Qwen 3.6 27B @ temp 0.6 | Native MTP Speculative Decoding On Apple Silicon With No External Drafter.☆1,056Updated this week
- LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar☆17,950Updated this week
- OpenAI and Anthropic compatible server for Apple Silicon. Run LLMs and vision-language models (Llama, Qwen-VL, LLaVA) with continuous bat…☆1,446Jun 28, 2026Updated 3 weeks ago
- Lossless DFlash speculative decoding for MLX on Apple Silicon☆752Jun 11, 2026Updated last month
- 🔥 The fastest local AI engine for Apple Silicon. Optimised for agentic use.☆82May 24, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- MLX Studio - Home of JANG_Q - Image Gen/Edit + Chat/Code All in one - + OpenClaw (Anthropic API)☆908Updated this week
- Exact speculative decoding on Apple Silicon, powered by MLX.☆380Apr 20, 2026Updated 3 months ago
- Own your AI. The native macOS harness for AI agents -- any model, persistent memory, autonomous execution, cryptographic identity. Built …☆7,234Updated this week
- vMLX - JANGTQ Uber Compressed MLX Models - L2 Disk Cache (survives restart) + L1 Paged (super fast ttft) + Hybrid SSM Scheduler + Cont B…☆774Updated this week
- ⚡ Native MLX Swift LLM inference server for Apple Silicon. OpenAI-compatible API, SSD streaming for 100B+ MoE models, TurboQuant KV cache…☆722May 19, 2026Updated 2 months ago
- Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unslot…☆1,363Jun 23, 2026Updated 3 weeks ago
- MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.☆5,175Updated this week
- Run LLMs with MLX☆6,335Jul 11, 2026Updated last week
- DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm☆18,842Jul 3, 2026Updated 2 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.☆5,647Jun 30, 2026Updated 2 weeks ago
- DFlash: Block Diffusion for Flash Speculative Decoding☆5,496May 10, 2026Updated 2 months ago
- The headless browser for AI agents and web scraping☆19,433Updated this week
- A tool for creating and running Linux containers using lightweight virtual machines on a Mac. It is written in Swift, and optimized for A…☆48,021Updated this week
- 🎨 The open-source Claude Design alternative. 🖥️ Local-first desktop app. 🖼️ Your coding agent becomes the design engine: prototypes, l…☆79,670Updated this week
- The agent that grows with you☆217,147Updated this week
- Production-grade engineering skills for AI coding agents.☆79,265Updated this week
- The open-source managed agents platform. Turn coding agents into real teammates — assign tasks, track progress, compound skills.☆41,053Updated this week
- Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.☆20,237Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!☆70,136Updated this week
- A slide framework built for agents.☆5,875Updated this week
- 🐹 Clean, uninstall, analyze, optimize, and monitor your Mac from the terminal.☆59,337Updated this week
- #1 Persistent memory for AI coding agents based on real-world benchmarks☆25,354Updated this week
- Open-Source Frontier Voice AI☆50,143Updated this week
- PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.☆27,475Updated this week
- A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speec…☆7,577Jul 10, 2026Updated last week
- Garry's Opinionated OpenClaw/Hermes Agent Brain☆26,603Updated this week
- VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning☆33,775Jul 8, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Desktop app to manage markdown knowledge bases☆18,740Updated this week
- Your Personal AI super intelligence. A brain that builds a local-first memory of your life, a fantastic orchestrator of agent fleets and …☆35,080Updated this week
- Graphs that teach > graphs that impress. Turn any code into an interactive knowledge graph you can explore, search, and ask questions abo…☆75,148Updated this week
- AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI☆72,762Updated this week
- Open source Ghostty-based macOS terminal with vertical tabs and notifications for AI coding agents. Built for multitasking, organization,…☆24,779Updated this week
- 📡 Your own AI-powered news radar. Generates daily briefings in English & Chinese. | 用 AI 构建你专属的新闻雷达☆8,279Updated this week
- CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies☆71,836Updated this week