The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
☆3,419Aug 8, 2026Updated this week
Alternatives and similar repositories for Rapid-MLX
Users that are interested in Rapid-MLX are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 3x decode TPS increase On Qwen 3.6 27B @ temp 0.6 | Native MTP Speculative Decoding On Apple Silicon With No External Drafter.☆1,144Updated this week
- LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar☆18,529Updated this week
- OpenAI and Anthropic compatible server for Apple Silicon. Run LLMs and vision-language models (Llama, Qwen-VL, LLaVA) with continuous bat…☆1,497Updated this week
- Lossless DFlash speculative decoding for MLX on Apple Silicon☆758Jun 11, 2026Updated last month
- 🔥 The fastest local AI engine for Apple Silicon. Optimised for agentic use.☆88May 24, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- MLX Studio - Home of JANG_Q - Image Gen/Edit + Chat/Code All in one - + OpenClaw (Anthropic API)☆930Updated this week
- Exact speculative decoding on Apple Silicon, powered by MLX.☆382Apr 20, 2026Updated 3 months ago
- Own your AI. The native macOS harness for AI agents -- any model, persistent memory, autonomous execution, cryptographic identity. Built …☆7,566Updated this week
- ⚡ Native MLX Swift LLM inference server for Apple Silicon. OpenAI-compatible API, SSD streaming for 100B+ MoE models, TurboQuant KV cache…☆734Updated this week
- vMLX - JANGTQ Uber Compressed MLX Models - L2 Disk Cache (survives restart) + L1 Paged (super fast ttft) + Hybrid SSM Scheduler + Cont B…☆792Updated this week
- Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unslot…☆1,377Jun 23, 2026Updated last month
- MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.☆5,308Updated this week
- DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm☆20,977Updated this week
- Run LLMs with MLX☆6,544Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- DFlash: Block Diffusion for Flash Speculative Decoding☆5,575May 10, 2026Updated 2 months ago
- ⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.☆5,736Updated this week
- The headless browser for AI agents and web scraping☆20,725Updated this week
- A tool for creating and running Linux containers using lightweight virtual machines on a Mac. It is written in Swift, and optimized for A…☆48,760Updated this week
- 🎨 The open-source Claude Design alternative. 🖥️ Local-first desktop app. 🖼️ Your coding agent becomes the design engine: prototypes, l…☆84,517Updated this week
- The agent that grows with you☆227,383Updated this week
- Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.☆21,036Updated this week
- Assign issues to Claude Code, Codex, Cursor, and 17 more coding agents like teammates — open-source and self-hostable.☆44,783Updated this week
- Production-grade engineering skills for AI coding agents.☆84,559Updated this week
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- 🐹 Clean, uninstall, analyze, optimize, and monitor your Mac from the terminal.☆62,598Updated this week
- 🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!☆73,102Updated this week
- A slide framework built for agents.☆6,108Updated this week
- Open-Source Frontier Voice AI☆52,216Jul 24, 2026Updated 2 weeks ago
- A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speec…☆7,694Updated this week
- #1 Persistent memory for AI coding agents based on real-world benchmarks☆26,745Updated this week
- PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.☆28,269Updated this week
- Desktop app to manage markdown knowledge bases☆19,359Updated this week
- Garry's Opinionated OpenClaw/Hermes Agent Brain☆28,002Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning☆35,093Jul 8, 2026Updated last month
- Your Personal AI super intelligence. A brain that builds a local-first memory of your life, a fantastic orchestrator of agent fleets and …☆36,100Updated this week
- Graphs that teach > graphs that impress. Turn any code into an interactive knowledge graph you can explore, search, and ask questions abo…☆77,971Jul 30, 2026Updated last week
- MLX: An array framework for Apple silicon☆27,876Updated this week
- AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI☆85,615Updated this week
- Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.☆13,631Jul 24, 2026Updated 2 weeks ago
- Open source Ghostty-based macOS terminal with vertical tabs and notifications for AI coding agents. Built for multitasking, organization,…☆25,783Updated this week