π₯ The fastest local AI engine for Apple Silicon. Optimised for agentic use.
β90May 24, 2026Updated 3 months ago
Alternatives and similar repositories for lightning-mlx
Users that are interested in lightning-mlx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 3x faster speeds on MLX | Qwen 3.8 27B | Native MTP Speculative Decoding On Apple Silicon With No External Drafter.β1,598Updated this week
- DFlash block-diffusion speculative decoding running on Apple Silicon via MLX, with an ANE execution path that explores heterogeneous acceβ¦β61Apr 18, 2026Updated 4 months ago
- Easily expose services on your Tailscale network under their own DNS names.β16Updated this week
- The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cacβ¦β3,535Updated this week
- High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal modeβ¦β1,534Updated this week
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A native Mac App for LLM fine-tuning on Apple Silicon β fully on-device, fully open source.β257Jul 16, 2026Updated last month
- MLX Studio - Home of JANG_Q - Image Gen/Edit + Chat/Code All in one - + OpenClaw (Anthropic API)β948Updated this week
- JANG β GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Siliconβ223Aug 16, 2026Updated last week
- Lossless DFlash speculative decoding for MLX on Apple Siliconβ769Updated this week
- REAP expert pruning for MoE LLMs on Apple Silicon via MLXβ58Mar 16, 2026Updated 5 months ago
- Moshi-Finetune-MLX lets you fine-tune Moshi (Native, Real-Time, Speech-to-Speech) models all on Apple Silicon.β26Apr 21, 2026Updated 4 months ago
- A terminal coding agent, and a Python SDK for embedding on-device models in your own apps.β50Updated this week
- Explore your Tailwind config visually directly in your browser while developing.β15Mar 2, 2026Updated 5 months ago
- Uses conversation history to audit important decisions and changes.β17Jul 13, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- LLM inference server with continuous batching & SSD caching for Apple Silicon β managed from the macOS menu barβ20,448Updated this week
- Implements harmful/harmless refusal removal using pure HF Transformersβ25May 8, 2025Updated last year
- Guardrails service for AI agents. Default-deny tool call evaluation with LLM safety analysis, priority-ordered decision matrix, and humanβ¦β21Aug 16, 2026Updated last week
- SuperOptiX: Full Stack Agentic AI Frameworkβ24May 2, 2026Updated 3 months ago
- Enterprise-grade distributed AI agent framework | Develop β Deploy β Observe | K8s-native | Dynamic DI | Auto-failover | Multi-LLM | Pythβ¦β40Updated this week
- NuxtJS module for Nhostβ10Jan 26, 2022Updated 4 years ago
- Exact speculative decoding on Apple Silicon, powered by MLX.β387Apr 20, 2026Updated 4 months ago
- Chatons is a desktop AI workspace for coding and project workflows: it lets you chat with multiple AI providers, pick scoped or full modβ¦β18May 27, 2026Updated 2 months ago
- vMLX - JANGTQ Uber Compressed MLX Models - L2 Disk Cache (survives restart) + L1 Paged (super fast ttft) + Hybrid SSM Scheduler + Cont Bβ¦β826Updated this week
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Pricing and calibration modelsβ13Mar 28, 2025Updated last year
- β15Jan 8, 2025Updated last year
- The harness engineering framework for coding agentsβ38Updated this week
- π Web search for pi β 12 providers, one tool. Pick what you need.β48Jun 10, 2026Updated 2 months ago
- Lightweight Docker image for launching multiple MCP servers via MCPO with unified OpenAPI accessβ40May 26, 2025Updated last year
- β12Jan 2, 2024Updated 2 years ago
- β26Updated this week
- Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agentβ¦β779Updated this week
- Hand-made Improved Nuclei Templates!πͺ΄β13Jun 12, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Is a high-performance Augmented Recovery-Generation (RAG) solution based on Redis, Qdrant or PostgreSQL. It offers a high-level interfaceβ¦β30Jan 6, 2026Updated 7 months ago
- β15Dec 14, 2025Updated 8 months ago
- Comprehensive Python Cheatsheetβ11Oct 12, 2022Updated 3 years ago
- β47Jul 14, 2024Updated 2 years ago
- Code Intelligence Engine β indexes your codebase and gives AI assistants deep understanding via MCP (semantic search, call graphs, 20+ toβ¦β19Feb 14, 2026Updated 6 months ago
- Review Pi agent work, run visible tmux workers, and save evidence outside model contextβ17Aug 10, 2026Updated 2 weeks ago
- AI-powered MCP server that continuously improves Claude Code by learning from high-signal GitHub repositories β auto-evolving developer iβ¦β23Mar 31, 2026Updated 4 months ago