π₯ The fastest local AI engine for Apple Silicon. Optimised for agentic use.
β98May 24, 2026Updated 3 months ago
Alternatives and similar repositories for lightning-mlx
Users that are interested in lightning-mlx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 3x faster speeds on MLX | Qwen 3.8 27B | Native MTP Speculative Decoding On Apple Silicon With No External Drafter.β2,279Sep 6, 2026Updated last week
- Testing a few local models on ability to understand a research paper and accompanying code.β17May 13, 2026Updated 3 months ago
- The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cacβ¦β3,731Updated this week
- The open-source security layer for AI agents. Deterministic guardrails, PII redaction, and EU AI Act compliance in one line of code.β24May 12, 2026Updated 4 months ago
- A native Mac App for LLM fine-tuning on Apple Silicon β fully on-device, fully open source.β265Aug 26, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- MLX Studio - Easiest way to run LLM's on your Mac. All in one engine.β970Updated this week
- Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 60-95% fewer tokens, same answersβ30May 19, 2026Updated 3 months ago
- JANG β GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Siliconβ225Sep 4, 2026Updated last week
- Lossless DFlash speculative decoding for MLX on Apple Siliconβ779Aug 20, 2026Updated 3 weeks ago
- REAP expert pruning for MoE LLMs on Apple Silicon via MLXβ58Mar 16, 2026Updated 5 months ago
- Moshi-Finetune-MLX lets you fine-tune Moshi (Native, Real-Time, Speech-to-Speech) models all on Apple Silicon.β26Apr 21, 2026Updated 4 months ago
- Virtual piano controls for your MIDI projectsβ12Sep 27, 2025Updated 11 months ago
- Uses conversation history to audit important decisions and changes.β17Jul 13, 2025Updated last year
- LLM inference server with continuous batching & SSD caching for Apple Silicon β managed from the macOS menu barβ21,651Updated this week
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Implements harmful/harmless refusal removal using pure HF Transformersβ25May 8, 2025Updated last year
- Guardrails service for AI agents. Default-deny tool call evaluation with LLM safety analysis, priority-ordered decision matrix, and humanβ¦β22Updated this week
- Enterprise-grade distributed AI agent framework | Develop β Deploy β Observe | K8s-native | Dynamic DI | Auto-failover | Multi-LLM | Pythβ¦β40Updated this week
- Exact speculative decoding on Apple Silicon, powered by MLX.β388Apr 20, 2026Updated 4 months ago
- IP Fabric ChatOpsβ11Sep 6, 2023Updated 3 years ago
- vMLX - Use MLX models easily - JANGQ (GGUF for MLX) - Not dependant on mlx_vlmβ849Updated this week
- β15Jan 8, 2025Updated last year
- π Web search for pi β 12 providers, one tool. Pick what you need.β48Jun 10, 2026Updated 3 months ago
- openpilot is an open source driver assistance system. openpilot performs the functions of Automated Lane Centering and Adaptive Cruise Coβ¦β12Jul 14, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The harness interoperability layer for coding agent. Agent to Agent connect over ACP, A2A and UHPβ44Updated this week
- A tmux plugin that copies words from your pane buffer using flash.nvim style labels.β47Sep 2, 2026Updated last week
- It's an script that helps you to stream video files or video contents to instagram as a live post.β12Aug 17, 2019Updated 7 years ago
- A Kotlin Multiplatform library for emitting a system media sessionβ11May 22, 2024Updated 2 years ago
- Python SDK for Modaicβ28Updated this week
- Behind the scene, project based mark management πͺβ16Dec 1, 2025Updated 9 months ago
- Hand-made Improved Nuclei Templates!πͺ΄β13Jun 12, 2023Updated 3 years ago
- Is a high-performance Augmented Recovery-Generation (RAG) solution based on Redis, Qdrant or PostgreSQL. It offers a high-level interfaceβ¦β30Jan 6, 2026Updated 8 months ago
- β31Sep 4, 2026Updated last week
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Backtesting Library for Strategy Ideationβ15Nov 12, 2024Updated last year
- Connector for transmission-daemon and influxdbβ12Jun 1, 2020Updated 6 years ago
- β16Sep 19, 2020Updated 5 years ago
- Splunk Technology Add-On for pfsenseβ11Mar 30, 2022Updated 4 years ago
- Agent Skill to help convert transformer LLMs to mlx-lmβ51Jun 16, 2026Updated 2 months ago
- python wrapper for the emby rest apiβ15Dec 20, 2021Updated 4 years ago
- The best benchmark for LLMs on Apple's MLX framework knowledge and coding tasks.β38Jun 12, 2026Updated 3 months ago