LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
☆21,834Sep 17, 2026Updated this week
Alternatives and similar repositories for omlx
Users that are interested in omlx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The fastest way to run Qwen 3.8 Flash Next and Qwen 3.8 27B on a Mac: 125 tok/s in OpenCode on an M5 Max. Native MTP speculative decoding…☆2,347Updated this week
- Run LLMs with MLX☆7,040Updated this week
- MLX Studio - Easiest way to run LLM's on your Mac. All in one engine.☆975Updated this week
- The agent that grows with you☆246,441Updated this week
- Lossless DFlash speculative decoding for MLX on Apple Silicon☆784Aug 20, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.☆5,507Updated this week
- Own your AI. The native macOS harness for AI agents -- any model, persistent memory, autonomous execution, cryptographic identity. Built …☆7,955Updated this week
- vMLX - Use MLX models easily - JANGQ (GGUF for MLX) - Not dependant on mlx_vlm