Start a server from the MLX library.
☆198Jul 26, 2024Updated 2 years ago
Alternatives and similar repositories for mlxserver
Users that are interested in mlxserver are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The easiest way to run the fastest MLX-based LLMs locally☆327Oct 30, 2024Updated last year
- mlx image models for Apple Silicon machines☆100Apr 8, 2026Updated 3 months ago
- FastMLX is a high performance production ready API to host MLX models.☆363Mar 18, 2025Updated last year
- For inferring and serving local LLMs using the MLX framework☆115Mar 24, 2024Updated 2 years ago
- Run large models from the terminal using Apple MLX.☆32Mar 18, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A simple script to enhance text editing across your Mac, leveraging the power of MLX. Designed for seamless integration, it offers real-t…☆109Mar 4, 2024Updated 2 years ago
- Gradio chat interface for FastMLX☆12Sep 22, 2024Updated last year
- SiLLM simplifies the process of training and running Large Language Models (LLMs) on Apple Silicon by leveraging the MLX framework.☆283Jun 16, 2025Updated last year
- A simple UI / Web / Frontend for MLX mlx-lm using Streamlit.☆263Oct 25, 2025Updated 9 months ago
- A multi-platform SwiftUI frontend for running local LLMs with Apple's MLX framework.☆435Oct 27, 2024Updated last year
- MLX Omni Server is a local inference server powered by Apple's MLX framework, specifically designed for Apple Silicon (M-series) chips. I…☆735May 9, 2026Updated 2 months ago
- huggingface chat-ui integration with mlx-lm server☆62Feb 13, 2024Updated 2 years ago
- An example implementation of RLHF (or, more accurately, RLAIF) built on MLX and HuggingFace.☆37Jun 21, 2024Updated 2 years ago
- Minimal, clean code implementation of RAG with mlx using gguf model weights☆52Apr 27, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Explore a simple example of utilizing MLX for RAG application running locally on your Apple Silicon device.☆180Jan 31, 2024Updated 2 years ago
- Transcribe and summarize videos using whisper and llms on apple mlx framework☆80Jan 28, 2024Updated 2 years ago
- An all-in-one LLMs Chat UI for Apple Silicon Mac using MLX Framework.☆1,598Sep 6, 2024Updated last year
- Chat with MLX is a high-performance macOS application that connects your local documents to a personalized large language model (LLM).☆178Mar 8, 2024Updated 2 years ago
- Scripts to create your own moe models using mlx☆92Feb 26, 2024Updated 2 years ago
- Examples in the MLX framework☆8,851Apr 6, 2026Updated 3 months ago
- Large Language Models (LLMs) applications and tools running on Apple Silicon in real-time with Apple MLX.☆465Jan 29, 2025Updated last year
- ☆44Jun 27, 2025Updated last year
- ☆11Aug 25, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- On-device Image Generation for Apple Silicon☆705Apr 11, 2025Updated last year
- Dagger functions to import Hugging Face GGUF models into a local ollama instance and optionally push them to ollama.com.☆120May 22, 2024Updated 2 years ago
- Implementation of nougat that focuses on processing pdf locally.☆85Jan 15, 2025Updated last year
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated 11 months ago
- Distributed Inference for mlx LLm☆102Aug 1, 2024Updated last year
- Generate train.jsonl and valid.jsonl files to use for fine-tuning Mistral and other LLMs.☆97Feb 5, 2024Updated 2 years ago
- ☆14Oct 18, 2023Updated 2 years ago
- A reinforcement learning framework based on MLX.☆260Jul 1, 2026Updated 3 weeks ago
- A python package for serving LLM on OpenAI-compatible API endpoints with prompt caching using MLX.☆103Jun 29, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆284Jun 4, 2024Updated 2 years ago
- ☆20Oct 25, 2025Updated 9 months ago
- tiny_fnc_engine is a minimal python library that provides a flexible engine for calling functions extracted from a LLM.☆37Sep 11, 2024Updated last year
- A simple Python sandbox for helpful LLM data agents☆312Apr 22, 2026Updated 3 months ago
- Examples using MLX Swift☆2,639Updated this week
- MLX Transformers is a library that provides model implementation in MLX. It uses a similar model interface as HuggingFace Transformers an…☆78Mar 23, 2026Updated 4 months ago
- An extremely fast implementation of whisper optimized for Apple Silicon using MLX.☆953May 8, 2024Updated 2 years ago