REAP expert pruning for MoE LLMs on Apple Silicon via MLX
☆58Mar 16, 2026Updated 6 months ago
Alternatives and similar repositories for reap-mlx
Users that are interested in reap-mlx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The best benchmark for LLMs on Apple's MLX framework knowledge and coding tasks.☆38Jun 12, 2026Updated 3 months ago
- Minimalist repo to do mlx vlm constrained decoding in batch mode☆19Apr 11, 2026Updated 5 months ago
- DFlash block-diffusion speculative decoding running on Apple Silicon via MLX, with an ANE execution path that explores heterogeneous acce…☆61Apr 18, 2026Updated 5 months ago
- Train Embedding Models on MLX.☆17Jun 2, 2026Updated 3 months ago
- Moshi-Finetune-MLX lets you fine-tune Moshi (Native, Real-Time, Speech-to-Speech) models all on Apple Silicon.☆26Apr 21, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Triton‑style kernel toolkit for MLX plus a small upstream incubator: prototype, benchmark, and upstream fusions for Apple Silicon☆53Mar 31, 2026Updated 5 months ago
- An MLX port of Meta's Coconut reasoning model☆16Sep 2, 2025Updated last year
- Find the hidden meaning of LLMs☆42Nov 13, 2025Updated 10 months ago
- Minimal Claude Code alternative powered by MLX☆47Jan 11, 2026Updated 8 months ago
- ☆223Mar 24, 2026Updated 5 months ago
- ☆18May 27, 2025Updated last year
- ollama like cli tool for MLX models on huggingface (pull, rm, list, show, serve etc.)☆152Sep 11, 2026Updated last week
- this repo has all official MLX-LM-LoRA example notebooks for training on Apple Silicon☆42Updated this week
- ☆17Sep 1, 2026Updated 2 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Automated hyperparameter search for optimal Gabliteration configurations on large language models☆49Mar 10, 2026Updated 6 months ago
- How much experts do we need to serve a model?☆151Mar 18, 2026Updated 6 months ago
- Mini AI Developer☆20Mar 17, 2026Updated 6 months ago
- ☆182Mar 30, 2026Updated 5 months ago
- Flash-MoE sidecar slot-bank runtime for large GGUF MoE models on Apple Silicon — llama.cpp fork☆134Sep 4, 2026Updated 2 weeks ago
- MLX Implementation of Recursive Reasoning with Tiny Networks☆79Oct 11, 2025Updated 11 months ago
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 5 months ago
- A native Mac App for LLM fine-tuning on Apple Silicon — fully on-device, fully open source.☆267Aug 26, 2026Updated 3 weeks ago
- Exploratory stuff on RLMs for video workflows☆36Apr 20, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- MLX-Video is the best package for inference and finetuning of Image-Video-Audio generation models on your Mac using MLX.☆301May 13, 2026Updated 4 months ago
- Lossless DFlash speculative decoding for MLX on Apple Silicon☆784Aug 20, 2026Updated 3 weeks ago
- This repo maintains a 'cheat sheet' for LLMs that are undertrained on mlx☆33Mar 12, 2026Updated 6 months ago
- The ultimate training toolkit for finetuning diffusion models☆34Jan 22, 2026Updated 7 months ago
- Running a big model on a small laptop☆57Mar 28, 2026Updated 5 months ago
- ☆50Aug 11, 2026Updated last month
- CLI for Recursive Language Models (arXiv:2512.24601)☆210Jun 17, 2026Updated 3 months ago
- Exact speculative decoding on Apple Silicon, powered by MLX.☆390Apr 20, 2026Updated 4 months ago
- MLX binary vectors and associated algorithms.☆14Mar 13, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- mlx-lm server wrapper for agentic harness☆21Jan 26, 2026Updated 7 months ago
- MLX-Embeddings is the best package for running Vision and Language Embedding models locally on your Mac using MLX.☆444May 13, 2026Updated 4 months ago
- Train Large Language Models on MLX.☆414Updated this week
- On-device semantic search over Apple WWDC 2025 docs using MLX embeddings — SwiftUI app (WWDC OMT 2025)☆76Jun 12, 2025Updated last year
- ☆13Jun 29, 2024Updated 2 years ago
- alternative way to calculating self attention☆18May 25, 2024Updated 2 years ago
- Native Mac OS GUI for Using mlx-lm-lora.☆66Dec 19, 2025Updated 8 months ago