fastest runtime for apple silicon.
☆88Apr 16, 2026Updated 3 months ago
Alternatives and similar repositories for bodega-inference-engine
Users that are interested in bodega-inference-engine are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Compressed KV cache as cross-backend wire format for Metal + CUDA split inference over Thunderbolt 5☆16Apr 14, 2026Updated 3 months ago
- vMLX - JANGTQ Uber Compressed MLX Models - L2 Disk Cache (survives restart) + L1 Paged (super fast ttft) + Hybrid SSM Scheduler + Cont B…☆777Updated this week
- vLLM Metal plugin powered by mlx-swift — high-performance LLM inference on Apple Silicon☆275Jun 3, 2026Updated last month
- ☆36Mar 30, 2026Updated 3 months ago
- Exact speculative decoding on Apple Silicon, powered by MLX.☆379Apr 20, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- AI code quality toolkit — deterministic linter for the AI coding era. 22 detectors, GitHub Action PR gate, zero LLM required.☆55Apr 8, 2026Updated 3 months ago
- ☆23Mar 6, 2026Updated 4 months ago
- ⚡ Native MLX Swift LLM inference server for Apple Silicon. OpenAI-compatible API, SSD streaming for 100B+ MoE models, TurboQuant KV cache…☆725May 19, 2026Updated 2 months ago
- ☆204Jun 11, 2026Updated last month
- A multi-provider AI coding agent with the persona of a Tech-Priest☆18Nov 1, 2025Updated 8 months ago
- Compile time decorator pattern via IL rewriting☆15Apr 19, 2015Updated 11 years ago
- AI-powered personal knowledge vault☆14Jul 15, 2026Updated last week
- MLX Studio - Home of JANG_Q - Image Gen/Edit + Chat/Code All in one - + OpenClaw (Anthropic API)☆915Updated this week
- Lossless DFlash speculative decoding for MLX on Apple Silicon☆753Jun 11, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A browser in the agent's hand. 9 MCP tools, self-diagnosing verdicts, single Go binary.☆27Updated this week
- Chatons is a desktop AI workspace for coding and project workflows: it lets you chat with multiple AI providers, pick scoped or full mod…☆18May 27, 2026Updated last month
- Agent Skill for exploring Obsidian vaults with Enzyme — self-contained, cross-agent compatible☆64Jul 6, 2026Updated 2 weeks ago
- OpenAI and Anthropic compatible server for Apple Silicon. Run LLMs and vision-language models (Llama, Qwen-VL, LLaVA) with continuous bat…☆1,458Jun 28, 2026Updated 3 weeks ago
- DFlash block-diffusion speculative decoding running on Apple Silicon via MLX, with an ANE execution path that explores heterogeneous acce…☆62Apr 18, 2026Updated 3 months ago
- An LLM-native project management system — a self-hosted, lightweight Jira alternative where AI agents are first-class citizens via 23 MCP…☆16Oct 5, 2025Updated 9 months ago
- Moshi-Finetune-MLX lets you fine-tune Moshi (Native, Real-Time, Speech-to-Speech) models all on Apple Silicon.☆24Apr 21, 2026Updated 3 months ago
- Decentralized, encrypted storage for social media data. Built on Blockstack.☆19Jan 4, 2023Updated 3 years ago
- A Python tool that automatically converts OpenAPI(Swagger, ETAPI) compatible specifications into fully functional Model Context Protocol …☆28May 23, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Stop prompting your agents to behave. Start engineering them to.☆25Apr 17, 2026Updated 3 months ago
- mlx-lm server wrapper for agentic harness☆21Jan 26, 2026Updated 6 months ago
- RSS Crawler MCP Server☆19Mar 31, 2025Updated last year
- Agentic coding harness with persistent memory and a REPL body. Built on Ori Mnemos. Open source must win.☆17May 8, 2026Updated 2 months ago
- Apple Silicon (MLX) port of Karpathy's autoresearch — autonomous AI research loops on Mac, no PyTorch required.☆1,749Jul 2, 2026Updated 3 weeks ago
- Trails of values over time in SwiftUI☆16Feb 4, 2024Updated 2 years ago
- BrainAPI is a knowledge graph–powered AI memory layer that transforms unstructured data into structured knowledge, enabling intelligent s…☆236Updated this week
- C.O.N.T.EX.T is designed to compress complex, multi-domain conversations into machine-optimized "Carry-Packets." These packets achieve a …☆30Mar 7, 2026Updated 4 months ago
- RAG + Semantic Search for Apple Notes☆15Feb 27, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Flash-MoE sidecar slot-bank runtime for large GGUF MoE models on Apple Silicon — llama.cpp fork☆117Jul 15, 2026Updated last week
- A minimal iOS/macOS app to run LLMs and VLMs on-device with MLX Swift.☆42Mar 5, 2025Updated last year
- Multi-LoRA inference server for Apple Silicon -- one base model, many adapters, zero reload☆18Apr 13, 2026Updated 3 months ago
- ☆22Jan 22, 2026Updated 6 months ago
- CLIP Interrogator, fully in HuggingFace Transformers 🤗, with LongCLIP & CLIP's own words and / or *your* own words!☆19Jul 18, 2025Updated last year
- LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar☆18,169Updated this week
- A PowerShell module that allows you to easily create a terse re-pave script for a Windows Machine making heavy use of Chocolatey.☆20Apr 10, 2015Updated 11 years ago