fastest runtime for apple silicon.
☆90Apr 16, 2026Updated 5 months ago
Alternatives and similar repositories for bodega-inference-engine
Users that are interested in bodega-inference-engine are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A small Python library and CLI for working with Cloudflare Turnstile on your own pages - read the sitekey, build the token, verify the re…☆124Sep 6, 2026Updated 2 weeks ago
- A small Python library and CLI for working with Cloudflare Turnstile on your own pages - read the sitekey, build the token, verify the re…☆322Sep 15, 2026Updated last week
- axe - a precision agentic coder. large codebases. zero bloat. terminal-native. precise retrieval. powerful inference.☆36Apr 16, 2026Updated 5 months ago
- vMLX - Use MLX models easily - JANGQ (GGUF for MLX) - Not dependant on mlx_vlm☆868Updated this week
- Local LLM benchmark tool for comparing engines (MLX vs llama.cpp), scenarios on Apple Silicon☆25Sep 11, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆38Mar 30, 2026Updated 5 months ago
- Structural code intelligence for LLM coding agents. Agents that understand the architecture finish in fewer turns. Fewer turns = less cac…☆31Apr 5, 2026Updated 5 months ago
- Candy Dungeon Music Forge (CDMF) is a local-first AI music workstation for Windows. It runs on your PC, uses your GPU, and keeps your pro…☆15Dec 17, 2025Updated 9 months ago
- Exact speculative decoding on Apple Silicon, powered by MLX.☆390Apr 20, 2026Updated 5 months ago
- ☆23Mar 6, 2026Updated 6 months ago
- ⚡ Native MLX Swift LLM inference server for Apple Silicon. OpenAI-compatible API, SSD streaming for 100B+ MoE models, TurboQuant KV cache…☆772Updated this week
- ☆230Jun 11, 2026Updated 3 months ago
- Self-hosted meeting transcription portal — speech-to-text, speaker diarization, LLM-corrected transcripts, structured summaries and Word …☆42Aug 26, 2026Updated 3 weeks ago
- Complexity-aware agentic coding pipeline for Claude Code☆36Jul 26, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- MLX Studio - Easiest way to run LLM's on your Mac. All in one engine.☆973Updated this week
- The fastest way to run Qwen 3.8 Flash Next, Qwen 3.8 27B and Ternary Bonsai 2 27B on a Mac: 125 tok/s in OpenCode on an M5 Max, and a 27B…☆2,421Updated this week
- Lossless DFlash speculative decoding for MLX on Apple Silicon☆786Aug 20, 2026Updated last month
- A lightweight transformer language model built from scratch in PyTorch, trained on a single consumer GPU with a full pipeline for data pr…☆42Aug 20, 2026Updated last month
- A browser in the agent's hand. 9 MCP tools, self-diagnosing verdicts, single Go binary.☆30Jul 24, 2026Updated 2 months ago
- Agent Skill for exploring Obsidian vaults with Enzyme — self-contained, cross-agent compatible☆62Updated this week
- Uses google fonts data to generate a minimal list of typeface names, variants, and makes SVGs of each☆18Jan 24, 2026Updated 8 months ago
- Seamlessly interact with CLI coders (Claude, Gemini, Codex) via tmux TUI sessions.☆165Apr 13, 2026Updated 5 months ago
- High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal mode…☆1,591Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- DFlash block-diffusion speculative decoding running on Apple Silicon via MLX, with an ANE execution path that explores heterogeneous acce…☆61Apr 18, 2026Updated 5 months ago
- ComfyUI prompt expansion node☆15Jan 27, 2025Updated last year
- Manage Emacs cursor styles using presets☆15Jul 1, 2026Updated 2 months ago
- mlx-lm server wrapper for agentic harness☆21Jan 26, 2026Updated 7 months ago
- Optimized Ollama LLM server configuration for Mac Studio and other Apple Silicon Macs. Headless setup with automatic startup, resource op…☆356Jan 24, 2026Updated 8 months ago
- Agentic coding harness with persistent memory and a REPL body. Built on Ori Mnemos. Open source must win.☆18May 8, 2026Updated 4 months ago
- Apple Silicon (MLX) port of Karpathy's autoresearch — autonomous AI research loops on Mac, no PyTorch required.☆1,845Jul 2, 2026Updated 2 months ago
- Trails of values over time in SwiftUI☆16Feb 4, 2024Updated 2 years ago
- Flash-MoE sidecar slot-bank runtime for large GGUF MoE models on Apple Silicon — llama.cpp fork☆134Sep 4, 2026Updated 2 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- C.O.N.T.EX.T is designed to compress complex, multi-domain conversations into machine-optimized "Carry-Packets." These packets achieve a …☆32Updated this week
- BrainAPI is a knowledge graph–powered AI memory layer that transforms unstructured data into structured knowledge, enabling intelligent s…☆466Sep 6, 2026Updated 2 weeks ago
- A minimal iOS/macOS app to run LLMs and VLMs on-device with MLX Swift.☆43Mar 5, 2025Updated last year
- CLIP Interrogator, fully in HuggingFace Transformers 🤗, with LongCLIP & CLIP's own words and / or *your* own words!☆19Jul 18, 2025Updated last year
- LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar☆22,180Updated this week
- The bedrock layer for AI coding agents. One governance.md. Any project. Never stale. Universal skills + cross-agent compilation (Claude, …☆40Jul 18, 2026Updated 2 months ago
- A Blueprint-style visual node editor for creating FastMCP servers. Build MCP tools, resources, and prompts by connecting nodes - no codin…☆26Dec 8, 2025Updated 9 months ago