Control panel for VLLM, Sglang, llama.cpp, exllamav3
☆1,753Sep 2, 2026Updated this week
Alternatives and similar repositories for local-studio
Users that are interested in local-studio are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Your AI friend right in your browser☆542Updated this week
- extract all your personal data history from cursor, codex, claude-code, windsurf, and trae☆1,272Updated this week
- Ship your repo + live coding-agent session (Claude Code / Codex / pi / Droid) to another machine over Tailscale; it resumes in tmux and k…☆171Updated this week
- Rust & GPUI Native Agent CLIs manager for macOS. Ghostty Terminals + Codex App Features/UX = Ghostex! Embedded browser & IDE. Tons of use…☆755Updated this week
- LLM speculative inference server for heterogeneous hardware & consumer GPUs☆2,832Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Local Responses-API shim that exposes Factory BYOK models (and optional ChatGPT GPT-5.5 passthrough) to Codex Desktop.☆1,065Aug 25, 2026Updated last week
- A vLLM patch + hand‑written SM120 SASS kernels: 2‑bit MoE experts + an FP4 "delta" cache that recovers precision — matching the official …☆538Aug 28, 2026Updated last week
- ☆180Mar 30, 2026Updated 5 months ago
- ☆279Jan 16, 2026Updated 7 months ago
- How much experts do we need to serve a model?☆150Mar 18, 2026Updated 5 months ago
- ☆2,497Updated this week
- Organise your AI's memories with graph database entries☆79Dec 2, 2025Updated 9 months ago
- Turn your PC, Mac, or Linux box into an AI server. LLM inference, chat UI, voice, agents, workflows, RAG, and image generation.☆6,118Updated this week
- Browse the world in the comfort of your terminal☆165Jan 8, 2026Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- DFlash: Block Diffusion for Flash Speculative Decoding☆6,049Aug 18, 2026Updated 2 weeks ago
- Autonomous self-improving 4x DGX Spark (GB10) MoA stack + LoRA loop (DSV4F router, Qwen3.6/Omni/TwoTower/Gemma). Hermes MoA routing, ~90%…☆22Aug 17, 2026Updated 2 weeks ago
- Test LLMs on real tasks. Compare models side-by-side.☆416Aug 10, 2026Updated 3 weeks ago
- ☆24Updated this week
- Native macOS menu bar app to use your Claude Code & ChatGPT subscriptions with AI coding tools - no API keys needed☆3,328Updated this week
- Fully uncensored, capability-enhanced abliteration of Qwen3.6-27B. NVFP4 + z-lab DFlash speculative decoding (n=12) on the unified ghcr.i…☆457Jul 3, 2026Updated 2 months ago
- ☆121May 5, 2026Updated 3 months ago
- The safest, simplest way to manage Hermes from your Mac. Pure SSH. No gateways, no exposed ports, no browser layer.☆2,017Jun 19, 2026Updated 2 months ago
- Agent memory infrastructure: provenance, rollback, lifecycle/supersession, three-layer model (working memory + session archive + wiki). M…☆292May 8, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Native web workspace for Hermes Agent — chat, terminal, memory, skills, inspector.☆6,561Aug 22, 2026Updated last week
- The agent that grows with you☆61Aug 7, 2026Updated 3 weeks ago
- Soul-driven AI agent with permission-hardened tools, token budgets, and multi-channel access. Runs 24/7 from CLI or Telegram.☆3,063Updated this week
- An AI assistant that lives in your browser. Built for collaboration, not autonomy theater. You guide, it executes. Automate repetitive we…☆793Mar 18, 2026Updated 5 months ago
- AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI☆101,758Updated this week
- Autonomous experiment loop extension for pi☆7,990Updated this week
- Mixed-capability LLM benchmark for DGX Spark — 57 scenarios, 10 domains, partial-credit grading, trial statistics☆161Updated this week
- Rust CLI for RBMEM (.rbmem), a structured Rust-Brain memory format with timestamp-protected sections, hierarchy, graph relations, Hermes …☆29Jun 14, 2026Updated 2 months ago
- ☆17May 10, 2026Updated 3 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Experimental llama.cpp fork for inference research and development☆819Updated this week
- The Pi desktop app you want to use.☆352Aug 20, 2026Updated 2 weeks ago
- Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.☆22,186Updated this week
- MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.☆5,469Updated this week
- Educational reference: NVIDIA Blackwell SM100 vs SM120, NVFP4, tcgen05, MoE inference on consumer Blackwell☆26Updated this week
- Own your AI. The native macOS harness for AI agents -- any model, persistent memory, autonomous execution, cryptographic identity. Built …☆7,785Updated this week
- Read-only observability plugin for Hermes Agent: journeys, crossings, guideposts, and reports.☆305May 4, 2026Updated 4 months ago