Autonomous self-improving 4x DGX Spark (GB10) MoA stack + LoRA loop (DSV4F router, Qwen3.6/Omni/TwoTower/Gemma). Hermes MoA routing, ~90% local, self-LoRA. Includes a SINGLE-SPARK recipe (all-Gemma-4: 31B brain + DiffusionGemma + 12B) to build the same on 1 GB10.
☆20Aug 17, 2026Updated last week
Alternatives and similar repositories for Keys-Setup-Autonomous-Self-Improving-Local-Inference-Stack
Users that are interested in Keys-Setup-Autonomous-Self-Improving-Local-Inference-Stack are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Recipe: GLM-5.2 (unpruned QuantTrio Int4-Int8Mix) at 200K ctx with MTP spec decode on a 4x NVIDIA DGX Spark (GB10) cluster☆74Jul 25, 2026Updated 3 weeks ago
- DeepSeek V4 Flash @ 1M token context on 2x NVIDIA DGX Spark — production-tested recipe (45 tok/s decode, real 800K prompts served)☆37Jul 13, 2026Updated last month
- Working recipe to serve DeepSeek-V4-Flash across two NVIDIA DGX Spark (GB10) nodes with vLLM (TP=2, FP8 KV, MTP) over a RoCE/RDMA link — …☆71Jul 13, 2026Updated last month
- Run on TWO-DGX-Spark - vLLm-0.24.0 dual cache optimized DSV4F+DSpark+NVFP4 KV (Concurrency 12 with 1.5M context/3M KV token Pool) >0.58-0…☆21Aug 17, 2026Updated last week
- MiMo-V2.5 Omni TP=2 on 2x DGX Spark · 1M context · NVFP4 4-bit KV (~1.97M-token KV pool @ 1M, ~30 tok/s) · 69-eval: thinking-OFF 97.8 bea…☆39Jul 13, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Recipe: GLM-5.2 (unpruned QuantTrio Int4-Int8Mix) at 655,360-token context + MTP spec decode (DCP4, fp8_ds_mla) on a 4x NVIDIA DGX Spark …☆30Jul 13, 2026Updated last month
- Give any blind LLM eyes — on any machine — in one shot. Tiny local VLM (Qwen3.5-0.8B) as an OpenAI-compatible vision service. Mac (MLX) /…☆29Jul 7, 2026Updated last month
- Run the AEON Bench suite on your own hardware: verified HuggingFace pull → serve → benchmark (text · agentic ×3 harnesses · vision · audi…☆23Aug 16, 2026Updated last week
- Deploy DeepSeek V4 Flash (MoE reasoning model) on dual DGX Spark nodes with 1M token context, InfiniBand, and FP8 KV-cache☆103Updated this week
- Local diagnostic CLI for NVIDIA DGX Spark (GB10). Detects power caps, UMA pressure, thermal risk, CUDA 13/SM_121 wheel mismatches, Docker…☆101Jul 11, 2026Updated last month
- DFlash vLLM for DGX Spark — Plug & Play Block-Diffusion Speculative Decoding☆54Jun 28, 2026Updated last month
- Mixed-capability LLM benchmark for DGX Spark — 57 scenarios, 10 domains, partial-credit grading, trial statistics☆155Updated this week
- Open Source Desktop App to help with Previs for AI-native filmmaking — stage grey-box scenes, choreograph camera & cast with marks, expor…☆112Aug 5, 2026Updated 2 weeks ago
- ☆19Mar 21, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Homebridge plugin for Pentair IntelliCenter☆12Apr 13, 2023Updated 3 years ago
- ☆12Feb 19, 2017Updated 9 years ago
- ☆16Jul 6, 2026Updated last month
- A self-evolving agent in rust with one tool that only uses skills☆31Aug 12, 2026Updated last week
- Swift package for reading and writing Safetensors files.☆13Feb 6, 2026Updated 6 months ago
- Control panel for VLLM, Sglang, llama.cpp, exllamav3☆1,735Updated this week
- ☆67Jun 4, 2026Updated 2 months ago
- Sensirion SVM30 on Arduino, ESPxx, 32U4, Lorawan, DUE☆10Oct 30, 2023Updated 2 years ago
- Serial buffer for Particle Photon and Electron☆13Nov 2, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆15Feb 3, 2026Updated 6 months ago
- ☆16May 21, 2026Updated 3 months ago
- Local benchmarking UI for LLMs and AI agents☆23Apr 13, 2026Updated 4 months ago
- AseeVR fix for Droolon Pi1☆14Mar 16, 2024Updated 2 years ago
- Unified agent memory and context compression stack for 2026 NVIDIA + edge (Vera CPU, Grace, Jetson Thor, 3090). Glues busyBee-cpu, honey-…☆36Updated this week
- If I had my druthers....a programming language made during Khan Academy's 2014 Hack Week.☆15Jan 26, 2015Updated 11 years ago
- ☆23Jun 16, 2026Updated 2 months ago
- A ComfyUI custom node that brings a DAW-style interactive video timeline directly into the node graph. Upload any video, scrub through it…☆18Apr 21, 2026Updated 4 months ago
- AI autoprompter for Ideogram 4's structured JSON caption format — local Qwen3-VL, Ollama or Gemini build the caption, then edit it on a v…☆39Jun 15, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Swift Package Manager package for embedding Python in Swift app☆21Feb 15, 2024Updated 2 years ago
- ☆22Mar 10, 2022Updated 4 years ago
- A Meteor implementation of the Particle Photon SoftAP.☆16Mar 30, 2016Updated 10 years ago
- A minimalistic logging framework for Cocoa☆17Jan 26, 2017Updated 9 years ago
- ☆41Updated this week
- Resources for building agentic apps in Swift☆15Apr 6, 2026Updated 4 months ago
- Control Civilization VI using natural voice commands. You make the strategy — the agent executes it.☆29May 17, 2026Updated 3 months ago