One-command vLLM installation for NVIDIA DGX Spark with Blackwell GB10 GPUs (sm_121 architecture)
☆106Oct 28, 2025Updated 10 months ago
Alternatives and similar repositories for dgx-spark-vllm-setup
Users that are interested in dgx-spark-vllm-setup are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Run vLLM on 1-to-N NVIDIA DGX Spark servers (single Spark, 2 via direct cable, or 3+ via switched fabric) to serve or benchmark LLMs☆129Jun 22, 2026Updated 2 months ago
- LLM fine-tuning with LoRA + NVFP4/MXFP8 on NVIDIA DGX Spark (Blackwell GB10)☆21Dec 22, 2025Updated 8 months ago
- Docker configuration for running VLLM on dual DGX Sparks☆2,263Updated this week
- A dedicated effort to make an optimized, bleeding edge vLLM image using Docker to support DGX comprehensively☆125Feb 22, 2026Updated 6 months ago
- An Enhanced TOP program to monitor your Nvidia DGX SPARK's Hardware☆37Jan 6, 2026Updated 8 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Complete guide to running Qwen3.5-35B-A3B on NVIDIA DGX Spark (GB10) with vLLM - installation, benchmarks, vision features, and troublesh…☆98Mar 11, 2026Updated 6 months ago
- Setup guide for ML training on NVIDIA DGX Spark (GB10 Blackwell, CUDA 13, aarch64)☆181Feb 26, 2026Updated 6 months ago
- Collection of step-by-step playbooks for setting up AI/ML workloads on NVIDIA DGX Spark devices with Blackwell architecture.☆1,353Updated this week
- Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for no…☆51Sep 4, 2026Updated last week
- FlashQLA TileLang GDN kernels ported to NVIDIA Blackwell consumer (GB10 / DGX Spark)☆18Jun 5, 2026Updated 3 months ago
- DGX Spark / GB10 vLLM Docker stack for large-model serving, presets, patches, and validation notes.☆58Updated this week
- antirez/ds4-style hybrid quant DeepSeek V4 Flash on a single DGX Spark via vLLM☆17May 11, 2026Updated 4 months ago
- sparkrun - launch, manage, and stop LLM inference workloads on NVIDIA DGX Spark systems. Live support: https://discord.com/invite/GH5kRg…☆505Updated this week
- Local diagnostic CLI for NVIDIA DGX Spark (GB10). Detects power caps, UMA pressure, thermal risk, CUDA 13/SM_121 wheel mismatches, Docker…☆105Sep 5, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Qwen3.6-35B-A3B-heretic NVFP4 + DFlash speculative decoding on DGX Spark (GB10/sm_121a). Source-built vLLM image + 7 patches + comprehens…☆143Jun 28, 2026Updated 2 months ago
- Pure Rust Inference Engine☆679Updated this week
- ☆74Feb 27, 2026Updated 6 months ago
- DFlash vLLM for DGX Spark — Plug & Play Block-Diffusion Speculative Decoding☆54Jun 28, 2026Updated 2 months ago
- Agentic Workflow☆15Jun 12, 2026Updated 3 months ago
- Codex Lattice: installable Codex skills, hooks, custom agents, MCP search, and docs-first workflow harness.☆19May 18, 2026Updated 3 months ago
- Storybook for flowcharts. Auto-discovers Mermaid diagram files from your codebase, organizes them by category, and renders them in a brow…☆17Mar 9, 2026Updated 6 months ago
- Fully uncensored, capability-enhanced abliteration of Qwen3.6-27B. NVFP4 + z-lab DFlash speculative decoding (n=12) on the unified ghcr.i…☆462Jul 3, 2026Updated 2 months ago
- Production ready, bleeding edge vLLM Docker image for the NVIDIA DGX Spark (GB10 / sm_121a).☆48Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A private AI text summarization tool that runs 100% locally. No data leaves your machine.☆17Jan 18, 2026Updated 7 months ago
- Guide Docs for Harness Engineering☆15Apr 7, 2026Updated 5 months ago
- Manage your self-defined cheat sheets & knowledge base in Alfred☆14Dec 25, 2023Updated 2 years ago
- One dashboard for all your AI assistants☆17Updated this week
- ☆24Feb 14, 2026Updated 6 months ago
- Deploy concurrent Hermes Agent workers on unified-memory GPUs (GB10, DGX Spark) for maximum total tok/s. Profile-isolated, kanban-coordin…☆80Jul 18, 2026Updated last month
- Qwen3.5-122B-A10B on DGX Spark: 28.3 → 51 tok/s (+80%)☆314Aug 23, 2026Updated 3 weeks ago
- Run and manage local Codex workflows from trusted Discord channels.☆20Apr 27, 2026Updated 4 months ago
- Qwen3.5-122B-A10B on a DGX Spark with DFlash speculative decode. One-shot Docker/vLLM installer. 80+ tok/s!☆60Jun 29, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- reTerminalEM is an open-source expansion module platform for reTerminal's high speed interface (PCIe X2, USB 2.0, POE, 2x I2C, 2x SPI, UA…☆23Sep 29, 2025Updated 11 months ago
- YouTube 영상에 AssemblyAI 음성인식 + Claude 번역으로 한글 자막을 생성·오버레이하는 웹 서비스 (에이전트 하네스 기반)☆24Jul 12, 2026Updated 2 months ago
- A Python script that saves your CrewAI agents crew output to a Notion Database☆15Feb 17, 2024Updated 2 years ago
- No code tool for finetuning embedding models☆34Sep 2, 2026Updated last week
- Cloud-independent, self-hosted AI agent runtime — no vendor lock-in. Slack · Telegram · Discord · Web. 9 provider-neutral backends (Claud…☆22Mar 24, 2026Updated 5 months ago
- ⚠️ DEPRECATED - Use openclaw-true-recall-base instead☆30Mar 4, 2026Updated 6 months ago
- ☆24Apr 24, 2026Updated 4 months ago