Run Qwen3.5-35B-A3B with llama.cpp and openclaw on NVIDIA DGX Spark (GB10)
☆72Mar 1, 2026Updated 5 months ago
Alternatives and similar repositories for Qwen3.5-35B-A3B-openclaw-dgx-spark
Users that are interested in Qwen3.5-35B-A3B-openclaw-dgx-spark are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FlashQLA TileLang GDN kernels ported to NVIDIA Blackwell consumer (GB10 / DGX Spark)☆17Jun 5, 2026Updated 2 months ago
- Complete guide to running Qwen3.5-35B-A3B on NVIDIA DGX Spark (GB10) with vLLM - installation, benchmarks, vision features, and troublesh…☆97Mar 11, 2026Updated 4 months ago
- Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for no…☆51Updated this week
- MiniMax M2 inference server for NVIDIA DGX Spark☆17Jan 24, 2026Updated 6 months ago
- Qwen3.5-122B-A10B on a DGX Spark with DFlash speculative decode. One-shot Docker/vLLM installer. 80+ tok/s!☆55Jun 29, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- vLLM + Qwen3.5-122B-A10B-NVFP4 on NVIDIA DGX Spark (GB10/SM121) — single-GPU NVFP4 W4A4 with MTP speculative decoding, self-contained Doc…☆40Mar 12, 2026Updated 4 months ago
- One-command vLLM installation for NVIDIA DGX Spark with Blackwell GB10 GPUs (sm_121 architecture)☆105Oct 28, 2025Updated 9 months ago
- ☆20Apr 7, 2026Updated 4 months ago
- Some benchmark results of small models and quants that fit on DGX Spark☆49Updated this week
- An Enhanced TOP program to monitor your Nvidia DGX SPARK's Hardware☆34Jan 6, 2026Updated 7 months ago
- Headless Matrix WebRTC voice AND video agent — auto-answers calls, bridges audio to any AI agent via PipeWire, optional camera-frame visi…☆18Jun 28, 2026Updated last month
- comfyui optimizations for the dgx spark☆33Apr 30, 2026Updated 3 months ago
- vLLM for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆26Jul 2, 2026Updated last month
- ☆72Feb 27, 2026Updated 5 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Docker configuration for running VLLM on dual DGX Sparks☆2,008Updated this week
- Collection of step-by-step playbooks for setting up AI/ML workloads on NVIDIA DGX Spark devices with Blackwell architecture.☆1,246Jul 29, 2026Updated last week
- DGX Spark research and tests - containers, benchmarks, and investigation notes for running models on GB10 (SM 12.1)☆21Updated this week
- LLM fine-tuning with LoRA + NVFP4/MXFP8 on NVIDIA DGX Spark (Blackwell GB10)☆21Dec 22, 2025Updated 7 months ago
- Puppet Certificate Auto-signer for AWS (and other clouds eventually)☆10May 27, 2020Updated 6 years ago
- Qwen3.5-122B-A10B on DGX Spark: 28.3 → 51 tok/s (+80%)☆308Updated this week
- Kubernetes Initializer that injects the Istio sidecar into pods.☆24Jul 13, 2017Updated 9 years ago
- Qwen3.6-35B-A3B-heretic NVFP4 + DFlash speculative decoding on DGX Spark (GB10/sm_121a). Source-built vLLM image + 7 patches + comprehens…☆135Jun 28, 2026Updated last month
- linux allwinner h616 kernel based on the orangepi zero 2 version of the legacy allwinner 4.9 bsp kernel☆10Jan 25, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Local benchmarking UI for LLMs and AI agents☆22Apr 13, 2026Updated 3 months ago
- A dedicated effort to make an optimized, bleeding edge vLLM image using Docker to support DGX comprehensively☆124Feb 22, 2026Updated 5 months ago
- Kubeconfig Generator is a tool to generate kubeconfig.☆12Jun 28, 2019Updated 7 years ago
- modified version of https://automatedhome.party/2017/07/15/wifi-controlled-car-with-a-self-hosted-htmljs-joystick-using-a-wemos-d1-minies…☆13Nov 27, 2017Updated 8 years ago
- Linux hwmon driver for the NVIDIA DGX Spark (GB10 SoC) that exposes full system power telemetry via standard sensors / sysfs interfaces.☆30Mar 2, 2026Updated 5 months ago
- Profile repo — categorized index of NVFP4 model releases, DGX Spark inference stacks, Apple Silicon MLX builds, the voice-AI stack, and t…☆39Jul 9, 2026Updated last month
- NVFP4 Gemma-4 26B-A4B MoE for DGX Spark — optimal recipe: DFlash n=10 (flex) on AEON vLLM Ultimate. 144 tok/s single / 1,724 peak (Coding…☆42Jun 28, 2026Updated last month
- ☆15May 8, 2021Updated 5 years ago
- Dynamic LLM model swapping system with Docker, vLLM integration, and GPU acceleration. Supports GGUF & Hugging Face models with automatic…☆22Mar 6, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- High-performance C++ inference engine for Diffusion Language Models (LLaDA, SEDD, MDLM)☆17Apr 5, 2026Updated 4 months ago
- ☆12Oct 7, 2018Updated 7 years ago
- ComfyUI-KokoroTTS: A text-to-speech model that utilizes the Kokoro TTS framework to convert text into natural-sounding speech. It suppor…☆13Apr 18, 2025Updated last year
- ☆21Apr 3, 2025Updated last year
- A Puppet policy-based autosigner which uses AWS API as validation source☆13Jul 17, 2018Updated 8 years ago
- Real-time hardware and LLM inference monitoring — GPU, CPU, memory, and vLLM metrics streamed to a dashboard.☆89Jul 28, 2026Updated 2 weeks ago
- DTMF, Blue, and US/UK Red Box tone generator for the CardPuter device in Arduino Sketch☆14Mar 1, 2024Updated 2 years ago