Run Qwen3.5-35B-A3B with llama.cpp and openclaw on NVIDIA DGX Spark (GB10)
☆71Mar 1, 2026Updated 6 months ago
Alternatives and similar repositories for Qwen3.5-35B-A3B-openclaw-dgx-spark
Users that are interested in Qwen3.5-35B-A3B-openclaw-dgx-spark are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- antirez/ds4-style hybrid quant DeepSeek V4 Flash on a single DGX Spark via vLLM☆17May 11, 2026Updated 4 months ago
- Complete guide to running Qwen3.5-35B-A3B on NVIDIA DGX Spark (GB10) with vLLM - installation, benchmarks, vision features, and troublesh…☆98Mar 11, 2026Updated 6 months ago
- MiniMax M2 inference server for NVIDIA DGX Spark☆17Jan 24, 2026Updated 8 months ago
- Qwen3.5-122B-A10B on a DGX Spark with DFlash speculative decode. One-shot Docker/vLLM installer. 80+ tok/s!☆60Jun 29, 2026Updated 2 months ago
- vLLM + Qwen3.5-122B-A10B-NVFP4 on NVIDIA DGX Spark (GB10/SM121) — single-GPU NVFP4 W4A4 with MTP speculative decoding, self-contained Doc…☆41Mar 12, 2026Updated 6 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- One-command vLLM installation for NVIDIA DGX Spark with Blackwell GB10 GPUs (sm_121 architecture)☆105Oct 28, 2025Updated 10 months ago
- Some benchmark results of small models and quants that fit on DGX Spark☆50Aug 23, 2026Updated last month
- 4-5x faster Qwen3.5 on ASUS GX10 / DGX Spark — Hybrid INT4+FP8 + MTP via one shell script☆32Apr 16, 2026Updated 5 months ago
- Local setup of Kubernetes using VMware Fusion, ansible and Kubespray☆16Feb 12, 2019Updated 7 years ago
- comfyui optimizations for the dgx spark☆36Apr 30, 2026Updated 4 months ago
- ☆25Aug 3, 2026Updated last month
- vLLM for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆39Aug 25, 2026Updated 3 weeks ago
- ☆74Feb 27, 2026Updated 6 months ago
- Docker configuration for running VLLM on dual DGX Sparks☆2,309Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Collection of step-by-step playbooks for setting up AI/ML workloads on NVIDIA DGX Spark devices with Blackwell architecture.☆1,374Sep 10, 2026Updated 2 weeks ago
- LLM fine-tuning with LoRA + NVFP4/MXFP8 on NVIDIA DGX Spark (Blackwell GB10)☆21Dec 22, 2025Updated 9 months ago
- ☆30Dec 12, 2025Updated 9 months ago
- ☆17Aug 13, 2026Updated last month
- Setup guide for ML training on NVIDIA DGX Spark (GB10 Blackwell, CUDA 13, aarch64)☆182Feb 26, 2026Updated 6 months ago
- Qwen3.5-122B-A10B on DGX Spark: 28.3 → 51 tok/s (+80%)☆315Aug 23, 2026Updated last month
- Sample web application integrating MarkLogic 7 Java Client API with Spring Boot.☆11Jun 16, 2015Updated 11 years ago
- A command-line client for Bluesky☆22May 29, 2026Updated 3 months ago
- Qwen3.6-35B-A3B-heretic NVFP4 + DFlash speculative decoding on DGX Spark (GB10/sm_121a). Source-built vLLM image + 7 patches + comprehens…☆144Jun 28, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆11Mar 23, 2016Updated 10 years ago
- Local benchmarking UI for LLMs and AI agents☆24Apr 13, 2026Updated 5 months ago
- Linux hwmon driver for the NVIDIA DGX Spark (GB10 SoC) that exposes full system power telemetry via standard sensors / sysfs interfaces.☆34Mar 2, 2026Updated 6 months ago
- The DevOps Opportunity: Balancing Security & Velocity in Automation // Ansible Playbooks from October 24th's Webinar☆17May 29, 2019Updated 7 years ago
- NVFP4 Gemma-4 26B-A4B MoE for DGX Spark — optimal recipe: DFlash n=10 (flex) on AEON vLLM Ultimate. 144 tok/s single / 1,724 peak (Coding…☆44Jun 28, 2026Updated 2 months ago
- Dynamic LLM model swapping system with Docker, vLLM integration, and GPU acceleration. Supports GGUF & Hugging Face models with automatic…☆23Mar 6, 2026Updated 6 months ago
- ☆16May 8, 2021Updated 5 years ago
- ☆12Oct 7, 2018Updated 7 years ago
- ComfyUI-KokoroTTS: A text-to-speech model that utilizes the Kokoro TTS framework to convert text into natural-sounding speech. It suppor…☆13Apr 18, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆12Jan 6, 2022Updated 4 years ago
- ☆24Nov 12, 2025Updated 10 months ago
- ChatPDF is a Streamlit app allowing users to query PDF & DOCX content via natural language. It indexes documents for conversational inter…☆17Sep 5, 2023Updated 3 years ago
- ☆19Oct 5, 2025Updated 11 months ago
- Real-time hardware and LLM inference monitoring — GPU, CPU, memory, and vLLM metrics streamed to a dashboard.☆123Sep 10, 2026Updated 2 weeks ago
- Training materials for Operating a Platform: BOSH and Everything Else☆13Nov 3, 2015Updated 10 years ago
- Demonstration of HTTP/2 Push using Jetty Server☆12Apr 5, 2024Updated 2 years ago