☆16Jun 27, 2026Updated 3 months ago
Alternatives and similar repositories for spark-auto-round
Users that are interested in spark-auto-round are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18May 31, 2026Updated 4 months ago
- Web UI for sparkrun — launch and monitor inference workloads on NVIDIA DGX Spark☆28Jun 16, 2026Updated 3 months ago
- Some benchmark results of small models and quants that fit on DGX Spark☆51Aug 23, 2026Updated last month
- A monitor of resources for DGX Spark☆25Feb 13, 2026Updated 7 months ago
- Qwen3.5-122B-A10B on a DGX Spark with DFlash speculative decode. One-shot Docker/vLLM installer. 80+ tok/s!☆59Jun 29, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Entrpi/ds4, a Blackwell CUDA perf fork of antirez/ds4 on NVIDIA DGX Spark: one-command install, ~3x upstream prefill, ~1.5x decode, DSpar…☆407Aug 27, 2026Updated last month
- Tool-calling quality benchmark for LLM serving stacks. 80+ deterministic scenarios testing multi-turn orchestration, safety boundaries, a…☆356Updated this week
- Operator-grade GPU monitor for NVIDIA GPUs with native GB10 / DGX Spark coherent UMA support — PSI pressure, clock detection, ConnectX-7 …☆34May 31, 2026Updated 4 months ago
- comfyui optimizations for the dgx spark☆36Apr 30, 2026Updated 5 months ago
- Real-time hardware and LLM inference monitoring — GPU, CPU, memory, and vLLM metrics streamed to a dashboard.☆128Sep 10, 2026Updated 3 weeks ago
- Browser-based model and inference management for NVIDIA DGX Spark - inventory local and Hugging Face models, manage Ollama and LiteLLM, g…☆52Aug 11, 2026Updated last month
- recipe for running Qwen3.8-Flash-Next on a single DGX Spark☆389Updated this week
- FlashQLA TileLang GDN kernels ported to NVIDIA Blackwell consumer (GB10 / DGX Spark)☆18Jun 5, 2026Updated 3 months ago
- Tencent Hunyuan 3 (295B MoE) on 2x NVIDIA DGX Spark: NVFP4 W4A16 + native MTP speculative decoding. First published MTP-on-GB10 numbers, …☆19Jul 13, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- sparkrun - launch, manage, and stop LLM inference workloads on NVIDIA DGX Spark systems. Live support: https://discord.com/invite/GH5kRg…☆527Updated this week
- Multi-model support for group chats☆21Jun 11, 2026Updated 3 months ago
- ☆74Feb 27, 2026Updated 7 months ago
- Mixed-precision quantization for LLMs. Every layer refracts into a different format based on its sensitivity. Native compressed-tensors e…☆102Updated this week
- MCP server and Codex skill for Microsoft FastContext repository exploration☆19Jun 17, 2026Updated 3 months ago
- 🔮 A powerful and stylish Prompt Generator powered by OpenAI and Python. Includes a built-in JSON editor, modular prompt libraries, and f…☆20Jul 12, 2025Updated last year
- MiMo-V2.5 Omni TP=2 on 2x DGX Spark · 1M context · NVFP4 4-bit KV (~1.97M-token KV pool @ 1M, ~30 tok/s) · 69-eval: thinking-OFF 97.8 bea…☆39Jul 13, 2026Updated 2 months ago
- 4-5x faster Qwen3.5 on ASUS GX10 / DGX Spark — Hybrid INT4+FP8 + MTP via one shell script☆32Apr 16, 2026Updated 5 months ago
- Fused BF16 Huffman GEMV Inference kernel☆22Apr 22, 2026Updated 5 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Docker compose serving stack for DeepSeek v4 Flash DSpark for NVIDIA Spark GB10 system using Aidendle94 image☆34Aug 30, 2026Updated last month
- ☆11Jun 21, 2023Updated 3 years ago
- DeepSeek V4 Flash DSpark 1M NVFP4 KV recipe for 2x DGX Spark☆491Sep 10, 2026Updated 3 weeks ago
- 瀏覽器內 HTML 簡報編輯器 · 點擊編輯、拖曳移動、AI 改寫(Claude / Codex CLI)· Claude Design 後續迭代利器 · 單檔 Python 零依賴☆27Aug 5, 2026Updated last month
- LINE Channel for Claude Code - 讓 Claude Code 即時收發 LINE 訊息:LINE Messaging API + MCP server、多用戶配對、遠端權限核准、圖片備份 Google Drive☆26Apr 13, 2026Updated 5 months ago
- Linux hwmon driver for the NVIDIA DGX Spark (GB10 SoC) that exposes full system power telemetry via standard sensors / sysfs interfaces.☆34Mar 2, 2026Updated 7 months ago
- Catch local LLMs spilling into the CPU. One run shows GPU placement, prefill, decode, memory, power and a verdict.☆18Updated this week
- OpenCC for Swift:內建完整字典、locale preset、自訂轉換器與 XML-compatible HTML 轉換的 SwiftPM 實作。☆26Jun 8, 2026Updated 3 months ago
- MCP Server for RSS, Atom, and JSON Feeds☆36Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- High-Resolution Differential Z-Belt Mod for V0 (with optional Kirigami support)☆12May 22, 2022Updated 4 years ago
- Ruuvi Gateway ESP32 code☆29Sep 23, 2026Updated last week
- UniFi MCP server that makes the official UniFi API documentation (Network, Protect, Site Manager, InnerSpace) queryable by AI agents — en…☆38Updated this week
- ☆33Sep 4, 2026Updated last month
- ☆10Jan 22, 2023Updated 3 years ago
- Run Faster-Qwen3-TTS on NVIDIA DGX Spark GB10 (ARM64/SM121/CUDA13) - OpenAI-compatible TTS API with CUDA graph acceleration☆22Sep 21, 2026Updated last week
- A solution for mounting 9mm and 6mm gt2 belts to mgn12 carriage with M2-SHCS or 2mm pin.☆10Jan 28, 2025Updated last year