☆17Jun 27, 2026Updated 2 months ago
Alternatives and similar repositories for spark-auto-round
Users that are interested in spark-auto-round are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Web UI for sparkrun — launch and monitor inference workloads on NVIDIA DGX Spark☆25Jun 16, 2026Updated 2 months ago
- A monitor of resources for DGX Spark☆25Feb 13, 2026Updated 7 months ago
- Tool-calling quality benchmark for LLM serving stacks. 80+ deterministic scenarios testing multi-turn orchestration, safety boundaries, a…☆340Updated this week
- Operator-grade GPU monitor for NVIDIA GPUs with native GB10 / DGX Spark coherent UMA support — PSI pressure, clock detection, ConnectX-7 …☆31May 31, 2026Updated 3 months ago
- comfyui optimizations for the dgx spark☆36Apr 30, 2026Updated 4 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Real-time hardware and LLM inference monitoring — GPU, CPU, memory, and vLLM metrics streamed to a dashboard.☆115Updated this week
- ☆16Jul 28, 2026Updated last month
- Browser-based model and inference management for NVIDIA DGX Spark - inventory local and Hugging Face models, manage Ollama and LiteLLM, g…☆44Aug 11, 2026Updated last month
- Home Assistant integration for EcoFlow energy devices: PowerOcean, PowerOcean Plus, Delta 2 Max, Delta 3, Smart Plug, Stream, Stream Micr…☆35Updated this week
- FlashQLA TileLang GDN kernels ported to NVIDIA Blackwell consumer (GB10 / DGX Spark)☆18Jun 5, 2026Updated 3 months ago
- rs3gw (Rust S3 Gateway) is an ultra-lightweight, high-throughput object storage gateway designed for AI model training and scientific com…☆25Updated this week
- sparkrun - launch, manage, and stop LLM inference workloads on NVIDIA DGX Spark systems. Live support: https://discord.com/invite/GH5kRg…☆505Updated this week
- ☆74Feb 27, 2026Updated 6 months ago
- kagent and kMCP website and documentation | Join Discord: https://bit.ly/kagentdiscord☆20Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Mixed-precision quantization for LLMs. Every layer refracts into a different format based on its sensitivity. Native compressed-tensors e…☆104Updated this week
- Docker configuration for running VLLM on dual DGX Sparks☆2,263Updated this week
- 🔮 A powerful and stylish Prompt Generator powered by OpenAI and Python. Includes a built-in JSON editor, modular prompt libraries, and f…☆20Jul 12, 2025Updated last year
- Forecasting extension for OpenBB Platform☆20Jul 19, 2024Updated 2 years ago
- ☆73Apr 1, 2026Updated 5 months ago
- An optimized setup for running ComfyUI on DGX Spark☆50Aug 2, 2026Updated last month
- Recipe: GLM-5.2 (unpruned QuantTrio Int4-Int8Mix) at 655,360-token context + MTP spec decode (DCP4, fp8_ds_mla) on a 4x NVIDIA DGX Spark …☆31Jul 13, 2026Updated 2 months ago
- 4-5x faster Qwen3.5 on ASUS GX10 / DGX Spark — Hybrid INT4+FP8 + MTP via one shell script☆32Apr 16, 2026Updated 4 months ago
- Fused BF16 Huffman GEMV Inference kernel☆22Apr 22, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Local-LLM-first agentic coding assistant, with everything you need out of the box.☆292Updated this week
- Docker compose serving stack for DeepSeek v4 Flash DSpark for NVIDIA Spark GB10 system using Aidendle94 image☆33Aug 30, 2026Updated 2 weeks ago
- ☆11Jun 21, 2023Updated 3 years ago
- 瀏覽器內 HTML 簡報編輯器 · 點擊編輯、拖曳移動、AI 改寫(Claude / Codex CLI)· Claude Design 後續迭代利器 · 單檔 Python 零依賴☆27Aug 5, 2026Updated last month
- Linux hwmon driver for the NVIDIA DGX Spark (GB10 SoC) that exposes full system power telemetry via standard sensors / sysfs interfaces.☆33Mar 2, 2026Updated 6 months ago
- Catch local LLMs spilling into the CPU. One run shows GPU placement, prefill, decode, memory, power and a verdict.☆18Aug 23, 2026Updated 2 weeks ago
- MCP Server for RSS, Atom, and JSON Feeds☆34Updated this week
- High-Resolution Differential Z-Belt Mod for V0 (with optional Kirigami support)☆12May 22, 2022Updated 4 years ago
- UniFi MCP server that makes the official UniFi API documentation (Network, Protect, Site Manager, InnerSpace) queryable by AI agents — en…☆34Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆32Sep 4, 2026Updated last week
- media.ccc.de media library application for webOS. Think of it as a Netflix for hackers☆17Aug 25, 2026Updated 2 weeks ago
- ☆10Jan 22, 2023Updated 3 years ago
- Run Faster-Qwen3-TTS on NVIDIA DGX Spark GB10 (ARM64/SM121/CUDA13) - OpenAI-compatible TTS API with CUDA graph acceleration☆22Aug 26, 2026Updated 2 weeks ago
- A solution for mounting 9mm and 6mm gt2 belts to mgn12 carriage with M2-SHCS or 2mm pin.☆10Jan 28, 2025Updated last year
- This is a repo to compile all the 3d printer modifications☆11Sep 24, 2022Updated 3 years ago
- 🤵 AIfred-Intelligence — self-hosted Multi-Agent Assistant with Debate Modes (Symposion/Tribunal), Voice (STT + Streaming-TTS), RAG with …☆37Updated this week