A Prometheus metrics exporter for NVIDIA DGX Spark clusters.
☆18Feb 16, 2026Updated 5 months ago
Alternatives and similar repositories for dgx-spark-prometheus
Users that are interested in dgx-spark-prometheus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-agent orchestration framework for AI applications - build, deploy, and manage AI agents across the full lifecycle with Forge, Conve…☆33Mar 28, 2026Updated 3 months ago
- Headless 4K remote desktop for the NVIDIA DGX Spark (GB10): one-command installer for Sunshine + Moonlight low-latency game streaming wit…☆43Jun 3, 2026Updated last month
- ☆18Dec 1, 2025Updated 7 months ago
- bf16 LoRA fine-tuning of [Qwen3.5-35B-A3B](https://huggingface.co/unsloth/Qwen3.5-35B-A3B) (a 35B-total / 3B-active Mixture-of-Experts vi…☆15Mar 12, 2026Updated 4 months ago
- Lightweight API Specification for Intelligent Systems☆15Feb 16, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Some benchmark results of small models and quants that fit on DGX Spark☆47Updated this week
- ☆20Jan 3, 2026Updated 6 months ago
- Dual-engine (llama.cpp + vLLM) LLM benchmarking pipeline for GGUF & safetensors on NVIDIA GPUs — speed, quality, live dashboard, publisha…☆24Updated this week
- Inspect LLM's logprobs and perplexity over a piece of text, or compare two LLMs (like a git diff)☆20Mar 23, 2026Updated 3 months ago
- A miniaturized version of the Kimi-K2 model optimized for deployment on single H100 GPUs.☆35Jul 16, 2025Updated last year
- A Chrome extension that enables virtual fashion try-on and model swap using FASHN AI. Hover over fashion images on any website to: (1) tr…☆22Aug 14, 2025Updated 11 months ago
- Protocol for Augmented Memory of Project Artifacts (MCP compatible) - extended☆25Jan 24, 2026Updated 5 months ago
- Explore a wide range of computer vision projects and documentation covering everything from object detection, image segmentation, and tra…☆12Sep 9, 2025Updated 10 months ago
- A robust Python toolkit for converting video/audio content into accurate, multilingual subtitles using WhisperX for transcription and Goo…☆28Dec 2, 2025Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A bytebot variant that uses Holo 1.5 7b to control the desktop☆25Nov 4, 2025Updated 8 months ago
- High-performance batched Top-K selection for CPU inference. Up to 80x faster than PyTorch, optimized for LLM sampling with AVX2 SIMD.☆17Mar 20, 2026Updated 4 months ago
- An Enhanced TOP program to monitor your Nvidia DGX SPARK's Hardware☆34Jan 6, 2026Updated 6 months ago
- A quick-and-dirty load-tester for any OpenAI-style LLM API endpoint☆20May 29, 2025Updated last year
- ☆35Updated this week
- 4-5x faster Qwen3.5 on ASUS GX10 / DGX Spark — Hybrid INT4+FP8 + MTP via one shell script☆31Apr 16, 2026Updated 3 months ago
- A MCP stdio toolpack for local LLMs☆33Apr 6, 2026Updated 3 months ago
- A curated collection of persona-based mcp server & tool groupings.☆37Sep 11, 2025Updated 10 months ago
- An educational Rust project for exporting and running inference on Qwen3 LLM family☆44Aug 3, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- this is an easy way to make ai podcast useing ai loccaly like ollama and the tts of piper☆16Feb 17, 2026Updated 5 months ago
- A MCP server for reading and searching ZIM files☆17Jun 6, 2025Updated last year
- The Unreasonable Effectiveness of Synthetic Data☆16Mar 29, 2023Updated 3 years ago
- LLM training & inference in python/C++ with web UI☆37Updated this week
- sparkrun - launch, manage, and stop LLM inference workloads on NVIDIA DGX Spark systems☆395Updated this week
- open-source, local-first AI CLI for developers. It connects to your local Ollama models or remote providers like OpenAI, Anthropic, Gemin…☆60Jul 2, 2026Updated 2 weeks ago
- An AI assistant for PCs powered by Meta's LLaMA3 using Hugging Face, runs on voice recognition, text-to-speech. Send messages, voice/vide…☆19Jun 6, 2024Updated 2 years ago
- GPU-accelerated voice assistant — local LLM, fine-tuned Whisper, Kokoro TTS, AMD ROCm☆16Apr 2, 2026Updated 3 months ago
- Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for no…☆49Jul 12, 2026Updated last week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Speculative Decoding Implementations: MTP, EAGLE-3, Medusa-1, PARD, Draft Models, N-gram and Suffix Decoding from scratch☆15May 2, 2026Updated 2 months ago
- Zero-instrumentation LLM API and MCP tracer for your agents powered by eBPF — latency, tokens, and tool use in realtime☆18Mar 16, 2026Updated 4 months ago
- ☆43Aug 2, 2025Updated 11 months ago
- npm package template with typescript and tsup☆11Nov 27, 2025Updated 7 months ago
- Production-grade agent orchestration for Claude Code - 11 agents, 46 MCP tools, SQLite+FTS5, drift detection, consensus checkpoints☆52Jun 8, 2026Updated last month
- Run vLLM on 1-to-N NVIDIA DGX Spark servers (single Spark, 2 via direct cable, or 3+ via switched fabric) to serve or benchmark LLMs☆122Jun 22, 2026Updated 3 weeks ago
- For the better CI as well as CD using gogs and drone base on kubernetes☆10Jul 31, 2021Updated 4 years ago