A Prometheus metrics exporter for NVIDIA DGX Spark clusters.
☆19Feb 16, 2026Updated 5 months ago
Alternatives and similar repositories for dgx-spark-prometheus
Users that are interested in dgx-spark-prometheus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-agent orchestration framework for AI applications - build, deploy, and manage AI agents across the full lifecycle with Forge, Conve…☆33Mar 28, 2026Updated 4 months ago
- Headless 4K remote desktop for the NVIDIA DGX Spark (GB10): one-command installer for Sunshine + Moonlight low-latency game streaming wit…☆46Jun 3, 2026Updated 2 months ago
- ☆18Dec 1, 2025Updated 8 months ago
- Lightweight API Specification for Intelligent Systems☆15Feb 16, 2026Updated 5 months ago
- bf16 LoRA fine-tuning of [Qwen3.5-35B-A3B](https://huggingface.co/unsloth/Qwen3.5-35B-A3B) (a 35B-total / 3B-active Mixture-of-Experts vi…☆16Mar 12, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 3.34× faster inference on Apple Silicon — native MLX port of DFlash speculative decoding☆19Apr 11, 2026Updated 3 months ago
- Some benchmark results of small models and quants that fit on DGX Spark☆49Updated this week
- DocFinder is a local-first indexing and searching documents using semantic embeddings stored in SQLite. Everything runs on your machine, …☆27Updated this week
- Dual-engine (llama.cpp + vLLM) LLM benchmarking pipeline for GGUF & safetensors on NVIDIA GPUs — speed, quality, live dashboard, publisha…☆25Updated this week
- Inspect LLM's logprobs and perplexity over a piece of text, or compare two LLMs (like a git diff)☆19Updated this week
- Protocol for Augmented Memory of Project Artifacts (MCP compatible) - extended☆24Jan 24, 2026Updated 6 months ago
- A robust Python toolkit for converting video/audio content into accurate, multilingual subtitles using WhisperX for transcription and Goo…☆28Dec 2, 2025Updated 8 months ago
- A bytebot variant that uses Holo 1.5 7b to control the desktop☆25Nov 4, 2025Updated 9 months ago
- High-performance batched Top-K selection for CPU inference. Up to 80x faster than PyTorch, optimized for LLM sampling with AVX2 SIMD.☆18Mar 20, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- An Enhanced TOP program to monitor your Nvidia DGX SPARK's Hardware☆34Jan 6, 2026Updated 7 months ago
- 🔱 Modern SNMP management platform - Simulate agents, walk devices, manage traps, browse MIBs. Replace Net-SNMP CLI, iReasoning, snmpsim …☆21May 26, 2026Updated 2 months ago
- ☆39Jul 13, 2026Updated 3 weeks ago
- A MCP stdio toolpack for local LLMs☆34Apr 6, 2026Updated 4 months ago
- 4-5x faster Qwen3.5 on ASUS GX10 / DGX Spark — Hybrid INT4+FP8 + MTP via one shell script☆31Apr 16, 2026Updated 3 months ago
- Tool-calling quality benchmark for LLM serving stacks. 80+ deterministic scenarios testing multi-turn orchestration, safety boundaries, a…☆285Updated this week
- A curated collection of persona-based mcp server & tool groupings.☆38Sep 11, 2025Updated 10 months ago
- An educational Rust project for exporting and running inference on Qwen3 LLM family☆43Aug 3, 2025Updated last year
- A miniaturized version of the Kimi-K2 model optimized for deployment on single H100 GPUs.☆36Jul 16, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- this is an easy way to make ai podcast useing ai loccaly like ollama and the tts of piper☆16Feb 17, 2026Updated 5 months ago
- LLM training & inference in python/C++ with web UI☆37Updated this week
- open-source, local-first AI CLI for developers. It connects to your local Ollama models or remote providers like OpenAI, Anthropic, Gemin…☆61Jul 2, 2026Updated last month
- IPTV_multicast monitoring system.☆11Apr 2, 2023Updated 3 years ago
- An AI assistant for PCs powered by Meta's LLaMA3 using Hugging Face, runs on voice recognition, text-to-speech. Send messages, voice/vide…☆19Jun 6, 2024Updated 2 years ago
- Serve the home! Inference stack for your Nvidia DGX Spark aka the Grace Blackwell AI supercomputer on your desk. Mostly vLLM based for no…☆51Updated this week
- A MCP server for reading and searching ZIM files☆17Jun 6, 2025Updated last year
- Automated parameter sweep pipeline for finding optimal sampling settings for any local LLM on quantized weights☆19Feb 21, 2026Updated 5 months ago
- Performance comparison of Leaflet markercluster and supercluster☆12Apr 17, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Speculative Decoding Implementations: MTP, EAGLE-3, Medusa-1, PARD, Draft Models, N-gram and Suffix Decoding from scratch☆15May 2, 2026Updated 3 months ago
- Kubeconfig Generator is a tool to generate kubeconfig.☆12Jun 28, 2019Updated 7 years ago
- Docker swarm + IPv6 + nftables. Please submit Pull Requests to the GitLab repository. Mirror of☆11Apr 5, 2023Updated 3 years ago
- Zero-instrumentation LLM API and MCP tracer for your agents powered by eBPF — latency, tokens, and tool use in realtime☆18Mar 16, 2026Updated 4 months ago
- A tool to install and configure FreeRADIUS for use with Sonar.☆15Aug 12, 2024Updated last year
- ☆43Aug 2, 2025Updated last year
- npm package template with typescript and tsup☆11Nov 27, 2025Updated 8 months ago