Live dashboard: fire N parallel streaming coding-agent runs at any OpenAI-compatible endpoint — per-run TTFT/tok-s/E2E, real kill switch, per-GPU metrics, shareable summary. Zero deps.
☆27Jul 23, 2026Updated 2 months ago
Alternatives and similar repositories for 2Wild-Coding-Agent-Latency-Monitor
Users that are interested in 2Wild-Coding-Agent-Latency-Monitor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tencent Hunyuan 3 (295B MoE) on 2x NVIDIA DGX Spark: NVFP4 W4A16 + native MTP speculative decoding. First published MTP-on-GB10 numbers, …☆19Jul 13, 2026Updated 2 months ago
- DeepSeek V4 Flash DSpark 1M NVFP4 KV recipe for 2x DGX Spark☆490Sep 10, 2026Updated 3 weeks ago
- Production-ready vLLM deployment wrapper for Qwen3.6-27B (NVFP4) — self-hosted OpenAI-compatible inference☆61Jul 30, 2026Updated 2 months ago
- Mixed-capability LLM benchmark for DGX Spark — 57 scenarios, 10 domains, partial-credit grading, trial statistics☆164Aug 29, 2026Updated last month
- DeepSeek-V4-Flash-DSpark abliterated (uncensored) · ~100% refusal bypass · C1 ~57 tok/s · 1M ctx · 2× DGX Spark · HF weights☆51Aug 17, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Run on TWO-DGX-Spark - vLLm-0.24.0 dual cache optimized DSV4F+DSpark+NVFP4 KV (Concurrency 12 with 1.5M context/3M KV token Pool) >0.58-0…☆21Aug 17, 2026Updated last month
- Qwen3.8-Flash-Next (NVFP4) on DGX Spark in vLLM: one Spark 43.9 tok/s with our disk-backed n-gram table patch, staged gather and reduced-…☆126Sep 7, 2026Updated 3 weeks ago
- Deploy DeepSeek V4 Flash (MoE reasoning model) on dual DGX Spark nodes with 1M token context, InfiniBand, and FP8 KV-cache☆107Aug 20, 2026Updated last month
- vLLM deployment for Unsloth Qwen3.6-35B-A3B-NVFP4-Fast on NVIDIA DGX Spark☆70Jul 29, 2026Updated 2 months ago
- DeepSeek-V4.1-Flash (552B MoE, MXFP4 experts, 1M ctx) on four NVIDIA DGX Sparks with vLLM TP4: Engram-on-disk patch, sm121 kernel build, …☆89Sep 19, 2026Updated 2 weeks ago
- ☆19May 10, 2026Updated 4 months ago
- Run the AEON Bench suite on your own hardware: verified HuggingFace pull → serve → benchmark (text · agentic ×3 harnesses · vision · audi…☆27Sep 11, 2026Updated 3 weeks ago
- llama-server start/stop scripts for Qwen3.6-35B-A3B UD-Q8_K_XL GGUF on DGX Spark☆27Jul 6, 2026Updated 2 months ago
- MiMo-V2.5 Omni TP=2 on 2x DGX Spark · 1M context · NVFP4 4-bit KV (~1.97M-token KV pool @ 1M, ~30 tok/s) · 69-eval: thinking-OFF 97.8 bea…☆39Jul 13, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Algorithmic trading using IBKR Python API☆27Mar 1, 2025Updated last year
- ☆20Mar 29, 2026Updated 6 months ago
- AMD APU compatible Ollama. Get up and running with OpenAI gpt-oss, DeepSeek-R1, Gemma 3 and other models.☆16Aug 18, 2026Updated last month
- IB TWS API examples☆29Aug 22, 2026Updated last month
- ☆39Jul 17, 2026Updated 2 months ago
- DeepSeek V4 Flash @ 1M token context on 2x NVIDIA DGX Spark — production-tested recipe (45 tok/s decode, real 800K prompts served)☆36Jul 13, 2026Updated 2 months ago
- GLM-5.3 Flash EXL3 for 2-4x DGX Sparks☆691Updated this week
- A lightweight tool to optimize your C# project for LLM context windows by using a knowledge graph | Code structure visualization | Static…☆36Dec 3, 2024Updated last year
- Give any blind LLM eyes — on any machine — in one shot. Tiny local VLM (Qwen3.5-0.8B) as an OpenAI-compatible vision service. Mac (MLX) /…☆28Jul 7, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A clone of debian's package so I can submit patches☆22Feb 19, 2015Updated 11 years ago
- AEON vLLM Ultimate — vLLM 0.27.1 built from source for DGX Spark / Blackwell (sm_121a/GB10). DSpark quantized Markov heads, DFlash SWA on…☆143Sep 12, 2026Updated 3 weeks ago
- DeepSeek-v4-Flash 0731 recipe for 2x DGX Sparks☆1,439Sep 26, 2026Updated last week
- A Rust CLI tool that transforms your project files into perfectly formatted context blocks for Large Language Models. Ideal for code revi…☆13May 11, 2025Updated last year
- Dual-engine (llama.cpp + vLLM) LLM benchmarking pipeline for GGUF & safetensors on NVIDIA GPUs — speed, quality, live dashboard, publisha…☆40Updated this week
- One-click AI agent setup for NVIDIA DGX Spark, Jetson, and RTX hardware. OpenClaw + Ollama, fully local.☆28Apr 16, 2026Updated 5 months ago
- LLM inference decode throughput benchmark with Rich TUI dashboard. Measures token generation speed across concurrency levels and context …☆107Updated this week
- ConfigMgr LogFiler Opener automates the usage of CMTrace, CMLogViewer and OneTrace for opening single or multiple ConfigMgr Client Logfil…☆13May 22, 2023Updated 3 years ago
- Pure Rust Inference Engine☆700Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code & data helping to reverse engineer the Eora3D and produce an open source application that can use it☆37Dec 18, 2020Updated 5 years ago
- Browser-based model and inference management for NVIDIA DGX Spark - inventory local and Hugging Face models, manage Ollama and LiteLLM, g…☆52Aug 11, 2026Updated last month
- Tools to speed up migration from other SSE solutions to Microsoft's Global Secure Access☆20Sep 24, 2026Updated last week
- 🐣🕐📅 A simple utility to draft scheduling emails.☆12Sep 13, 2023Updated 3 years ago
- ☆19Feb 24, 2025Updated last year
- SUPERCHARGE AI-assisted development by using Git. Cross-model review gates, evidence-linked closure, verification profiles, model-tier ro…☆55Apr 22, 2026Updated 5 months ago
- A real estate agent CRM for managing clients' properties and profiles☆15May 6, 2024Updated 2 years ago