A physics-grounded, cost-aware optimization loop for vLLM
☆64Aug 22, 2026Updated this week
Alternatives and similar repositories for profile
Users that are interested in profile are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CLI secret management☆16Aug 13, 2026Updated last week
- Ultra-low-latency, high-throughput multiprocess transport over SHM and mmap. LMAX-Disruptor-style cross-process ring substrate.☆18Aug 6, 2026Updated 2 weeks ago
- A Kubernetes controller and webhook implementation that enables safe, staged rollouts of DaemonSets☆16Jul 31, 2025Updated last year
- The code implementation of HyGRAG, accepted by WWW'26.☆15May 31, 2026Updated 2 months ago
- Drop-in Prometheus / Loki / Tempo HTTP gateway for ClickHouse. Translate PromQL, LogQL, and TraceQL into optimized CH SQL — keep Grafana,…☆63Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- High-performance Rust benchmark client for vLLM serving endpoints.☆52Aug 3, 2026Updated 3 weeks ago
- For a number of years now, work has been proceeding in order to bring to perfection the crudely-conceived idea of a machine that would no…☆14Nov 12, 2025Updated 9 months ago
- Open-source MCP server — progressive tool discovery, code execution, intelligent routing & token optimization across 50+ tools☆15Jun 25, 2026Updated last month
- ☆98Aug 16, 2026Updated last week
- A local MCP server providing tools for exploring code change history and developer insights.☆28Jul 24, 2025Updated last year
- Command buffers for Go. Write automation that works everywhere: local, containers, remote hosts. See https://cmdbuf.io☆25Aug 9, 2026Updated 2 weeks ago
- Archives for Triton Inference Server Practices☆15Feb 28, 2022Updated 4 years ago
- [ACL 2026] WildGraphBench: Benchmarking GraphRAG with Wild-Source Corpora☆17May 11, 2026Updated 3 months ago
- Frontier-grade answers from any mix of models — a local MCP server bringing OpenRouter's Fusion panel architecture to any MCP client. Bri…☆34Aug 8, 2026Updated 2 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- vLLM Qwen3.5-122B NVFP4 on DGX Spark (SM121) — full Docker build with 15 patches☆17Mar 17, 2026Updated 5 months ago
- Tool that slows down I/O operations☆25May 17, 2025Updated last year
- Self-hosted S3 storage browser with cost intelligence, optimization recommendations, and AI-powered analytics. Works with AWS, MinIO, R2,…☆27Jul 14, 2026Updated last month
- FastMemory is a topological representation of text data using concepts as the primary input. It helps in improving the RAG(by replacing e…☆53Jun 8, 2026Updated 2 months ago
- Google Cast protocol v2 implementation for Sming allowing you to control your smart TV or cast device from a microcontroller.☆11Feb 13, 2026Updated 6 months ago
- Lightweight, RESTful microservice for IoT applications☆13Sep 22, 2018Updated 7 years ago
- Website and blog built with Hugo using a custom theme, made by Dann☆20Jul 20, 2026Updated last month
- Enhancing the convergence speed by 2x and improving the training success of Physics-Informed Neural Networks (PINNs).☆13Oct 14, 2024Updated last year
- A tiny PID 1 for containers, written in x86-64 NASM and ARM64 GAS.☆21Feb 23, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- an local MCP server you can run on your env, connecting you to Microsoft Graph, and the complete M365 eco system.☆54Apr 2, 2026Updated 4 months ago
- ☆22Jul 20, 2026Updated last month
- Functions for creating and analyzing word co-occurrence networks in Python and R☆12May 18, 2020Updated 6 years ago
- ☆10Jul 20, 2023Updated 3 years ago
- LLM inference in C/C++, but for GFX906!☆23Updated this week
- Crow is a Desktop AI Assistant☆33Aug 9, 2024Updated 2 years ago
- An associative memory you can run anywhere. Write facts in plain language, recall them by meaning. No tables, no schema, no embeddings, n…☆72Aug 15, 2026Updated last week
- CatRAG is a RAG framework builds on the HippoRAG 2 architecture and transforms the static KG into query-adaptive navigation structure. RA…☆29Updated this week
- Plugin based Mesh monitoring bridge for Meshtastic; including APRS-IS support, message logging, prometheus exporting and various other in…☆14Apr 6, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- SNDR Core Engine (Genesis) — vLLM runtime patch-overlay for Qwen3.6 + Gemma4 on consumer NVIDIA (Ampere sm_86, 2× A5000/3090). Qwen3.6-35…☆132Updated this week
- Single-file installer for a club-3090 webserver providing an admin control panel, a reverse proxy that automatically routes requests to t…☆24Jul 8, 2026Updated last month
- A database of knowledge around inference & training on GFX906 GPUs https://skyne98.github.io/wiki-gfx906/☆16Feb 21, 2026Updated 6 months ago
- RAG-QA Leaderboard☆28Jan 27, 2026Updated 6 months ago
- async compile on neovim☆12Aug 17, 2023Updated 3 years ago
- A tool to post-process json trace files for IBM-AIU performance analysis. It enhances the traces with additional statistics extracted fro…☆13Updated this week
- Governed memory runtime for AI assistants: policy-before-storage, context admission, memory usage trace, deletion proof, leakage evals, a…☆18Updated this week