local-inference-lab / llm-inference-benchView on GitHub
LLM inference decode throughput benchmark with Rich TUI dashboard. Measures token generation speed across concurrency levels and context lengths. Supports SGLang and vLLM engines.
62Jul 11, 2026Updated 3 weeks ago

Alternatives and similar repositories for llm-inference-bench

Users that are interested in llm-inference-bench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

Are these results useful?