These are performance benchmarks we did to prepare for our own privacy-preserving and NDA-compliant in-house AI coding assistant. If by any chance, you're a German KMU, and you want strong in-house AI, too, feel free to contact us.
☆32Apr 2, 2025Updated last year
Alternatives and similar repositories for llm-performance-tests
Users that are interested in llm-performance-tests are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Aug 1, 2025Updated last year
- Authenticated independently verifiable agent delegation.☆36Jun 26, 2026Updated last month
- [BROKEN] Zabbix agent (3.0) native plugin for Mongodb monitoring☆11Aug 28, 2018Updated 7 years ago
- ☆13Mar 3, 2023Updated 3 years ago
- Pi extension that tracks bash tool token usage with live stats, grouping, and export☆23Feb 10, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆12May 10, 2021Updated 5 years ago
- Comparing M2M and mT5 on a rare language pairs, blog post: https://medium.com/@abdessalemboukil/comparing-facebooks-m2m-to-mt5-in-low-re…☆16Jun 25, 2021Updated 5 years ago
- ☆12Mar 18, 2024Updated 2 years ago
- ☆24Mar 18, 2025Updated last year
- Docker images for LLM inference (SGLang + vLLM) on NVIDIA Blackwell GPUs (SM120, CUDA 13.2)☆66Updated this week
- How Safe is SF?☆10Aug 20, 2024Updated 2 years ago
- Official Gluon client (iOS, Android, Web)☆17Oct 1, 2018Updated 7 years ago
- Python poker library☆14Sep 9, 2023Updated 2 years ago
- one-click deepfake (face swap)☆10May 30, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Pacific Drive UEVR compatibility plugin☆24Mar 2, 2024Updated 2 years ago
- ☆12Aug 21, 2024Updated 2 years ago
- Cosine Similarity implementation in Nvidia CUDA☆21Sep 28, 2017Updated 8 years ago
- GPT-3 summarizations of new arxiv.org machine learning papers (frontend only)☆12Mar 28, 2023Updated 3 years ago
- WikiGenGPT creates fictional Wikipedia articles using GPT-4 API for imaginative exploration.☆13Nov 5, 2023Updated 2 years ago
- full scripts for highlighted samples in the Getting Started with iControl article series on DevCentral☆13Jun 18, 2016Updated 10 years ago
- fatt tries to find any purl in your project by looking at predefined fields in the supported packages. These fields describe using a purl…☆11Aug 10, 2026Updated 2 weeks ago
- A Github-contributions like visualisation for Strava activities☆17Jan 5, 2023Updated 3 years ago
- Script Execution service☆13Nov 21, 2016Updated 9 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Visa API☆12Apr 29, 2017Updated 9 years ago
- Javascript library which enables the use of drand's randomness in JavaScript-based applications☆16Dec 9, 2019Updated 6 years ago
- AdamW optimizer for bfloat16 models in pytorch 🔥.☆41Jun 16, 2024Updated 2 years ago
- An easy-to-use library to linguistically compare one sentence and its words to another, in the same language or a different one. For inst…☆26Nov 27, 2021Updated 4 years ago
- A simple way to connect your N26 bank account to a Google Sheets cell.☆12Oct 18, 2018Updated 7 years ago
- The Expert Orchestrator AI: Dynamically Adapting, Budget-Aware, and Precisely Tailored to Your Needs☆20Jun 26, 2025Updated last year
- An example repo demonstrating keyless signing with Github Actions☆11May 24, 2022Updated 4 years ago
- Google Container Analysis data import utility, supports OSS vulnerability scanner reports, SLSA provenance and sigstore attestations.☆12Dec 5, 2025Updated 8 months ago
- LLM inference decode throughput benchmark with Rich TUI dashboard. Measures token generation speed across concurrency levels and context …☆76Aug 14, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆11Nov 29, 2017Updated 8 years ago
- A small test for multithreaded C++ stack unwinding on unixes☆16Feb 24, 2020Updated 6 years ago
- Bilinear Pairings Components Library for Delphi☆12Dec 19, 2018Updated 7 years ago
- ⏱️ Run & deploy any Git repo in a single command☆11Dec 27, 2022Updated 3 years ago
- Interact with various LLMs in your browser (LangChain.js, Angular)☆17Updated this week
- ☆13Dec 26, 2022Updated 3 years ago
- A Scheduler for Batched LLM Inference☆19Oct 5, 2025Updated 10 months ago