A daily benchmark to regression-test cloud LLMs
☆21Aug 7, 2025Updated last year
Alternatives and similar repositories for daily-bench
Users that are interested in daily-bench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 李鲁鲁老师的 Copilot-Python 学习。和ChatGPT等大语言模型协同进化。☆10Jun 3, 2025Updated last year
- **NOT MAINTAINED** Airbrake.io notifier for Play 2.0☆16Jan 4, 2018Updated 8 years ago
- Examples for QinYan GLMs☆13Sep 3, 2024Updated 2 years ago
- Making AI & LLM APPs components reusable, replaceable, portable, and flexible.☆23Apr 28, 2024Updated 2 years ago
- ☆10Oct 24, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 5 months ago
- ☆13Jun 29, 2024Updated 2 years ago
- Pytorch implementation of our paper accepted by ICML 2024 -- CaM: Cache Merging for Memory-efficient LLMs Inference☆51Jun 19, 2024Updated 2 years ago
- Python package for extractive NLP using the OpenAI API☆17Aug 28, 2024Updated 2 years ago
- This repository will contain projects on multi-agent applications using frameworks such as crewai, langchain, gradio, hugging face etc.☆25Aug 17, 2024Updated 2 years ago
- A benchmarking harness for coding agents.☆17Updated this week
- My Gen AI research☆11Jun 3, 2024Updated 2 years ago
- ☆40Mar 17, 2025Updated last year
- Using modal.com to process FineWeb-edu data☆21Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- alternative way to calculating self attention☆18May 25, 2024Updated 2 years ago
- AI Agents Workshop with Red Hat AI☆13Feb 26, 2025Updated last year
- This is a personal learning repository for the book Hands-On Generative AI with Transformers and Diffusion Models. Here you'll find hand…☆22Jul 27, 2025Updated last year
- A Ray-based LLM server compatible with OpenAI API☆12Mar 12, 2024Updated 2 years ago
- ☆21Apr 27, 2025Updated last year
- Using DSPy to optimize Chat-to-SQL☆15Nov 17, 2025Updated 10 months ago
- General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.☆17Apr 26, 2026Updated 4 months ago
- a set of scripts to easily convert all training data from huggingface into alpaca instruct or sharegpt format, which should allow for eas…☆20Mar 14, 2025Updated last year
- ☆21Nov 28, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Real-time audit log explorer for OpenShift and Kubernetes with AI risk scoring☆18Mar 29, 2026Updated 5 months ago
- various experiments for scaling inference time compute with small reasoning models☆17Jan 16, 2025Updated last year
- ☆20Jan 27, 2024Updated 2 years ago
- ☆22Oct 14, 2024Updated last year
- Universal text artifact optimizer using LLM-powered iterative search☆20Mar 3, 2026Updated 6 months ago
- ☆21Feb 2, 2025Updated last year
- ☆28Aug 1, 2024Updated 2 years ago
- Evalution: evolve your LLMs with better evals.☆16Sep 8, 2026Updated last week
- High-performance late-interaction retrieval engine for on-prem AI. ColBERT/ColPali multi-vector search with Rust fused MaxSim, Triton GPU…☆17Jul 6, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official completion of “Training on the Benchmark Is Not All You Need”.☆40Dec 31, 2024Updated last year
- An advanced AI-powered conversational agent leveraging the Llama 3.2 model and Phidata framework. Features include reasoning, natural lan…☆15Oct 29, 2024Updated last year
- ☆14Jul 15, 2026Updated 2 months ago
- Portfolio REgret for Confidence SEquences☆21Jan 6, 2026Updated 8 months ago
- Core-to-core latency benchmark that works on Apple MacOS without hard affinity☆21May 9, 2026Updated 4 months ago
- Cuda implemenation of flash-kmeans, 2x faster☆23May 8, 2026Updated 4 months ago
- A graph of Reddit☆16Aug 24, 2024Updated 2 years ago