Agentic Research and Evaluation Suite
☆109Aug 12, 2026Updated this week
Alternatives and similar repositories for ares
Users that are interested in ares are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆235Aug 8, 2026Updated last week
- ☆22Jun 18, 2026Updated last month
- ☆137Mar 31, 2026Updated 4 months ago
- A curated list of awesome Harbor ecosystem projects☆52May 29, 2026Updated 2 months ago
- A compact high-signal benchmark for evaluating frontier agents☆25Aug 3, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Well documented examples of running distributed training jobs on Modal☆29Updated this week
- Trajectory Recording and Capture Environments☆19Jan 24, 2026Updated 6 months ago
- Framework for evaluating and improving agents☆4,243Updated this week
- Samaya AI's FrontierFinance Benchmark Grader☆18Jul 16, 2026Updated 3 weeks ago
- ☆31Nov 14, 2025Updated 9 months ago
- Coco is a proactive co-assistant that connects user workspace with a broader ecosystem of AI agents.☆31Updated this week
- ☆15Dec 12, 2024Updated last year
- ☆15Jun 19, 2026Updated last month
- Public benchmark results from Kernel Arena, a leaderboard for LLM-generated AI accelerator kernels.☆21Mar 11, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Repository for results and data (coming soon!) for ClawsBench☆33Apr 8, 2026Updated 4 months ago
- Data recipes and robust infrastructure for training AI agents☆279Updated this week
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆1,989Updated this week
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆17Mar 11, 2026Updated 5 months ago
- Navi-Bench: benchmarking web agents on everyday tasks directly on real websites☆19Updated this week
- Our library for RL environments + evals☆4,510Updated this week
- Atropos is a Language Model Reinforcement Learning Environments framework for collecting and evaluating LLM trajectories through diverse …☆1,345Jul 4, 2026Updated last month
- Agentic RL Training at Scale☆1,916Updated this week
- Convert GitHub PRs into Harbor tasks☆76Jul 13, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Small, simple agent task environments for training and evaluation☆20Nov 1, 2024Updated last year
- Lexical semantic change detection shared task at SemEval 2020: UiO-UVA team☆16Jan 10, 2023Updated 3 years ago
- ☆25Jan 7, 2026Updated 7 months ago
- OpenTinker is an RL-as-a-Service infrastructure for foundation models☆677Mar 21, 2026Updated 4 months ago
- An LLM agent framework for automated AI interpretability research☆19Apr 17, 2026Updated 3 months ago
- Accompanying Code for "Flipping Coins to Estimate Pseudocounts for Exploration in Reinforcement Learning", ICML 2023☆25Dec 29, 2023Updated 2 years ago
- context-efficient terminal agent powered by an RLM☆60Feb 7, 2026Updated 6 months ago
- Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours☆514Updated this week
- Applying SAEs for fine-grained control☆27Dec 15, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆38May 4, 2026Updated 3 months ago
- Package for calling different models with same interface☆33Jul 21, 2025Updated last year
- Agen is a minimalist language for agent loops and state machines.☆50Mar 30, 2026Updated 4 months ago
- Simple AI CLI that generates docs, unit tests and README.md files☆15Mar 8, 2026Updated 5 months ago
- ☆17Jun 15, 2026Updated 2 months ago
- Continual Learning Bench☆199Jul 19, 2026Updated 3 weeks ago
- Solidity contracts for the decentralized Prime Network protocol☆26Jul 6, 2025Updated last year