A curated list of awesome Harbor ecosystem projects
☆48May 29, 2026Updated last month
Alternatives and similar repositories for awesome-harbor
Users that are interested in awesome-harbor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Convert GitHub PRs into Harbor tasks☆72Jul 13, 2026Updated last week
- ☆19Jun 18, 2026Updated last month
- Measuring and evolving with the frontier of agent work☆387Updated this week
- Trajectory Recording and Capture Environments☆19Jan 24, 2026Updated 6 months ago
- SWE-Bench-plus-plus☆25Feb 5, 2026Updated 5 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Terminal-Bench 2.1☆52Updated this week
- 💻 SETA: Scaling Environments for Terminal Agents - Environments☆143Feb 16, 2026Updated 5 months ago
- Framework for evaluating and improving agents☆3,504Updated this week
- ☆36May 16, 2026Updated 2 months ago
- Agentic Research and Evaluation Suite☆107Updated this week
- Dataset of hackable TerminalBench-style tasks and exploit trajectories☆34Apr 18, 2026Updated 3 months ago
- OpenRouter Agent SDK☆19Updated this week
- Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours☆467Updated this week
- Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL training☆71Jul 28, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Benchmark of LLMs on real open-source projects against dependency hell, legacy toolchains, and complex build systems.☆58Jul 14, 2026Updated last week
- Formally proving the security of Fast Reed-Solomon interactive oracle proofs of proximity☆93Dec 11, 2025Updated 7 months ago
- Code for 'Inference Suboptimality in Variational Autoencoders'☆11May 22, 2020Updated 6 years ago
- Simple provider agnostic LLM gateway☆20Updated this week
- A benchmark for evaluating AI agents on realistic business workflows☆149Jul 16, 2026Updated last week
- SWE-Marathon: an ultra long-horizon SWE benchmark☆114Updated this week
- New testbed of interactive SWE tasks for coding agents, set in a realistic multi-turn developer driven environment☆24Jun 30, 2026Updated 3 weeks ago
- Download Web-10K data by querying Bing Image Search☆10Feb 1, 2022Updated 4 years ago
- AI Benchmark for Investment Banking Workflows☆36Jun 30, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆61Apr 7, 2026Updated 3 months ago
- A lightweight computational physics framework, based on the organization of turboWAVE. Implements a "Simulation, PhysicsModule, ComputeTo…☆12Updated this week
- Python package for extractive NLP using the OpenAI API☆17Aug 28, 2024Updated last year
- SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?☆487May 18, 2026Updated 2 months ago
- An agent framework for building and evaluating general digital agents.☆41Apr 21, 2026Updated 3 months ago
- ☆134Mar 31, 2026Updated 3 months ago
- FrontierSmith, a new system that uses AI to synthesize open-ended coding problems at scale☆48May 30, 2026Updated last month
- ☆12Mar 3, 2022Updated 4 years ago
- a benchmark to evaluate the situated inductive reasoning☆16Jan 7, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Tools for developing and optimizing background agents.☆33Jun 16, 2026Updated last month
- ☆13Apr 16, 2025Updated last year
- A Datasette instance for searching WebVid-10M☆15Sep 30, 2022Updated 3 years ago
- ☆14Mar 11, 2024Updated 2 years ago
- Scalable, cloud-native infrastructure for evaluating AI agents across any benchmark.☆27Updated this week
- AI agent benchmark hackability scanner — find evaluation vulnerabilities before they undermine your results☆40May 25, 2026Updated 2 months ago
- ☆11May 20, 2025Updated last year