A curated list of awesome Harbor ecosystem projects
☆52May 29, 2026Updated 3 months ago
Alternatives and similar repositories for awesome-harbor
Users that are interested in awesome-harbor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Convert GitHub PRs into Harbor tasks☆82Jul 13, 2026Updated last month
- ☆23Jun 18, 2026Updated 2 months ago
- Terminal-Bench-Science: Evaluating AI agents on research workflows across scientific domains☆514Updated this week
- Trajectory Recording and Capture Environments☆19Jan 24, 2026Updated 7 months ago
- 💻 SETA: Scaling Environments for Terminal Agents - Environments☆145Feb 16, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- NLSpec instruction following benchmark for https://factory.strongdm.ai/products/attractor☆21Feb 26, 2026Updated 6 months ago
- ☆397Apr 30, 2026Updated 4 months ago
- Framework for evaluating and improving agents☆4,952Updated this week
- ☆41May 16, 2026Updated 3 months ago
- Agentic Research and Evaluation Suite☆111Aug 12, 2026Updated 3 weeks ago
- A benchmark for evaluating AI agents on realistic business workflows☆259Aug 4, 2026Updated last month
- Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours☆546Updated this week
- Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL training☆74Jul 28, 2025Updated last year
- Realistic examples of building evals and optimizing agents with Harbor☆200Apr 23, 2026Updated 4 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Benchmark of LLMs on real open-source projects against dependency hell, legacy toolchains, and complex build systems.☆59Jul 14, 2026Updated last month
- SWE-Marathon: an ultra long-horizon SWE benchmark☆149Updated this week
- ICML'20: SIGUA: Forgetting May Make Learning with Noisy Labels More Robust☆17Dec 14, 2020Updated 5 years ago
- New testbed of interactive SWE tasks for coding agents, set in a realistic multi-turn developer driven environment☆26Jun 30, 2026Updated 2 months ago
- ☆39Apr 17, 2024Updated 2 years ago
- Download Web-10K data by querying Bing Image Search☆10Feb 1, 2022Updated 4 years ago
- Sandboxed code execution for AI agents, locally or on the cloud. Massively parallel, easy to extend. Powering SWE-agent and more.☆582Updated this week
- SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?☆518May 18, 2026Updated 3 months ago
- Data recipes and robust infrastructure for training AI agents☆286Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Training terminal-agents☆297Updated this week
- Benchmarking Open-Ended Inference Optimization by AI Agents☆43Jul 6, 2026Updated last month
- An agent framework for building and evaluating general digital agents.☆42Apr 21, 2026Updated 4 months ago
- ☆142Mar 31, 2026Updated 5 months ago
- FrontierSmith, a new system that uses AI to synthesize open-ended coding problems at scale☆52May 30, 2026Updated 3 months ago
- ☆12Mar 3, 2022Updated 4 years ago
- a benchmark to evaluate the situated inductive reasoning☆19Jan 7, 2025Updated last year
- Tools for developing and optimizing background agents.☆33Jun 16, 2026Updated 2 months ago
- ☆13Apr 16, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A Datasette instance for searching WebVid-10M☆15Sep 30, 2022Updated 3 years ago
- Scalable, cloud-native infrastructure for evaluating AI agents across any benchmark.☆34Updated this week
- AI agent benchmark hackability scanner — find evaluation vulnerabilities before they undermine your results☆45May 25, 2026Updated 3 months ago
- ☆11May 20, 2025Updated last year
- ☆17Nov 7, 2023Updated 2 years ago
- A research framework for evaluating proactive AI assistants through active user simulation☆39May 23, 2026Updated 3 months ago
- My productivity workstation.☆26Jan 30, 2026Updated 7 months ago