Measuring frontier coding agents on original, long-horizon engineering tasks
☆1,701Aug 26, 2026Updated 3 weeks ago
Alternatives and similar repositories for deep-swe
Users that are interested in deep-swe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pier is a Harbor fork built for DeepSWE, with stronger support for CLI agents in air-gapped (no-internet) tasks and more faithful, consis…☆192Aug 29, 2026Updated 3 weeks ago
- Can Language Models Rebuild Programs From Scratch?☆928Updated this week
- Framework for evaluating and improving agents☆5,411Updated this week
- FrontierSWE is an ultra long-horizon coding agent benchmark that tests implementation, performance eng and ML research☆230Aug 13, 2026Updated last month
- SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?☆527May 18, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- SWE-Marathon: an ultra long-horizon SWE benchmark☆159Updated this week
- ☆23,063Updated this week
- A benchmark for LLMs on complicated tasks in the terminal☆2,590Jul 11, 2026Updated 2 months ago
- The 100 line AI agent that solves GitHub issues or helps you in your command line. Radically simple, no huge configs, no giant monorepo—b…☆7,782Updated this week
- AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI☆107,258Updated this week
- Lightweight coding agent that runs in your terminal☆125,285Updated this week
- Convert GitHub PRs into Harbor tasks☆83Jul 13, 2026Updated 2 months ago
- open source SWE-Atlas☆72Aug 20, 2026Updated 3 weeks ago
- Verify Precision of all Kimi K2 API Vendor