A Continuous Task Evaluation Playground for AI Harness
☆69Aug 1, 2026Updated last week
Alternatives and similar repositories for SWE-Milestone
Users that are interested in SWE-Milestone are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20Aug 2, 2026Updated last week
- [IPDPS 2024] Adaptive neighbor sampling for temporal GNN☆16Feb 17, 2025Updated last year
- ☆55May 3, 2026Updated 3 months ago
- Code repo for paper: Effective Strategies for Asynchronous Software Engineering Agents☆66Apr 2, 2026Updated 4 months ago
- ☆17Feb 25, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Can Language Models Rebuild Programs From Scratch?☆884Jul 26, 2026Updated 2 weeks ago
- ☆15Sep 25, 2025Updated 10 months ago
- Implementation for "EpiCoder: Encompassing Diversity and Complexity in Code Generation" (ICML 2025)☆28May 16, 2025Updated last year
- [NeurIPS '25] GSO: Challenging Software Optimization Tasks for Evaluating SWE-Agents☆89Jul 12, 2026Updated 3 weeks ago
- ☆19Feb 13, 2026Updated 5 months ago
- [ACL 2025] Graph-guided agentic framework for code localization https://arxiv.org/abs/2503.09089☆624Aug 17, 2025Updated 11 months ago
- Fine-tuned MARL algorithms on SMAC (100% win rates on most scenarios)☆19Aug 20, 2023Updated 2 years ago
- [ASE 2025] CoSIL: Issue Localization via Iteritive Code Graph Searching☆24May 31, 2026Updated 2 months ago
- Must-read papers on Repository-level Code Generation & Issue Resolution 🔥☆325Jul 15, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Mar 5, 2025Updated last year
- SWE-CI: Evaluating Agent Capabilities in Maintaining Codebases via Continuous Integration☆175Jun 10, 2026Updated 2 months ago
- ☆20Apr 8, 2025Updated last year
- ☆20Oct 21, 2025Updated 9 months ago
- Examples of Verbalized Machine Learning (VML)☆16Mar 16, 2025Updated last year
- NetEaseCrowd dataset, a collection of data obtained from You Ling crowdsourcing platform, Fuxi AI Lab, NetEase.☆14Dec 19, 2024Updated last year
- Official code release for "TDFNet: An Efficient Audio-Visual Speech Separation Model with Top-down Fusion", accepted ICIST 2023☆14Mar 17, 2024Updated 2 years ago
- Concise tutorials for distributed training using PyTorch☆10Apr 18, 2023Updated 3 years ago
- SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation. A typed knowledge graph unifies data synthe…☆23Jul 8, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆153May 13, 2026Updated 2 months ago
- ☆95Feb 28, 2026Updated 5 months ago
- SlopCodeBench: Measuring Code Erosion Under Iterative Specification Refinement☆138Updated this week
- ☆19Aug 28, 2025Updated 11 months ago
- [ISSTA'25] A GitHub issue resolution benchmark with multi-aspect diversity in programming languages, repository domains and modality of i…☆17Jun 13, 2025Updated last year
- ☆63Oct 29, 2024Updated last year
- Tools for Ahocoder data processing and evaluation metrics☆15Apr 22, 2024Updated 2 years ago
- ☆25May 8, 2025Updated last year
- Pretty collections of tools for educational data mining.☆11Aug 1, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICML'26] Scaling Long-Horizon LLM Agent via Context-Folding☆179May 18, 2026Updated 2 months ago
- FrontierSWE is an ultra long-horizon coding agent benchmark that tests implementation, performance eng and ML research☆205Updated this week
- Official repo to paper☆12Jan 15, 2023Updated 3 years ago
- ☆49Mar 6, 2026Updated 5 months ago
- Official Repo: AutoResearchBench: Benchmarking AI Agents on Complex Scientific Literature Discovery☆61Apr 24, 2026Updated 3 months ago
- Replication package for ISSTA2023 paper - Towards Efficient Fine-tuning of Pre-trained Code Models: An Experimental Study and Beyond☆23Apr 9, 2023Updated 3 years ago
- 将 有道单词本/不背单词/轻听英语 同步到 墨墨背单词