A living collection of frontier research on code agents, from the code we build to the worlds we act in.
☆133Sep 21, 2026Updated this week
Alternatives and similar repositories for awesome-code-agents
Users that are interested in awesome-code-agents are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains a PyTorch implementation of the ICSE'26 paper "Scrub It Out! Erasing Sensitive Memorization in Code Language Mod…☆30Sep 18, 2025Updated last year
- Must-read papers on Repository-level Code Generation & Issue Resolution 🔥☆333Aug 21, 2026Updated last month
- Reproducing R1 for Code with Reliable Rewards☆318May 5, 2025Updated last year
- LLVM-based Static Analysis Framework☆42Updated this week
- LLM agent to automatically set up arbitrary projects and run their test suites☆73Sep 4, 2026Updated 3 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [NeurIPS'25] EffiBench-X: A Multi-Language Benchmark for Measuring Efficiency of LLM-Generated Code☆17Oct 24, 2025Updated 11 months ago
- Advances and Frontiers of LLM-based Issue Resolution in Software Engineering A Comprehensive Survey☆88Updated this week
- ☆91Jun 19, 2026Updated 3 months ago
- [ICLR 2026] JanusCoder: Towards a Foundational Visual-Programmatic Interface for Code Intelligence☆80May 9, 2026Updated 4 months ago
- Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL training☆74Jul 28, 2025Updated last year
- ☆17Jul 11, 2023Updated 3 years ago
- ☆18May 18, 2025Updated last year
- [ICML25] CODESYNC: Synchronizing Large Language Models with Dynamic Code Evolution at Scale☆24Jul 31, 2025Updated last year
- [ACL25] FEA-Bench: A Benchmark for Evaluating Repository-Level Code Generation for Feature Implementation☆62Jan 28, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [NeurIPS'24] SemCoder: Training Code Language Models with Comprehensive Semantics Reasoning☆32Nov 19, 2024Updated last year
- DiffSearch is a search engine for code changes. The input is a query that describes a code change and the output is a list of matching co…☆20Feb 9, 2024Updated 2 years ago
- Open-science repository on KTH's program repair experiment on the Quixbugs benchmark☆12May 23, 2022Updated 4 years ago
- Based on the R1-Zero method, using rule-based rewards and GRPO on the Code Contests dataset.☆18Apr 22, 2025Updated last year
- RTS technique for .NET☆12Oct 25, 2021Updated 4 years ago
- ☆177May 13, 2026Updated 4 months ago
- A comprehensive code domain benchmark review of LLM researches.☆245Sep 16, 2026Updated last week
- Repo2Run is an LLM-based agent that automates environment configuration by generating error-free Dockerfiles for Python repositories.☆197Jun 10, 2026Updated 3 months ago
- ☆18Mar 28, 2026Updated 5 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Automated Classification of Overfitting Patches with Statically Extracted Code Features http://arxiv.org/pdf/1910.12057☆11Jan 13, 2023Updated 3 years ago
- Enhancing AI Software Engineering with Repository-level Code Graph☆304Apr 1, 2025Updated last year
- ☆23Feb 27, 2025Updated last year
- 南京大学 软件学院 2019高级算法oj练习题答案java版 by wx☆11Dec 9, 2019Updated 6 years ago
- SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution☆105Sep 24, 2025Updated last year
- All-in-one benchmarking platform for evaluating LLM.☆15Nov 12, 2025Updated 10 months ago
- To mitigate position bias in LLMs, especially in long-context scenarios, we scale only one dimension of LLMs, reducing position bias and …☆12Jun 18, 2024Updated 2 years ago
- Official implementation of the paper "From Complex to Simple: Enhancing Multi-Constraint Complex Instruction Following Ability of Large L…☆55Jun 24, 2024Updated 2 years ago
- ☆25Aug 2, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆24Jun 17, 2021Updated 5 years ago
- Structural Reasoning About Program Correctness in Natural Language☆22Jun 23, 2025Updated last year
- ☆21May 30, 2025Updated last year
- Multi-source retrieval and function localization for repository repair☆37Jul 12, 2026Updated 2 months ago
- ☆17Feb 19, 2026Updated 7 months ago
- A Code System for Grammar Error Correction Method. Code Repo for ACL 24 Main "Detection-Correction Structure via General Language Model f…☆24Sep 17, 2024Updated 2 years ago
- ☆12Jul 9, 2025Updated last year