Repo2Run is an LLM-based agent that automates environment configuration by generating error-free Dockerfiles for Python repositories.
☆195Jun 10, 2026Updated last month
Alternatives and similar repositories for Repo2Run
Users that are interested in Repo2Run are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS'25] The official code of "PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning"☆30Mar 30, 2026Updated 3 months ago
- [DL4C @ ICLR 2025] A Benchmark for Automated Environment Setup☆38Nov 9, 2025Updated 8 months ago
- All-in-one benchmarking platform for evaluating LLM.☆15Nov 12, 2025Updated 8 months ago
- SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution☆105Sep 24, 2025Updated 9 months ago
- Advances and Frontiers of LLM-based Issue Resolution in Software Engineering A Comprehensive Survey☆85Apr 22, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [NeurIPS 2025 D&B] 🚀 SWE-bench Goes Live!☆210Jun 11, 2026Updated last month
- CodeRepoQA dataset☆15Feb 19, 2025Updated last year
- tool of llm-based indirect-call analyzer☆31Feb 18, 2025Updated last year
- [COLM 2025] Official repository for R2E-Gym: Procedural Environment Generation and Hybrid Verifiers for Scaling Open-Weights SWE Agents☆308Jul 13, 2025Updated last year
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]☆708Jul 29, 2025Updated 11 months ago
- ☆15Jan 14, 2026Updated 6 months ago
- Automate the build, execution and test of GitHub repositories across programming languages and operating systems.☆124Jun 16, 2026Updated last month
- [FSE'2026] SWE-Factory: Your Automated Factory for Issue Resolution Training Data and Evaluation Benchmarks☆183May 12, 2026Updated 2 months ago
- ☆31Apr 7, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Generating Proof-of-Concept Exploits for npm Vulnerabilities☆18Jun 27, 2026Updated 3 weeks ago
- [NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agents☆710Updated this week
- ☆139May 8, 2025Updated last year
- Implementation and datasets for "Training Language Models to Generate Quality Code with Program Analysis Feedback"☆42Jul 21, 2025Updated last year
- ☆22Jul 16, 2024Updated 2 years ago
- LLM agent to automatically set up arbitrary projects and run their test suites☆73Updated this week
- [ACL25] FEA-Bench: A Benchmark for Evaluating Repository-Level Code Generation for Feature Implementation☆57Jan 28, 2026Updated 5 months ago
- SWE-Bench-plus-plus☆25Feb 5, 2026Updated 5 months ago
- A living collection of frontier research on code agents, from the code we build to the worlds we act in.☆114Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ASE'23] When Less is Enough: Positive-Unlabeled Learning Model for Vulnerability Detection☆16Jan 12, 2024Updated 2 years ago
- A benchmark for evaluating how well AI coding agents can cooperate on software engineering tasks with potential conflicts.☆18Jun 30, 2026Updated 3 weeks ago
- A minimal AI coding agent powered by Anthropic's Claude. Interactive terminal interface with tool execution. Visit https://lldong.github.…☆16Jul 15, 2025Updated last year
- Reproducing R1 for Code with Reliable Rewards☆13Apr 9, 2025Updated last year
- 漏洞规则库是一个致力于帮助开发者识别和避免常见安全漏洞的开源项目。我们收集、整理和分析各类编程语言和常用库中的安全漏洞模式,并提供相应的防范措施和最佳实践。☆39Aug 12, 2025Updated 11 months ago
- Archer: Agentic Code Review for LLVM PRs☆25Jul 13, 2026Updated last week
- Official repository for our paper "FullStack Bench: Evaluating LLMs as Full Stack Coders"☆122May 7, 2025Updated last year
- Parsing-based Analyzer☆78Jun 8, 2025Updated last year
- [NeurIPS 2024] Evaluation harness for SWT-Bench, a benchmark for evaluating LLM repository-level test-generation☆85Apr 28, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ACL2026 Main] AgencyBench: Benchmarking the Frontiers of Autonomous Agents in 1M-Token Real-World Contexts☆89Jan 23, 2026Updated 5 months ago
- Tools and prompt templates used to build and evaluate SWE-rebench-v2 tasks for the paper.☆71Mar 12, 2026Updated 4 months ago
- ☆92Feb 28, 2026Updated 4 months ago
- ☆52Oct 28, 2025Updated 8 months ago
- Framework for evaluating and improving agents☆3,348Updated this week
- SWE-Lego: Pushing the Limits of Supervised Fine-tuning for Software Issue Resolving☆71Feb 28, 2026Updated 4 months ago
- Must-read papers on Repository-level Code Generation & Issue Resolution 🔥☆319Jul 15, 2026Updated last week