PatchEval: A New Benchmark for Evaluating LLMs on Patching Real-World Vulnerabilities
☆230Aug 31, 2026Updated 2 weeks ago
Alternatives and similar repositories for PatchEval
Users that are interested in PatchEval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TeLL: Log Level Suggestions via Modeling Multi-Level Code Block Information, ISSTA'22☆14Jul 14, 2022Updated 4 years ago
- FLOWMATRIX: GPU-Assisted Information-Flow Analysis through Matrix-Based Representation, USENIX Security'22☆28Apr 17, 2023Updated 3 years ago
- An standalone execution trace library built on DynamoRIO.☆23Jul 4, 2022Updated 4 years ago
- ☆45Sep 8, 2023Updated 3 years ago
- GAINS: Getting stArted wIth biNary analysiS☆32Feb 23, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆26Mar 27, 2026Updated 5 months ago
- This is an evaluation set for the problem of directed/targeted test input generation. We use it to benchmark the ability of Large Languag…☆34Mar 11, 2025Updated last year
- Bilingual Resume Template in Latex. 中英双语Latex简历模板☆20Jun 4, 2024Updated 2 years ago
- CVE-Bench: A Benchmark for AI Agents’ Ability to Exploit Real-World Web Application Vulnerabilities☆288Updated this week
- A practical fuzzing tool for SMT solvers☆11Nov 26, 2025Updated 9 months ago
- Source code of AsiaCCS'22 paper - RecIPE: Revisiting the Evaluation of Memory Error Defenses☆14Sep 19, 2023Updated 3 years ago
- PalanTír: Optimizing Attack Provenance with Hardware-enhanced System Observability, ACM CCS'22☆25Nov 11, 2024Updated last year
- xAST评价体系,让安全工具不再“黑盒”. The xAST evaluation benchmark makes security tools no longer a "black box".☆494May 21, 2026Updated 3 months ago
- CVE-Factory☆182Mar 27, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A continuously updated collection of papers on agentic SE☆637Jul 17, 2026Updated 2 months ago
- [SOSP'25] Automatic checker synthesis for system-level static analysis☆189Oct 26, 2025Updated 10 months ago
- To detect logic bugs in graph database engines by mutating graph query patterns. ICSE'24.☆37Jan 24, 2024Updated 2 years ago
- SecCodeBench is a benchmark suite focusing on evaluating the security of code generated by large language models (LLMs).☆132Jun 10, 2026Updated 3 months ago
- Automated Benchmarking of LLM Agents on Real-World Software Security Tasks [NeurIPS 2025]☆97Jan 27, 2026Updated 7 months ago
- MCPCorpus is a comprehensive dataset for analyzing the Model Context Protocol (MCP) ecosystem, containing ~14K MCP servers and 300 MCP cl…☆34Sep 1, 2025Updated last year
- Learning graph-based code representations for source-level functional similarity detection. ICSE'23☆64Mar 27, 2023Updated 3 years ago
- The repo of "BugLens"☆42Jul 4, 2026Updated 2 months ago
- ☆14Oct 14, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Artifact for TOSEM Submission: GiantRepair☆12Jun 26, 2024Updated 2 years ago
- ☆18Sep 24, 2025Updated 11 months ago
- ConcoLLMic: the first language- and theory-agonistic concolic execution engine via LLM agents☆158Sep 8, 2026Updated last week
- ☆25Sep 9, 2026Updated last week
- [ASE2024] Mutual Learning-Based Framework for Enhancing Robustness of Code Models via Adversarial Training☆11Sep 13, 2024Updated 2 years ago
- A portable framework to map DFG (dataflow graph, representing an application) on spatial accelerators.☆41Oct 31, 2022Updated 3 years ago
- An autonomous LLM-agent for large-scale, repository-level code auditing☆443Mar 12, 2026Updated 6 months ago
- Generates executable Proof-of-Concept for any bug in any project. AI agents discover and reproduce vulnerabilities — verified, not halluc…☆34May 5, 2026Updated 4 months ago
- A manually vetted dataset for security vulnerability detection in Java projects☆112Aug 12, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [USENIX Security 25] PatchAgent is a LLM-based practical program repair agent that mimics human expertise.☆127Feb 25, 2026Updated 6 months ago
- Tai-e assignments for static program analysis☆1,228Aug 28, 2025Updated last year
- An easy-to-learn/use static analysis framework for Java and Android☆1,818Updated this week
- CVEfixes: Automated Collection of Vulnerabilities and Their Fixes from Open-Source Software☆363Jul 30, 2024Updated 2 years ago
- ☆19May 27, 2025Updated last year
- A Dataflow-Driven and Automated Fuzzer for the PHP Interpreter☆49Jun 19, 2025Updated last year
- UIHash: Detecting Similar Android UIs through Grid-Based Visual Appearance Representation, USENIX Security '24☆12Dec 5, 2024Updated last year