This is an evaluation set for the problem of directed/targeted test input generation. We use it to benchmark the ability of Large Language Models for generating inputs to reach a certain code location or produce a particular result.
☆34Mar 11, 2025Updated last year
Alternatives and similar repositories for PathEval
Users that are interested in PathEval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ISSTA 2025] A Large-scale Empirical Study on Fine-tuning Large Language Models for Unit Testing☆13Feb 9, 2025Updated last year
- A practical fuzzing tool for SMT solvers☆11Nov 26, 2025Updated 9 months ago
- ☆164May 27, 2025Updated last year
- Evaluation code of ASE24 accepted paper "On the Evaluation of LLM in Unit Test Generation"☆13Dec 9, 2024Updated last year
- ☆13Sep 12, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆36Jan 27, 2025Updated last year
- A learning-guided approach for executing arbitrary Python code snippets☆16Mar 4, 2024Updated 2 years ago
- Dataset of Codex generated tests for the CodaMosa project☆19Jun 2, 2023Updated 3 years ago
- WhiteFox: White-Box Compiler Fuzzing Empowered by Large Language Models (OOPSLA 2024)☆85Aug 5, 2025Updated last year
- ☆13Nov 20, 2024Updated last year
- An implementation of the ACL 2024 Findings paper "Generalization-Enhanced Code Vulnerability Detection via Multi-Task Instruction Fine-Tu…☆77Oct 29, 2025Updated 10 months ago
- A continuously updated collection of papers on agentic SE☆633Jul 17, 2026Updated last month
- The Z3-Noodler String Solver☆27Updated this week
- Fuzzing Deep-Learning Libraries via Automated Relational API Inference (ESEC/FSE 2022)☆39May 17, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆24Jul 2, 2024Updated 2 years ago
- Automatic AI-powered test suite generator☆112Apr 5, 2026Updated 4 months ago
- PatchEval: A New Benchmark for Evaluating LLMs on Patching Real-World Vulnerabilities☆227Updated this week
- ☆175Jul 25, 2025Updated last year
- ☆19Jan 17, 2024Updated 2 years ago
- Create CFGs and compute complexity metrics for Python, C++, and Java code.☆46May 10, 2024Updated 2 years ago
- Generates executable Proof-of-Concept for any bug in any project. AI agents discover and reproduce vulnerabilities — verified, not halluc…☆30May 5, 2026Updated 3 months ago
- Cottontail: A LLM-Driven Concolic Execution Engine (Accepted by IEEE S&P'26)☆46Dec 4, 2025Updated 8 months ago
- MetaMut is a mutation operator generator to facilitate compiler fuzzing.☆32Dec 29, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ConcoLLMic: the first language- and theory-agonistic concolic execution engine via LLM agents☆155May 25, 2026Updated 3 months ago
- The First International Workshop on Large Language Models for Code 2024 (Co-Located with ICSE 2024)☆18Oct 4, 2024Updated last year
- Empc: Effective Path Prioritization for Symbolic Execution with Path Cover☆35May 11, 2025Updated last year
- tool of llm-based indirect-call analyzer☆31Feb 18, 2025Updated last year
- A prototype to write blog posts with executable ocaml code blocks☆10Apr 25, 2025Updated last year
- OCaml bindings to Minisat☆12May 6, 2024Updated 2 years ago
- A GPT-Based Fuzz Driver Generator☆48Nov 19, 2023Updated 2 years ago
- Official repository for PraPR source code☆14May 11, 2021Updated 5 years ago
- This repository contains several examples of logic bomb.☆115Dec 23, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆25Updated this week
- ☆94Sep 10, 2023Updated 2 years ago
- The source code of project "LLift" (Enhancing static analysis with LLM)☆88Mar 5, 2024Updated 2 years ago
- ☆11Sep 28, 2022Updated 3 years ago
- C++ realisation of gnfs algorithm☆13Aug 5, 2020Updated 6 years ago
- ☆13Dec 31, 2024Updated last year
- Recent symbolic execution papers and tools.☆186May 16, 2025Updated last year