☆103Jan 30, 2026Updated 7 months ago
Alternatives and similar repositories for baxbench
Users that are interested in baxbench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Simultaneous evaluation on both functionality and security of LLM-generated code.☆44Jul 30, 2026Updated last month
- [NeurIPS 2024] Evaluation harness for SWT-Bench, a benchmark for evaluating LLM repository-level test-generation☆91Jul 23, 2026Updated 2 months ago
- ☆134Jul 14, 2024Updated 2 years ago
- "SALLM: Security Assessment of Generated Code" accepted at ASYDE workshop co-located with ASE'24.☆17Jul 21, 2026Updated 2 months ago
- Enhacing Code Pre-trained Models by Contrastive Learning☆41Mar 8, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The official codes for our paper at COLING 2022: Semantic-Preserving Adversarial Code Comprehension☆12Oct 23, 2022Updated 3 years ago
- CC: Causality-Aware Coverage Criterion for Deep Neural Networks☆12Feb 15, 2023Updated 3 years ago
- ☆13Oct 11, 2024Updated last year
- This repository contains the replication package of our paper "Assessing the Security of GitHub Copilot’s Generated Code - A Targeted Rep…☆10Nov 16, 2023Updated 2 years ago
- Guardrails for secure and robust agent development☆463Jan 12, 2026Updated 8 months ago
- ToolFuzz is a fuzzing framework designed to test your LLM Agent tools.☆43Jul 20, 2025Updated last year
- ☆21Feb 3, 2025Updated last year
- ☆39Jan 13, 2023Updated 3 years ago
- [SANER 2023] MixCode: Enhancing Code Classification by Mixup-Based Data Augmentation☆15Jul 13, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Synthesized models for PHOG to make the results reproducible by the research community☆11Jan 23, 2020Updated 6 years ago
- A repository of Language Model Vulnerabilities and Exposures (LVEs).☆113Mar 12, 2024Updated 2 years ago
- [ICSE'24 Industry Challenge Track] "ReposVul: A Repository-Level High-Quality Vulnerability Dataset".☆111Nov 24, 2024Updated last year
- [NAACL 2025] Benchmark for Repository-Level Code Generation, focus on Executability, Correctness from Test Cases and Usage of Contexts fr…☆47Jan 8, 2026Updated 8 months ago
- Certifying Geometric Robustness of Neural Networks☆16Mar 24, 2023Updated 3 years ago
- An implementation of the ACL 2024 Findings paper "Generalization-Enhanced Code Vulnerability Detection via Multi-Task Instruction Fine-Tu…☆79Oct 29, 2025Updated 10 months ago
- CVE-Bench: A Benchmark for AI Agents’ Ability to Exploit Real-World Web Application Vulnerabilities☆291Updated this week
- This repository includes various baseline techniques for label-free model evaluation task for the VDU2023 competition.☆19Mar 8, 2023Updated 3 years ago
- Constrained Decoding of Diffusion LLMs with Context-Free Grammars.☆57Dec 17, 2025Updated 9 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆28Sep 15, 2024Updated 2 years ago
- [EMNLP'22] Code for 'Exploring Representation-level Augmentation for Code Search'☆28Oct 9, 2023Updated 2 years ago
- ☆116Aug 6, 2026Updated last month
- Code for the paper "Firewalls to Secure Dynamic LLM Agentic Networks"☆30Jun 6, 2025Updated last year
- ☆70Dec 15, 2024Updated last year
- ☆16Aug 26, 2023Updated 3 years ago
- Baseline rules files to improve the security of AI-generated code (Claude, Cursor, Copilot + more)☆241Dec 23, 2025Updated 9 months ago
- Implementation for "RigorLLM: Resilient Guardrails for Large Language Models against Undesired Content"☆24Jul 28, 2024Updated 2 years ago
- Code release for the ICML 2019 paper "Are generative classifiers more robust to adversarial attacks?"☆24May 10, 2019Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ESLint plugin to detect and stop Trojan Source attacks☆80Jun 22, 2026Updated 3 months ago
- GemAI - A Free RAG CLI ChatBot 🤖☆16Aug 26, 2024Updated 2 years ago
- ☆32Oct 28, 2023Updated 2 years ago
- The library for symbolic interval☆23Jun 23, 2020Updated 6 years ago
- Generating Adversarial Examples for Holding Robustness of Source Code Processing Models☆17Dec 2, 2021Updated 4 years ago
- Repository for PrimeVul Vulnerability Detection Dataset☆274Sep 7, 2024Updated 2 years ago
- Towards Robustness of Deep Program Processing Models – Detection, Estimation and Enhancement☆22Oct 29, 2022Updated 3 years ago