lakeraai/pint-benchmark

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/lakeraai/pint-benchmark)

lakeraai / pint-benchmark

A benchmark for prompt injection detection systems.

☆198

Alternatives and similar repositories for pint-benchmark

Users that are interested in pint-benchmark are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

lakeraai / dsec-gandalf
View on GitHub
☆24Mar 18, 2025Updated last year
leolee99 / PIGuard
View on GitHub
[ACL 2025] The official implementation of the paper "PIGuard: Prompt Injection Guardrail via Mitigating Overdefense for Free".
☆79Dec 4, 2025Updated 7 months ago
microsoft / BIPIA
View on GitHub
A benchmark for evaluating the robustness of LLMs and defenses to indirect prompt injection attacks.
☆147Apr 15, 2024Updated 2 years ago
ethz-spylab / agentdojo
View on GitHub
A Dynamic Environment to Evaluate Attacks and Defenses for LLM Agents.
☆674Jun 2, 2026Updated last month
SaFo-Lab / ReasoningBomb
View on GitHub
[CCS 2026] The official implementation of our CCS 2026 paper "ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathological…
☆15Jun 24, 2026Updated 3 weeks ago
Deploy on Railway without the complexity - Free Credits Offer • Ad
Connect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
SaFo-Lab / DynAuditClaw
View on GitHub
DynAuditClaw — A security audit skill that dynamically discovers your OpenClaw agent's real configuration, designs targeted attack scenar…
☆15Apr 6, 2026Updated 3 months ago
aibom-squad / AIBOM-RSA-2024
View on GitHub
AIBOM Workshop RSA 2024
☆15May 20, 2024Updated 2 years ago
Yu-Fangxu / COLD-Attack
View on GitHub
[ICML 2024] COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability
☆176Dec 18, 2024Updated last year
microsoft / TaskTracker
View on GitHub
TaskTracker is an approach to detecting task drift in Large Language Models (LLMs) by analysing their internal activations. It provides a…
☆92Sep 1, 2025Updated 10 months ago
corca-ai / awesome-llm-security
View on GitHub
A curation of awesome tools, documents and projects about LLM Security.
☆1,659Aug 20, 2025Updated 11 months ago
SaFo-Lab / JailBreakV_28K
View on GitHub
[COLM 2024] JailBreakV-28K: A comprehensive benchmark designed to evaluate the transferability of LLM jailbreak attacks to MLLMs, and fur…
☆96May 9, 2025Updated last year
CryptoAILab / FigStep
View on GitHub
[AAAI'25 (Oral)] Jailbreaking Large Vision-language Models via Typographic Visual Prompts
☆211Jun 26, 2025Updated last year
facebookresearch / SecAlign
View on GitHub
Repo for the research paper "SecAlign: Defending Against Prompt Injection with Preference Optimization"
☆98Jul 2, 2026Updated 2 weeks ago
xampla / carapace
View on GitHub
🦞 Carapace - The hard shell that protects your OpenClaw from prompt injection
☆22Feb 2, 2026Updated 5 months ago
Deploy to Railway using AI coding agents - Free Credits Offer • Ad
Use Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
agencyenterprise / PromptInject
View on GitHub
PromptInject is a framework that assembles prompts in a modular fashion to provide a quantitative analysis of the robustness of LLMs to a…
☆510Apr 27, 2026Updated 2 months ago
corca-ai / LLMFuzzAgent
View on GitHub
[Corca / ML] Automatically solved Gandalf AI with LLM
☆51Jul 11, 2023Updated 3 years ago
protectai / rebuff
View on GitHub
LLM Prompt Injection Detector
☆1,513Aug 7, 2024Updated last year
HumanCompatibleAI / tensor-trust-data
View on GitHub
Dataset for the Tensor Trust project
☆49Mar 17, 2024Updated 2 years ago
YihanWang617 / llm-jailbreaking-defense
View on GitHub
A lightweight library for large laguage model (LLM) jailbreaking defense.
☆61Sep 11, 2025Updated 10 months ago
postmodern / npm_scan
View on GitHub
Scans npmjs.org for npm packages that can be taken over
☆19Jun 6, 2022Updated 4 years ago
ShiftLeftSecurity / scan-docs
View on GitHub
☆27Sep 15, 2022Updated 3 years ago
nccgroup / http-mcp-bridge
View on GitHub
☆17May 1, 2026Updated 2 months ago
centerforaisafety / HarmBench
View on GitHub
HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal
☆1,011Aug 16, 2024Updated last year
AI Agents on DigitalOcean Gradient AI Platform • Ad
Build production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
lapisrocks / rpo
View on GitHub
Official repository for "Robust Prompt Optimization for Defending Language Models Against Jailbreaking Attacks"
☆62Aug 8, 2024Updated last year
yuplin2333 / representation-space-jailbreak
View on GitHub
Code repo of our paper Towards Understanding Jailbreak Attacks in LLMs: A Representation Space Analysis (https://arxiv.org/abs/2406.10794…
☆24Jul 26, 2024Updated last year
kriti-hippo / red_queen
View on GitHub
Red Queen Dataset and data generation template
☆27Dec 26, 2025Updated 6 months ago
aws-samples / aws-amplify-cloud-assistant-app
View on GitHub
☆13Jan 28, 2024Updated 2 years ago
aira-security / Vulnerable-AI-Chatbot
View on GitHub
An intentionally vulnerable AI chatbot to learn and practice AI Security.
☆21Nov 6, 2025Updated 8 months ago
aws-samples / detecting-data-drift-in-nlp-using-amazon-sagemaker-custom-model-monitor
View on GitHub
☆11Dec 20, 2023Updated 2 years ago
nccgroup / SFPolDevChk
View on GitHub
Salesforce Policy Deviation Checker
☆30Sep 30, 2020Updated 5 years ago
enferex / sataniccanary
View on GitHub
A GCC plugin implementing various stack canaries.
☆14Sep 7, 2012Updated 13 years ago
cybozu / prompt-hardener
View on GitHub
Prompt Hardener analyzes prompt-injection-originated risk in LLM-based agents and applications.
☆54May 12, 2026Updated 2 months ago
Managed hosting for WordPress and PHP on Cloudways • Ad
Managed hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
microsoft / PyRIT
View on GitHub
The Python Risk Identification Tool for generative AI (PyRIT) is an open source framework built to empower security professionals and eng…
☆4,158Updated this week
chawins / llm-sp
View on GitHub
Papers and resources related to the security and privacy of LLMs 🤖
☆579Jun 8, 2025Updated last year
rmunro / headlines
View on GitHub
Practical example from Human-in-the-Loop Machine Learning book
☆11Oct 28, 2021Updated 4 years ago
OWASP / www-project-artificial-intelligence-vulnerability-scoring-system
View on GitHub
OWASP Foundation web repository
☆77Apr 10, 2026Updated 3 months ago
sherdencooper / GPTFuzz
View on GitHub
Official repo for GPTFUZZER : Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts
☆601Feb 27, 2026Updated 4 months ago
OSU-NLP-Group / EIA_against_webagent
View on GitHub
☆40Oct 2, 2024Updated last year
suzgunmirac / marnns
View on GitHub
MARNNs Can Learn Generalized Dyck Languages
☆12Nov 11, 2019Updated 6 years ago