☆23Jul 26, 2025Updated last year
Alternatives and similar repositories for SQL-Injection-Jailbreak
Users that are interested in SQL-Injection-Jailbreak are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Robust Provably Secure Linguistic Steganography Method with Diffusion Language Model☆17Dec 8, 2025Updated 8 months ago
- [Neurips 2025]StegoZip: Enhancing Linguistic Steganography Payload in Practice with Large Language Models☆33Dec 4, 2025Updated 8 months ago
- LiveSecBench:动态中文大模型安全榜单☆29Mar 9, 2026Updated 5 months ago
- [USENIX Security 2026] Membership Inference Attacks on Tokenizers of Large Language Models☆21May 22, 2026Updated 3 months ago
- [NeurIPS 2025] The official implementation of "T2SMark: Balancing Robustness and Diversity in Noise-as-Watermark for Diffusion Models"☆51May 9, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [TDSC 2025] InferDPT: Privacy-Preserving Inference for Closed-box Large Language Model☆45Nov 16, 2025Updated 9 months ago
- The source code of QueryAttack.☆27Feb 23, 2025Updated last year
- ☆31Feb 19, 2025Updated last year
- [AAAI 2026 Oral] AEDR: Training-Free AI-Generated Image Attribution via Autoencoder Double-Reconstruction☆24Apr 21, 2026Updated 4 months ago
- [NDSS'25] The official implementation of safety misalignment.☆19Jan 8, 2025Updated last year
- ☆132Jul 29, 2026Updated last month
- Official implementation of paper: DrAttack: Prompt Decomposition and Reconstruction Makes Powerful LLM Jailbreakers☆67Aug 25, 2024Updated 2 years ago
- ☆55Feb 24, 2024Updated 2 years ago
- Welcome to the official repository for Siren, a project aimed at understanding and mitigating harmful behaviors in large language models …☆15Jun 14, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Red Queen Dataset and data generation template☆29Dec 26, 2025Updated 8 months ago
- [IEEE T-IFS] AutoPT: How Far Are We from the Fully Automated Web Penetration Testing?☆47Jun 1, 2026Updated 2 months ago
- ☆37Dec 2, 2023Updated 2 years ago
- Code repo of our paper Towards Understanding Jailbreak Attacks in LLMs: A Representation Space Analysis (https://arxiv.org/abs/2406.10794…☆24Jul 26, 2024Updated 2 years ago
- [ICCV 2025] The official code of the paper "Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration R…☆113Jul 9, 2025Updated last year
- [ICML 2025] An official source code for paper "FlipAttack: Jailbreak LLMs via Flipping".☆181Aug 11, 2026Updated 2 weeks ago
- [ACL 2024] CodeAttack: Revealing Safety Generalization Challenges of Large Language Models via Code Completion☆60Oct 1, 2025Updated 10 months ago
- ☆11May 18, 2025Updated last year
- ☆40May 17, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- The official code for "Steering Dialogue Dynamics for Robustness against Multi-turn Jailbreaking Attacks".☆19Jun 24, 2026Updated 2 months ago
- [ACL 2025] The official implementation of the paper "PIGuard: Prompt Injection Guardrail via Mitigating Overdefense for Free".☆81Dec 4, 2025Updated 8 months ago
- offical implementation of MTSA: Multi-turn Safety Alignment for LLMs through Multi-round Red-teaming☆18Jun 2, 2025Updated last year
- Ferret: Faster and Effective Automated Red Teaming with Reward-Based Scoring Technique☆19Aug 22, 2024Updated 2 years ago
- The official repository for guided jailbreak benchmark☆32Jul 28, 2025Updated last year
- ☆21Apr 7, 2025Updated last year
- Code Implementation of Adversarial Prompt Evaluation paper☆14Sep 18, 2025Updated 11 months ago
- ☆141Dec 3, 2025Updated 8 months ago
- Provably Secure Steganography in Practice Based on “Distribution Copies”☆51Jun 1, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [AAAI'25 (Oral)] Jailbreaking Large Vision-language Models via Typographic Visual Prompts☆212Jun 26, 2025Updated last year
- ☆39Oct 14, 2021Updated 4 years ago
- This repository includes main notebook of the code for our proposed RCGAN☆12Apr 10, 2020Updated 6 years ago
- [ICLR 2025] A Closer Look at Machine Unlearning for Large Language Models☆49Dec 4, 2024Updated last year
- [ECCV'24 Oral] The official GitHub page for ''Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking …☆38Oct 23, 2024Updated last year
- [ICML 2025] Speak Easy: Eliciting Harmful Jailbreaks from LLMs with Simple Interactions☆16Mar 7, 2026Updated 5 months ago
- ☆15Jun 28, 2025Updated last year