☆23Jul 26, 2025Updated last year
Alternatives and similar repositories for SQL-Injection-Jailbreak
Users that are interested in SQL-Injection-Jailbreak are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [MM'23] ProTegO: Protect Text Content against OCR Extraction Attack☆14Mar 12, 2024Updated 2 years ago
- A Robust Provably Secure Linguistic Steganography Method with Diffusion Language Model☆17Dec 8, 2025Updated 9 months ago
- [NeurIPS 2025] Official Implementation of paper "LD-RoViS: Training-free robust video steganography for deterministic latent diffusion mo…☆34Mar 11, 2026Updated 6 months ago
- [Neurips 2025]StegoZip: Enhancing Linguistic Steganography Payload in Practice with Large Language Models☆33Dec 4, 2025Updated 9 months ago
- A provably secure disambiguating steganography method based on grouping ambiguous pools and synchronous sampling.☆17Apr 24, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- LiveSecBench:动态中文大模型安全榜单☆29Mar 9, 2026Updated 6 months ago
- ☆24May 14, 2025Updated last year
- [NeurIPS 2025] The official implementation of "T2SMark: Balancing Robustness and Diversity in Noise-as-Watermark for Diffusion Models"☆51May 9, 2026Updated 4 months ago
- [TDSC 2025] InferDPT: Privacy-Preserving Inference for Closed-box Large Language Model☆46Nov 16, 2025Updated 10 months ago
- The source code of QueryAttack.☆27Feb 23, 2025Updated last year
- ☆31Feb 19, 2025Updated last year
- [AAAI 2024] Data-Free Hard-Label Robustness Stealing Attack☆16Mar 29, 2024Updated 2 years ago
- [AAAI 2026 Oral] AEDR: Training-Free AI-Generated Image Attribution via Autoencoder Double-Reconstruction☆24Apr 21, 2026Updated 4 months ago
- [NDSS'25] The official implementation of safety misalignment.☆19Jan 8, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆133Jul 29, 2026Updated last month
- Official implementation of paper: DrAttack: Prompt Decomposition and Reconstruction Makes Powerful LLM Jailbreakers☆68Sep 12, 2026Updated last week
- Welcome to the official repository for Siren, a project aimed at understanding and mitigating harmful behaviors in large language models …☆15Jun 14, 2026Updated 3 months ago
- Red Queen Dataset and data generation template☆29Dec 26, 2025Updated 8 months ago
- [IEEE T-IFS] AutoPT: How Far Are We from the Fully Automated Web Penetration Testing?☆48Jun 1, 2026Updated 3 months ago
- ☆37Dec 2, 2023Updated 2 years ago
- Code repo of our paper Towards Understanding Jailbreak Attacks in LLMs: A Representation Space Analysis (https://arxiv.org/abs/2406.10794…☆24Jul 26, 2024Updated 2 years ago
- [ICCV 2025] The official code of the paper "Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration R…☆113Jul 9, 2025Updated last year
- [ICML 2025] An official source code for paper "FlipAttack: Jailbreak LLMs via Flipping".☆181Aug 11, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ACL 2024] CodeAttack: Revealing Safety Generalization Challenges of Large Language Models via Code Completion☆60Oct 1, 2025Updated 11 months ago
- ☆11May 18, 2025Updated last year
- ☆40May 17, 2025Updated last year
- The official code for "Steering Dialogue Dynamics for Robustness against Multi-turn Jailbreaking Attacks".☆19Jun 24, 2026Updated 2 months ago
- [ACL 2025] The official implementation of the paper "PIGuard: Prompt Injection Guardrail via Mitigating Overdefense for Free".☆87Dec 4, 2025Updated 9 months ago
- offical implementation of MTSA: Multi-turn Safety Alignment for LLMs through Multi-round Red-teaming☆18Jun 2, 2025Updated last year
- Ferret: Faster and Effective Automated Red Teaming with Reward-Based Scoring Technique☆19Aug 22, 2024Updated 2 years ago
- The official repository for guided jailbreak benchmark☆32Aug 31, 2026Updated 2 weeks ago
- ☆21Apr 7, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [AAAI'25 (Oral)] Jailbreaking Large Vision-language Models via Typographic Visual Prompts☆216Jun 26, 2025Updated last year
- ☆40Oct 14, 2021Updated 4 years ago
- This repository includes main notebook of the code for our proposed RCGAN☆12Apr 10, 2020Updated 6 years ago
- [ICLR 2025] A Closer Look at Machine Unlearning for Large Language Models☆49Dec 4, 2024Updated last year
- Source code of NAACL 2025 Findings "Scaling Up Membership Inference: When and How Attacks Succeed on Large Language Models"☆16Dec 16, 2025Updated 9 months ago
- [ECCV'24 Oral] The official GitHub page for ''Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking …☆38Oct 23, 2024Updated last year
- [ICML 2025] Speak Easy: Eliciting Harmful Jailbreaks from LLMs with Simple Interactions☆16Mar 7, 2026Updated 6 months ago