The opensoure repository of FuzzLLM
☆37May 4, 2024Updated 2 years ago
Alternatives and similar repositories for FuzzLLM
Users that are interested in FuzzLLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The repo for paper: Exploiting the Index Gradients for Optimization-Based Jailbreaking on Large Language Models.☆15Dec 16, 2024Updated last year
- MASTERKEY is a framework designed to explore and exploit vulnerabilities in large language model chatbots by automating jailbreak attacks…☆42Sep 12, 2024Updated 2 years ago
- Implementation of SLIM, a framework of dynamics skill lifecycle management for agentic reinforcement learning☆22Sep 25, 2026Updated 2 weeks ago
- ☆25Jan 17, 2025Updated last year
- Multi-dimensional analysis of orthogonal safety directions in LLM alignment☆23Jun 12, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for Chain-of-Thought Hijacking☆29Nov 10, 2025Updated 11 months ago
- ☆16Jun 15, 2024Updated 2 years ago
- Code to the paper: The Geometry of Refusal in Large Language Models: Concept Cones and Representational Independence☆36Jul 31, 2025Updated last year
- ☆122Apr 27, 2025Updated last year
- Official Code Repository for [AutoScale📈: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*…☆14Aug 8, 2025Updated last year
- ☆20Dec 13, 2021Updated 4 years ago
- [ACL 25] SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities☆30Apr 2, 2025Updated last year
- [EMNLP 2025] The code repo of paper "X-Boundary: Establishing Exact Safety Boundary to Shield LLMs from Multi-Turn Jailbreaks without Com…☆40Nov 24, 2025Updated 10 months ago
- [ICLR 2024] The official implementation of our ICLR2024 paper "AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language M…☆466Jan 22, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for the 2025 ACL publication "Fine-Tuning on Diverse Reasoning Chains Drives Within-Inference CoT Refinement in LLMs"☆33Jun 25, 2025Updated last year
- csl: PyTorch-based Constrained Learning☆11Jun 1, 2022Updated 4 years ago
- [Tensorflow] A Game Theoretic approach using GAN for Phishing URL synthesis and detection☆11Nov 14, 2022Updated 3 years ago
- ☆21Jun 27, 2024Updated 2 years ago
- Jailbreak artifacts for JailbreakBench☆113Nov 6, 2024Updated last year
- All in How You Ask for It: Simple Black-Box Method for Jailbreak Attacks☆17Apr 24, 2024Updated 2 years ago
- The dataset CoLan-150K and the concept decomposition in the paper Concept Lancet (CVPR 2025)☆21Jan 18, 2026Updated 8 months ago
- [CCS 2026] The official implementation of our CCS 2026 paper "ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathological…☆18Aug 5, 2026Updated 2 months ago
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆49Feb 26, 2025Updated last year
- ☆13Oct 6, 2022Updated 4 years ago
- JailbreakBench: An Open Robustness Benchmark for Jailbreaking Language Models [NeurIPS 2024 Datasets and Benchmarks Track]☆687Apr 4, 2025Updated last year
- Jailbreaking Multimodal Large Language Models via Shuffle Inconsistency☆16Aug 6, 2025Updated last year
- DSN jailbreak Attack & Evaluation Ensemble☆17Feb 7, 2026Updated 8 months ago
- ☆103Mar 20, 2025Updated last year
- [USENIX Security '24] An LLM-Assisted Easy-to-Trigger Backdoor Attack on Code Completion Models: Injecting Disguised Vulnerabilities agai…☆61Mar 22, 2025Updated last year
- 春秋云境靶场解题攻略☆12Jul 31, 2023Updated 3 years ago
- ☆24Jun 16, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [AAAI 2026] The official code for ``LoopLLM: Transferable Energy-Latency Attacks in LLMs via Repetitive Generation''☆17Mar 20, 2026Updated 6 months ago
- RAG Hallucination Detecting By LRP.☆13Mar 31, 2025Updated last year
- [AST'26] LLAMAFUZZ: Large Language Model Enhanced Greybox Fuzzing☆23Dec 3, 2024Updated last year
- Q-Probe: A Lightweight Approach to Reward Maximization for Language Models☆41Jun 10, 2024Updated 2 years ago
- Your finetuned model's back to its original safety standards faster than you can say "SafetyLock"!☆11Oct 16, 2024Updated last year
- Extensible Platform for Malware Analysis☆17Jan 14, 2021Updated 5 years ago
- BDC is the eBPF powered DNS caching mechanism in kernel inspired by BMC☆10May 13, 2022Updated 4 years ago