AgentLeak: Open benchmark for privacy leakage in LLM agents — 7 channels, multi-agent, multi-framework.
☆30Jul 1, 2026Updated last month
Alternatives and similar repositories for AgentLeak
Users that are interested in AgentLeak are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 文言文信息抽取(实体识别+关系抽取)☆10Feb 24, 2023Updated 3 years ago
- 中英文神经网络机器翻译☆14Jan 17, 2021Updated 5 years ago
- LobotoMl is a set of scripts and tools to assess production deployments of ML services☆10May 16, 2022Updated 4 years ago
- A collection of sample code used in some experiments with Sliver C2☆17Mar 28, 2023Updated 3 years ago
- 自然语言处理大作业——方面情感分析(Aspect Based Sentiment Analysis)☆14Sep 18, 2019Updated 6 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆48Dec 9, 2025Updated 8 months ago
- ☆13Mar 28, 2024Updated 2 years ago
- TraitBasis applied to TauBench☆18Nov 11, 2025Updated 9 months ago
- AI Security Research☆16Jun 21, 2023Updated 3 years ago
- Generative web directory fuzzer,crawling and subdomain checker based on chatgpt☆15May 15, 2024Updated 2 years ago
- A security-first linter for code that shouldn't need linting☆19Sep 12, 2023Updated 2 years ago
- 📄 [Talk] OFFZONE 2022 / ODS Data Halloween 2022: Black-box attacks on ML models + with use of open-source tools☆14May 23, 2023Updated 3 years ago
- A simple way to exfiltrate data using spotify API☆18Aug 19, 2026Updated last week
- Cache Me, Catch You: Cache Related Security Threats in LLM Serving Frameworks (NDSS 2026)☆19Dec 18, 2025Updated 8 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official PyTorch implementation of "MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks"☆18Dec 4, 2025Updated 8 months ago
- ☆16Jun 15, 2024Updated 2 years ago
- DynAuditClaw — A security audit skill that dynamically discovers your OpenClaw agent's real configuration, designs targeted attack scenar…☆15Apr 6, 2026Updated 4 months ago
- Awesome Jailbreak, red teaming arxiv papers (Automatically Update Every 12th hours)☆120Updated this week
- Official Implementation of "ToolSafe: Enhancing Tool Invocation Safety of LLM-based Agents via Proactive Step-level Guardrail and Feedbac…☆76Mar 25, 2026Updated 5 months ago
- Artifact Evaluation of "GhostCache: Timer- and Counter-Free Cache Attacks Exploiting Weak Coherence on RISC-V and ARM Chips" (CCS 2025)☆23Jul 19, 2025Updated last year
- Multimodal Safety Awareness Benchmark for Large Language Models☆15Jun 3, 2025Updated last year
- AI fun☆27Feb 27, 2025Updated last year
- ☆14Feb 24, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆20Jul 18, 2024Updated 2 years ago
- ☆18Jun 19, 2023Updated 3 years ago
- The official implementation of the paper "AgentLAB: Benchmarking LLM Agents against Long-Horizon Attacks"☆30Jun 1, 2026Updated 2 months ago
- [VLM-Attack-Survey-2024] Paper list and projects for VLM attacks☆17Feb 12, 2025Updated last year
- Research on "Many-Shot Jailbreaking" in Large Language Models (LLMs). It unveils a novel technique capable of bypassing the safety mechan…☆17Aug 6, 2024Updated 2 years ago
- The repo for using the model https://huggingface.co/thu-coai/Attacker-v0.1☆13Apr 23, 2025Updated last year
- ☆74Dec 11, 2025Updated 8 months ago
- [ICML'25] MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents☆38Jul 31, 2025Updated last year
- github for Beyond ReAct: A Planner-Centric Framework for Complex \\ Tool-Augmented LLM Reasoning☆19Feb 27, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ASIDE: Architectural Separation of Instructions and Data in Language Models [ICLR 2026]☆17Jun 10, 2026Updated 2 months ago
- Code for our NAACL2025 accepted paper: Attention Tracker: Detecting Prompt Injection Attacks in LLMs☆29Sep 19, 2025Updated 11 months ago
- [NeurIPS2024] Official code for (IMA) Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs☆23Oct 15, 2024Updated last year
- pytorch实现实时人脸识别,使用opencv+facenet+mtcnn+knn☆17Jul 15, 2022Updated 4 years ago
- [EMNLP 2026] PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization☆18Aug 21, 2026Updated last week
- ☆18May 17, 2025Updated last year
- ☆24Jun 13, 2024Updated 2 years ago