☆17Jul 2, 2026Updated 3 months ago
Alternatives and similar repositories for AMemGuard
Users that are interested in AMemGuard are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆48May 29, 2026Updated 4 months ago
- [EMNLP 2026] PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization☆20Aug 21, 2026Updated last month
- LLM Safeguarding with Internal Representations☆21Apr 27, 2026Updated 5 months ago
- Official Repo of Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents☆99Jun 2, 2026Updated 4 months ago
- This repo is the official implementation of “Are Your Agents Upward Deceivers?”. The paper is accepted by ICML 2026.☆24Dec 15, 2025Updated 9 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Memory Injection Attacks on LLM Agents via Query-Only Interaction☆37Aug 11, 2026Updated last month
- Codes for our paper "AgentMonitor: A Plug-and-Play Framework for Predictive and Secure Multi-Agent Systems"☆14Dec 13, 2024Updated last year
- Official Codebase of the ACL 2026 Oral paper "Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contra…☆28Jun 25, 2026Updated 3 months ago
- Progent: Securing AI Agents with Privilege Control☆52Sep 24, 2026Updated last week
- On the Robustness of GUI Grounding Models Against Image Attacks☆12Apr 8, 2025Updated last year
- Official Implementation of "ToolSafe: Enhancing Tool Invocation Safety of LLM-based Agents via Proactive Step-level Guardrail and Feedbac…☆80Mar 25, 2026Updated 6 months ago
- This is the official code repository for the paper: Towards General Continuous Memory for Vision-Language Models.☆32Jul 3, 2025Updated last year
- DICE: Detecting In-distribution Data Contamination with LLM's Internal State☆11Sep 21, 2024Updated 2 years ago
- ☆12Sep 8, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ACL 2025] The official code for "AGrail: A Lifelong Agent Guardrail with Effective and Adaptive Safety Detection".☆45Aug 12, 2026Updated last month
- ☆47Jun 28, 2025Updated last year
- [NDSS'24] Inaudible Adversarial Perturbation: Manipulating the Recognition of User Speech in Real Time☆57Sep 28, 2024Updated 2 years ago
- [NeurIPS 2024] Official implementation for "AgentPoison: Red-teaming LLM Agents via Memory or Knowledge Base Backdoor Poisoning"☆244Updated this week
- This is the official repository for Cross-modality Information Check for Detecting Jailbreaking in Multimodal Large Language Models.☆17Jan 16, 2025Updated last year
- ☆37Apr 13, 2026Updated 5 months ago
- [NeurIPS 2025] The official implementation of the paper "DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agen…☆61Jul 16, 2026Updated 2 months ago
- The official implementation for "Mitigating Overthinking in Large Reasoning Models via Manifold Steering"☆15May 29, 2025Updated last year
- This repository includes code for our paper: ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning…☆17May 2, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A toolkit to assess data privacy in LLMs (under development)☆75Jan 2, 2025Updated last year
- Adversarial attack against DeepSpeech2 ASR pytorch model☆24Jan 15, 2021Updated 5 years ago
- [ICLR 2026 Oral] RAIN-Merging☆16Mar 9, 2026Updated 6 months ago
- ☆12Mar 7, 2021Updated 5 years ago
- ☆16Oct 15, 2025Updated 11 months ago
- MLLM, DeepResearch, Agentic AI☆23Jun 1, 2026Updated 4 months ago
- [ICLR 2025] FLAT: LLM Unlearning via Loss Adjustment with Only Forget Data☆14Feb 26, 2025Updated last year
- Agent Security Bench (ASB)☆308Updated this week
- Can audio-visual integration strengthen robustness under multimodal attacks?☆31Mar 31, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [NDSS 2026] Official repo for Odysseus: Jailbreaking Commercial Multimodal LLM-integrated Systems via Dual Steganography☆63Mar 14, 2026Updated 6 months ago
- ☆27Aug 6, 2026Updated last month
- A Dynamic Environment to Evaluate Attacks and Defenses for LLM Agents.☆888Jun 2, 2026Updated 4 months ago
- ☆27May 14, 2026Updated 4 months ago
- ☆33Jun 5, 2025Updated last year
- [ACL 2025] Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints☆19May 23, 2025Updated last year
- [ACL 2026] Dissecting Failure Dynamics in Large Language Model Reasoning☆19Apr 17, 2026Updated 5 months ago