Self-Evolving Defense for AI Agents — Protect against prompt injection, data exfiltration, and multi-stage attacks with adaptive security.
☆19Apr 8, 2026Updated 4 months ago
Alternatives and similar repositories for ClawArmor
Users that are interested in ClawArmor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [Machine Learning 2023] Imbalanced Gradients: A Subtle Cause of Overestimated Adversarial Robustness☆17Jul 5, 2024Updated 2 years ago
- ☆25Apr 26, 2023Updated 3 years ago
- ☆10Jul 28, 2022Updated 4 years ago
- A Unified Benchmark and Toolbox for Multimodal Jailbreak Attack–Defense Evaluation☆75May 8, 2026Updated 3 months ago
- Research on "Many-Shot Jailbreaking" in Large Language Models (LLMs). It unveils a novel technique capable of bypassing the safety mechan…☆17Aug 6, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ICCV 2023 - AdaptGuard: Defending Against Universal Attacks for Model Adaptation☆11Dec 23, 2023Updated 2 years ago
- ☆20Mar 14, 2022Updated 4 years ago
- Adversarial Images for Variational Autoencoders☆13Nov 30, 2016Updated 9 years ago
- Official code repository for the paper "Internal Activation as the Polar Star for Steering Unsafe LLM Behavior"☆15May 31, 2026Updated 2 months ago
- 100天写100个gpts,从基础到深入☆16Updated this week
- This repository is the replication package of the NeurIPS19 paper "MarginGAN: Adversarial Training in Semi-Supervised Learning"☆12Oct 27, 2019Updated 6 years ago
- [ACL 2025] LongSafety: Evaluating Long-Context Safety of Large Language Models☆16Jun 18, 2025Updated last year
- A simple implementation of ReasonGenRM.☆19Apr 21, 2025Updated last year
- This is a Pytorch implementation of contrastive Learning(CL) baselines.☆14Aug 29, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Github repo for One-shot Neural Backdoor Erasing via Adversarial Weight Masking (NeurIPS 2022)☆15Jan 3, 2023Updated 3 years ago
- Coupling rejection strategy against adversarial attacks (CVPR 2022)☆29Mar 2, 2022Updated 4 years ago
- General research for Dreadnode☆28Jun 17, 2024Updated 2 years ago
- LIMA: Language for Integrated Modeling and Analysis☆12Sep 8, 2018Updated 7 years ago
- Vstream - Video Analytics pipeline with Hardware based accelerations (dev - stage)☆10Feb 2, 2024Updated 2 years ago
- ☆18Mar 15, 2026Updated 4 months ago
- [CVPRW 2023] Diversity is Definitely Needed: Improving Model-Agnostic Zero-shot Classification via Stable Diffusion☆24Jan 24, 2024Updated 2 years ago
- [CVPR'24 Oral] Metacloak: Preventing Unauthorized Subject-driven Text-to-image Diffusion-based Synthesis via Meta-learning☆32Nov 19, 2024Updated last year
- Official Repository of Paper "Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs"☆15Sep 25, 2025Updated 10 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code for Transferable Unlearnable Examples☆22Mar 11, 2023Updated 3 years ago
- CVPR2023: Unlearnable Clusters: Towards Label-agnostic Unlearnable Examples☆22Apr 25, 2023Updated 3 years ago
- [CVPR 2023] The official implementation of our CVPR 2023 paper "Detecting Backdoors During the Inference Stage Based on Corruption Robust…☆25May 25, 2023Updated 3 years ago
- ☆29Jun 17, 2024Updated 2 years ago
- Precision Knowledge Editing (PKE): A novel method to reduce toxicity in LLMs while preserving performance, with robust evaluations and ha…☆12Nov 26, 2024Updated last year
- ☆25Jan 20, 2019Updated 7 years ago
- ☆42Sep 9, 2023Updated 2 years ago
- ☆19Jan 1, 2023Updated 3 years ago
- Max Mahalanobis Training (ICML 2018 + ICLR 2020)☆90Dec 21, 2020Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An implementation of MSSRM method☆10Mar 23, 2023Updated 3 years ago
- ☆27Aug 7, 2024Updated 2 years ago
- DETR tensor去除推理过程无用辅助头+fp16部署再次加速+解决转tensorrt 输出全为0问题的新方法。☆12Jan 9, 2024Updated 2 years ago
- Code of paper [CVPR'24: Can Protective Perturbation Safeguard Personal Data from Being Exploited by Stable Diffusion?]☆26Apr 2, 2024Updated 2 years ago
- A Diagnostic Guardrail Framework for AI Agent Safety and Security☆683Jun 8, 2026Updated 2 months ago
- ☆15Jul 23, 2025Updated last year
- Official code for EasyDrag (CVPR 2024)☆17Jun 18, 2024Updated 2 years ago