VILA-Lab/M-Attack

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/VILA-Lab/M-Attack)

VILA-Lab / M-Attack

[NeurIPS25 & ICML25 Workshop on Reliable and Responsible Foundation Models] A Simple Baseline Achieving Over 90% Success Rate Against the Strong Black-box Models of GPT-4.5/4o/o1. Paper at: https://arxiv.org/abs/2503.10635

☆100

Alternatives and similar repositories for M-Attack

Users that are interested in M-Attack are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

jiaxiaojunQAQ / FOA-Attack
View on GitHub
Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment (NeurIPS 2025)
☆67Nov 5, 2025Updated 8 months ago
jiamingzhang94 / AnyAttack
View on GitHub
CVPR 2025 - Anyattack: Towards Large-scale Self-supervised Adversarial Attacks on Vision-language Models
☆74Aug 7, 2025Updated 11 months ago
MetaAgentX / NextGen-CAPTCHAs
View on GitHub
[ICML 2026]A defense framework against MLLM-based web GUI agents. This repository provides both the generative CAPTCHA system and tools f…
☆21May 1, 2026Updated 2 months ago
Shelton1013 / Chain_of_Attack
View on GitHub
[CVPR'25]Chain of Attack: On the Robustness of Vision-Language Models Against Transfer-Based Adversarial Attacks
☆32Jun 12, 2025Updated last year
HanxunH / XTransferBench
View on GitHub
[ICML 2025] X-Transfer Attacks: Towards Super Transferable Adversarial Attacks on CLIP
☆48Feb 3, 2026Updated 5 months ago
Managed Kubernetes at scale on DigitalOcean • Ad
DigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
Jiacheng8 / CV-DD
View on GitHub
Dataset Distillation via Committee Voting
☆15Jul 28, 2025Updated 11 months ago
SensenGao / VLPTransferAttack
View on GitHub
[ECCV2024] Boosting Transferability in Vision-Language Attacks via Diversification along the Intersection Region of Adversarial Trajector…
☆32Nov 15, 2025Updated 8 months ago
Haochen-Luo / CroPA
View on GitHub
☆56Dec 7, 2024Updated last year
Zoky-2020 / SGA
View on GitHub
Set-level Guidance Attack: Boosting Adversarial Transferability of Vision-Language Pre-training Models. [ICCV 2023 Oral]
☆70Sep 6, 2023Updated 2 years ago
liudaizong / Awesome-LVLM-Attack
View on GitHub
😎 up-to-date & curated list of awesome Attacks on Large-Vision-Language-Models papers, methods & resources.
☆567Updated this week
RUCAIBox / HADES
View on GitHub
[ECCV'24 Oral] The official GitHub page for ''Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking …
☆39Oct 23, 2024Updated last year
yunqing-me / AttackVLM
View on GitHub
[NeurIPS-2023] Annual Conference on Neural Information Processing Systems
☆231Dec 22, 2024Updated last year
Summu77 / V-Attack
View on GitHub
[CVPR2026] V-Attack: Targeting Disentangled Value Features for Controllable Adversarial Attacks on LVLMs
☆20Dec 8, 2025Updated 7 months ago
thu-ml / Attack-Bard
View on GitHub
☆108Feb 16, 2024Updated 2 years ago
1-Click AI Models by DigitalOcean Gradient • Ad
Deploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
NY1024 / BAP-Jailbreak-Vision-Language-Models-via-Bi-Modal-Adversarial-Prompt
View on GitHub
☆61Jun 5, 2024Updated 2 years ago
erfanshayegani / Jailbreak-In-Pieces
View on GitHub
[ICLR 2024 Spotlight 🔥 ] - [ Best Paper Award SoCal NLP 2023 🏆] - Jailbreak in pieces: Compositional Adversarial Attacks on Multi-Modal…
☆93Jun 6, 2024Updated 2 years ago
MetaAgentX / OpenCaptchaWorld
View on GitHub
[NeurIPS 2025] The first web-based benchmark and platform to evaluate visual reasoning and interaction capabilities of MLLM powered agent…
☆82Feb 19, 2026Updated 5 months ago
jiaxiaojunQAQ / OmniSafeBench-MM
View on GitHub
A Unified Benchmark and Toolbox for Multimodal Jailbreak Attack–Defense Evaluation
☆75May 8, 2026Updated 2 months ago
jiah-li / magic
View on GitHub
The repo for paper: Exploiting the Index Gradients for Optimization-Based Jailbreaking on Large Language Models.
☆15Dec 16, 2024Updated last year
Trustworthy-AI-Group / TransferAttack
View on GitHub
TransferAttack is a pytorch framework to boost the adversarial transferability for image classification.
☆479May 12, 2026Updated 2 months ago
facebookresearch / jailbreak-objectives
View on GitHub
Code and data to go with the Zhu et al. paper "An Objective for Nuanced LLM Jailbreaks"
☆37Jul 2, 2026Updated 2 weeks ago
VILA-Lab / OD3
View on GitHub
[ICLR 2026] Optimization-free Dataset Distillation for Object Detection. Paper at: https://arxiv.org/abs/2506.01942
☆31Jan 26, 2026Updated 5 months ago
ChenWu98 / agent-attack
View on GitHub
[ICLR 2025] Dissecting adversarial robustness of multimodal language model agents
☆139Feb 19, 2025Updated last year
Open source password manager - Proton Pass • Ad
Securely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
PKU-ML / PAT
View on GitHub
Code for NeurIPS 2024 Paper "Fight Back Against Jailbreaking via Prompt Adversarial Tuning"
☆22May 6, 2025Updated last year
NY1024 / Jailbreak_GPT4o
View on GitHub
☆28Jun 5, 2024Updated 2 years ago
RylanSchaeffer / AstraFellowship-When-Do-VLM-Image-Jailbreaks-Transfer
View on GitHub
Code for ICLR 2025 Failures to Find Transferable Image Jailbreaks Between Vision-Language Models
☆37Jun 1, 2025Updated last year
chs20 / RobustVLM
View on GitHub
[ICML 2024] Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models
☆161Feb 19, 2026Updated 5 months ago
gq-max / AdvDiffVLM
View on GitHub
☆48Apr 7, 2025Updated last year
Alibaba-AAIG / Oyster
View on GitHub
The Oyster series is a set of safety models developed in-house by Alibaba-AAIG, devoted to building a responsible AI ecosystem. | Oyster …
☆62Apr 29, 2026Updated 2 months ago
hefeimei06 / VEAttack-LVLM
View on GitHub
[ICLR 2026] VEAttack: Downstream-agnostic Vision Encoder Attack against Large Vision Language Models
☆21Feb 22, 2026Updated 4 months ago
AI45Lab / ActorAttack
View on GitHub
☆135Jun 29, 2026Updated 3 weeks ago
shiyuchengTJU / PAR
View on GitHub
☆14Mar 23, 2023Updated 3 years ago
End-to-end encrypted email - Proton Mail • Ad
Special offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
UCSC-VLAA / AttnGCG-attack
View on GitHub
[TMLR 2025] Official implementation of AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation
☆27Jun 17, 2025Updated last year
HashmatShadab / Robust-LLaVA
View on GitHub
[ICCVW 2025 (Oral)] Robust-LLaVA: On the Effectiveness of Large-Scale Robust Image Encoders for Multi-modal Large Language Models
☆29Oct 20, 2025Updated 9 months ago
itsvaibhav01 / Immune
View on GitHub
[CVPR2025] Official Repository for IMMUNE: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment
☆28Jun 11, 2025Updated last year
weizeming / momentum-attack-llm
View on GitHub
☆25Jan 17, 2025Updated last year
DyMessi / VisCRA
View on GitHub
☆19Dec 23, 2025Updated 6 months ago
thunxxx / MLLM-Jailbreak-evaluation-MMJ-Bench
View on GitHub
☆81Mar 30, 2025Updated last year
lionel-w2 / FAP
View on GitHub
☆19Oct 20, 2024Updated last year