Code for paper "MAGIC: A Co-Evolving Attacker-Defender Adversarial Game for Robust LLM safety"
☆51May 11, 2026Updated 2 months ago
Alternatives and similar repositories for MAGIC
Users that are interested in MAGIC are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TRACE, a framework for turn-aware credit assignment for multi-turn jailbreak optimization☆20Jun 22, 2026Updated last month
- anti-老登,反登味的飞书机器人。拒绝内耗,从我做起,让职场再无登味☆29Apr 14, 2026Updated 3 months ago
- ☆29May 14, 2026Updated 2 months ago
- [🏆CVPR'26] Official Repo for IAG: Input-aware Backdoor Attack on VLM-based Visual Grounding☆33Jun 2, 2026Updated 2 months ago
- ☆28Apr 9, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Velaclaw — control plane for team AI☆121Jul 21, 2026Updated 2 weeks ago
- Build your agent from 200,000+ skills via skill RETRIEVAL & ORCHESTRATION☆564Mar 7, 2026Updated 5 months ago
- [CVPR 2026] SOTA Chemical Reaction Diagram Parsing Framework☆26Mar 24, 2026Updated 4 months ago
- [Pattern Recognition 2025 🌟]Unbiased Multiscale Modal Fusion Model for Multimodal Semantic Segmentation☆10Jun 12, 2024Updated 2 years ago
- [DAI 2025] Beyond GPT-5: Making LLMs Cheaper and Better via Performance–Efficiency Optimized Routing☆222Dec 11, 2025Updated 7 months ago
- A pure C++ library for simulations on complex networks. It follow the standard of C++11.☆14Nov 26, 2022Updated 3 years ago
- Open-source red teaming framework for MLLMs with 42+ attack methods☆261Jul 17, 2026Updated 3 weeks ago
- [AAAI 2026] The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants☆46Dec 11, 2025Updated 7 months ago
- A Diagnostic Guardrail Framework for AI Agent Safety and Security☆679Jun 8, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation☆54Jun 4, 2026Updated 2 months ago
- DARWIN is a self-evolving LLM jailbreak framework that grows a reusable strategy pool through external extraction, sandbox filtering, his…☆64Jun 5, 2026Updated 2 months ago
- ☆170Mar 20, 2026Updated 4 months ago
- Fine-Grained Pixel-Text Alignment for Open-Vocabulary Semantic Segmentation☆16Mar 28, 2026Updated 4 months ago
- The official implementation of A Unified Game-Theoretic Interpretation of Adversarial Robustness.☆22Jun 9, 2022Updated 4 years ago
- [ICLR 2026] TwinVLA : Data-Efficient Bimanual Manipulation with Twin Single-Arm Vision-Language-Action Models☆18May 29, 2026Updated 2 months ago
- MATLAB codes for several tensor based methods for hypergraph partitioning and subspace clustering☆15Sep 26, 2018Updated 7 years ago
- [AAAI 2026] The official code for ``LoopLLM: Transferable Energy-Latency Attacks in LLMs via Repetitive Generation''☆17Mar 20, 2026Updated 4 months ago
- [Findings@ACL'26] LLMRouterBench: A Massive Benchmark and Unified Framework for LLM Routing☆91Apr 6, 2026Updated 4 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- This repo is the official implementation of “Are Your Agents Upward Deceivers?”. The paper is accepted by ICML 2026.☆24Dec 15, 2025Updated 7 months ago
- Universal preflight security scanner for AI coding agents — Detects hooks injection, credential exfiltration & backdoors in .cursorrules,…☆75May 29, 2026Updated 2 months ago
- ☆18Mar 25, 2026Updated 4 months ago
- ☆27Mar 17, 2025Updated last year
- ☆24Nov 19, 2024Updated last year
- [ArXiv 2025] Imperceptible Jailbreaking against Large Language Models☆25Oct 7, 2025Updated 10 months ago
- Implementation of SBM-meet-GNN☆23May 12, 2019Updated 7 years ago
- ☆10May 31, 2018Updated 8 years ago
- Official code for FAccT'21 paper "Fairness Through Robustness: Investigating Robustness Disparity in Deep Learning" https://arxiv.org/abs…☆13Mar 9, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [🏆ECCV'26] Official Repo for SlowBA: An efficiency backdoor attack towards VLM-based GUI agents☆18Jul 1, 2026Updated last month
- All-in-One Safety Evaluation Framwork☆52Jul 15, 2026Updated 3 weeks ago
- Implementation for paper Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs, which is accepted by ACL 2026 (main con…☆16Oct 10, 2025Updated 10 months ago
- Solving the OpenAI Gym (MountainCarContinuous-v0) with DDPG☆21Jan 23, 2023Updated 3 years ago
- Agent Security Bench (ASB)☆277Apr 16, 2026Updated 3 months ago
- Environments to support https://github.com/sholtodouglas/learning_from_play and reinforcement learning for robotic manipulation.☆21Mar 28, 2021Updated 5 years ago
- MinT-2M: Long-context training system for resident-prefix GRPO☆44Jul 24, 2026Updated 2 weeks ago