Codes for paper "SafeAgentBench: A Benchmark for Safe Task Planning of \\ Embodied LLM Agents"
☆77Feb 25, 2025Updated last year
Alternatives and similar repositories for SafeAgentBench
Users that are interested in SafeAgentBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the official repository for the ICLR 2025 accepted paper Badrobot: Manipulating Embodied LLMs in the Physical World.☆48Jun 11, 2026Updated 3 months ago
- ☆24Mar 6, 2025Updated last year
- Focused on the safety and security of Embodied AI☆113Jun 8, 2026Updated 3 months ago
- [NeurIPS 2025 Spotlight] Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning.☆158Mar 31, 2026Updated 5 months ago
- ☆102Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Responsible Robotic Manipulation☆17Aug 31, 2025Updated last year
- [ICLR 2025] Official codebase for the ICLR 2025 paper "Multimodal Situational Safety"☆37Jun 23, 2025Updated last year
- https://arxiv.org/pdf/2506.06677☆78Apr 29, 2026Updated 4 months ago
- [AAAI 2026] Data and Code for Paper IS-Bench: Evaluating Interactive Safety of VLM-Driven Embodied Agents in Daily Household Tasks☆52Nov 24, 2025Updated 10 months ago
- LoTa-Bench: Benchmarking Language-oriented Task Planners for Embodied Agents (ICLR 2024)☆94Feb 8, 2026Updated 7 months ago
- Benchmarking Physical Risk Awareness of Foundation Model-based Embodied AI Agents☆24Nov 28, 2024Updated last year
- The official code for "Steering Dialogue Dynamics for Robustness against Multi-turn Jailbreaking Attacks".☆19Jun 24, 2026Updated 3 months ago
- ☆82Dec 19, 2024Updated last year
- [ICML 2025 Oral] Official repo of EmbodiedBench, a comprehensive benchmark designed to evaluate MLLMs as embodied agents.☆344May 30, 2026Updated 3 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆41Feb 3, 2026Updated 7 months ago
- [ICLR 2026] The official code for "Doxing via the Lens: Revealing Location-related Privacy Leakage on Multi-modal Large Reasoning Models"☆31Feb 7, 2026Updated 7 months ago
- A vision-language-safety action architecture, named AEGIS, which contains a plug-and-play safety constraint layer formulated via control …☆132Updated this week
- ☆17Sep 25, 2024Updated 2 years ago
- The Oyster series is a set of safety models developed in-house by Alibaba-AAIG, devoted to building a responsible AI ecosystem. | Oyster …☆63Apr 29, 2026Updated 4 months ago
- Evaluate Multimodal LLMs as Embodied Agents☆60Feb 14, 2025Updated last year
- Official repo of Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics☆86Mar 26, 2026Updated 6 months ago
- Official implementation of Visco-Attack (EMNLP 2025 Main). An open-source one-click reproduction script is also provided.☆31Apr 11, 2026Updated 5 months ago
- Official code for "From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation" (ICLR2026)☆39Mar 1, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICCV 2025] RoboFactory: Exploring Embodied Agent Collaboration with Compositional Constraints☆145Sep 2, 2025Updated last year
- ☆13Aug 19, 2025Updated last year
- ☆134Jul 29, 2026Updated last month
- ☆17Mar 30, 2025Updated last year
- Code repository for the paper "Heuristic Induced Multimodal Risk Distribution Jailbreak Attack for Multimodal Large Language Models"☆20Aug 7, 2025Updated last year
- [ICML 2024] Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast☆130Mar 26, 2024Updated 2 years ago
- ☆23Jul 26, 2025Updated last year
- A Framework for Evaluating AI Agent Safety in Realistic Environments☆38Aug 10, 2026Updated last month
- ☆24Jun 16, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- AmpleGCG: Learning a Universal and Transferable Generator of Adversarial Attacks on Both Open and Closed LLM☆87Nov 3, 2024Updated last year
- ☆21Mar 26, 2026Updated 6 months ago
- Embodied Agent Interface (EAI): Benchmarking LLMs for Embodied Decision Making (NeurIPS D&B 2024 Oral)☆299Mar 6, 2025Updated last year
- Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses | 500+ Papers | Perception, Cognition, Planning, Interaction, Agentic Sys…☆143Updated this week
- [ICLR 2025] Dissecting adversarial robustness of multimodal language model agents☆144Feb 19, 2025Updated last year
- [ICML 2024] Code for the paper "MoE-RBench: Towards Building Reliable Language Models with Sparse Mixture-of-Experts"☆11Jul 1, 2024Updated 2 years ago
- ☆17Nov 24, 2025Updated 10 months ago