Codes for paper "SafeAgentBench: A Benchmark for Safe Task Planning of \\ Embodied LLM Agents"
☆76Feb 25, 2025Updated last year
Alternatives and similar repositories for SafeAgentBench
Users that are interested in SafeAgentBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the official repository for the ICLR 2025 accepted paper Badrobot: Manipulating Embodied LLMs in the Physical World.☆48Jun 11, 2026Updated 2 months ago
- ☆25Mar 6, 2025Updated last year
- Focused on the safety and security of Embodied AI☆113Jun 8, 2026Updated 2 months ago
- [NeurIPS 2025 Spotlight] Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning.☆157Mar 31, 2026Updated 5 months ago
- ☆102Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Responsible Robotic Manipulation☆17Aug 31, 2025Updated last year
- [ICLR 2025] Official codebase for the ICLR 2025 paper "Multimodal Situational Safety"☆37Jun 23, 2025Updated last year
- https://arxiv.org/pdf/2506.06677☆77Apr 29, 2026Updated 4 months ago
- [AAAI 2026] Data and Code for Paper IS-Bench: Evaluating Interactive Safety of VLM-Driven Embodied Agents in Daily Household Tasks☆52Nov 24, 2025Updated 9 months ago
- LoTa-Bench: Benchmarking Language-oriented Task Planners for Embodied Agents (ICLR 2024)☆94Feb 8, 2026Updated 6 months ago
- Code for IROS 2024 paper "AutoNeRF: Training Implicit Scene Representations with Autonomous Agents"☆17Oct 24, 2024Updated last year
- Benchmarking Physical Risk Awareness of Foundation Model-based Embodied AI Agents☆24Nov 28, 2024Updated last year
- The official code for "Steering Dialogue Dynamics for Robustness against Multi-turn Jailbreaking Attacks".☆19Jun 24, 2026Updated 2 months ago
- [ICML 2025 Oral] Official repo of EmbodiedBench, a comprehensive benchmark designed to evaluate MLLMs as embodied agents.☆340May 30, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The official implementation of NeurlPS 2025 D&B paper: IndustryEQA: Pushing the frontiers of Embodied Question Answering in Industrial Sc…☆16Sep 25, 2025Updated 11 months ago
- ☆43Feb 3, 2026Updated 7 months ago
- [ICLR 2026] The official code for "Doxing via the Lens: Revealing Location-related Privacy Leakage on Multi-modal Large Reasoning Models"☆31Feb 7, 2026Updated 7 months ago
- A vision-language-safety action architecture, named AEGIS, which contains a plug-and-play safety constraint layer formulated via control …☆130Jun 30, 2026Updated 2 months ago
- ☆17Sep 25, 2024Updated last year
- The Oyster series is a set of safety models developed in-house by Alibaba-AAIG, devoted to building a responsible AI ecosystem. | Oyster …☆62Apr 29, 2026Updated 4 months ago
- Evaluate Multimodal LLMs as Embodied Agents☆60Feb 14, 2025Updated last year
- Official repo of Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics☆84Mar 26, 2026Updated 5 months ago
- ☆44Jan 13, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official code for "From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation" (ICLR2026)☆39Mar 1, 2026Updated 6 months ago
- [ICCV 2025] RoboFactory: Exploring Embodied Agent Collaboration with Compositional Constraints☆142Sep 2, 2025Updated last year
- Official repository for "On the Multi-modal Vulnerability of Diffusion Models"☆17Jul 15, 2024Updated 2 years ago
- [EMNLP 2025] Reasoning-to-Defend: Safety-Aware Reasoning Can Defend Large Language Models from Jailbreaking☆12Aug 22, 2025Updated last year
- ☆13Aug 19, 2025Updated last year
- ☆132Jul 29, 2026Updated last month
- Code repository for the paper "Heuristic Induced Multimodal Risk Distribution Jailbreak Attack for Multimodal Large Language Models"☆20Aug 7, 2025Updated last year
- ☆23Jul 26, 2025Updated last year
- Official GitHub repository for the paper "Adversarial Attacks on Robotic Vision Language Action Models"☆35May 28, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Framework for Evaluating AI Agent Safety in Realistic Environments☆38Aug 10, 2026Updated 3 weeks ago
- AmpleGCG: Learning a Universal and Transferable Generator of Adversarial Attacks on Both Open and Closed LLM☆87Nov 3, 2024Updated last year
- ☆21Mar 26, 2026Updated 5 months ago
- Embodied Agent Interface (EAI): Benchmarking LLMs for Embodied Decision Making (NeurIPS D&B 2024 Oral)☆299Mar 6, 2025Updated last year
- ☆25Jun 9, 2026Updated 2 months ago
- Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses | 500+ Papers | Perception, Cognition, Planning, Interaction, Agentic Sys…☆135Updated this week
- [ICLR 2025] Dissecting adversarial robustness of multimodal language model agents☆144Feb 19, 2025Updated last year