Codes for paper "SafeAgentBench: A Benchmark for Safe Task Planning of \\ Embodied LLM Agents"
☆74Feb 25, 2025Updated last year
Alternatives and similar repositories for SafeAgentBench
Users that are interested in SafeAgentBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the official repository for the ICLR 2025 accepted paper Badrobot: Manipulating Embodied LLMs in the Physical World.☆46Jun 11, 2026Updated last month
- ☆15Dec 24, 2025Updated 7 months ago
- ☆24Mar 6, 2025Updated last year
- Focused on the safety and security of Embodied AI☆113Jun 8, 2026Updated last month
- [NeurIPS 2025 Spotlight] Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning.☆154Mar 31, 2026Updated 3 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆102Updated this week
- Responsible Robotic Manipulation☆16Aug 31, 2025Updated 10 months ago
- https://arxiv.org/pdf/2506.06677☆73Apr 29, 2026Updated 3 months ago
- [AAAI 2026] Data and Code for Paper IS-Bench: Evaluating Interactive Safety of VLM-Driven Embodied Agents in Daily Household Tasks☆47Nov 24, 2025Updated 8 months ago
- LoTa-Bench: Benchmarking Language-oriented Task Planners for Embodied Agents (ICLR 2024)☆94Feb 8, 2026Updated 5 months ago
- Code for IROS 2024 paper "AutoNeRF: Training Implicit Scene Representations with Autonomous Agents"☆17Oct 24, 2024Updated last year
- Benchmarking Physical Risk Awareness of Foundation Model-based Embodied AI Agents☆23Nov 28, 2024Updated last year
- The official code for "Steering Dialogue Dynamics for Robustness against Multi-turn Jailbreaking Attacks".☆18Jun 24, 2026Updated last month
- ☆82Dec 19, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆41Feb 3, 2026Updated 5 months ago
- [ICLR 2026] The official code for "Doxing via the Lens: Revealing Location-related Privacy Leakage on Multi-modal Large Reasoning Models"☆30Feb 7, 2026Updated 5 months ago
- A vision-language-safety action architecture, named AEGIS, which contains a plug-and-play safety constraint layer formulated via control …☆115Jun 30, 2026Updated 3 weeks ago
- ☆17Sep 25, 2024Updated last year
- The Oyster series is a set of safety models developed in-house by Alibaba-AAIG, devoted to building a responsible AI ecosystem. | Oyster …☆62Apr 29, 2026Updated 3 months ago
- Evaluate Multimodal LLMs as Embodied Agents☆59Feb 14, 2025Updated last year
- Official repo of Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics☆82Mar 26, 2026Updated 4 months ago
- Official implementation of Visco-Attack (EMNLP 2025 Main). An open-source one-click reproduction script is also provided.☆31Apr 11, 2026Updated 3 months ago
- ☆45Jan 13, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official code for "From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation" (ICLR2026)☆37Mar 1, 2026Updated 4 months ago
- [ICCV 2025] RoboFactory: Exploring Embodied Agent Collaboration with Compositional Constraints☆137Sep 2, 2025Updated 10 months ago
- Official repository for "On the Multi-modal Vulnerability of Diffusion Models"☆17Jul 15, 2024Updated 2 years ago
- [EMNLP 2025] Reasoning-to-Defend: Safety-Aware Reasoning Can Defend Large Language Models from Jailbreaking☆12Aug 22, 2025Updated 11 months ago
- ☆12Aug 19, 2025Updated 11 months ago
- ☆134Jun 29, 2026Updated last month
- ☆18Mar 30, 2025Updated last year
- ☆10May 8, 2024Updated 2 years ago
- [ICML 2024] Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast☆123Mar 26, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NeurIPS 2025] Official repository of RiOSWorld: Benchmarking the Risk of Multimodal Computer-Use Agents☆123Dec 2, 2025Updated 7 months ago
- ☆22Jul 26, 2025Updated last year
- Official GitHub repository for the paper "Adversarial Attacks on Robotic Vision Language Action Models"☆35May 28, 2025Updated last year
- A Framework for Evaluating AI Agent Safety in Realistic Environments☆38Oct 2, 2025Updated 9 months ago
- ☆23Jun 16, 2025Updated last year
- ☆24Jun 9, 2026Updated last month
- AmpleGCG: Learning a Universal and Transferable Generator of Adversarial Attacks on Both Open and Closed LLM☆87Nov 3, 2024Updated last year