☆21Dec 15, 2025Updated 7 months ago
Alternatives and similar repositories for TAMAS
Users that are interested in TAMAS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Evaluating Agent Safety in Realistic, High-Risk Simulations☆31Jul 6, 2026Updated 2 weeks ago
- The official implementation of the paper "AgentLAB: Benchmarking LLM Agents against Long-Horizon Attacks"☆26Jun 1, 2026Updated last month
- ☆41Jun 28, 2025Updated last year
- AgentLeak: Open benchmark for privacy leakage in LLM agents — 7 channels, multi-agent, multi-framework.☆25Jul 1, 2026Updated 2 weeks ago
- [ICML 2023] Protecting Language Generation Models via Invisible Watermarking☆13Sep 8, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆12Jul 15, 2020Updated 6 years ago
- DiffWA: Diffusion Models for Watermark Attack☆10Apr 23, 2024Updated 2 years ago
- ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis☆33Jul 10, 2026Updated last week
- Multi-view Reinforcement Learning☆11Feb 9, 2020Updated 6 years ago
- ☆39May 29, 2026Updated last month
- TensorFlow implementation of "A Relational Intervention Approach for Unsupervised Dynamics Generalization in Model-Based Reinforcement Le…☆16Jul 2, 2022Updated 4 years ago
- ☆18May 18, 2025Updated last year
- ☆13Jul 2, 2020Updated 6 years ago
- Code repository for the paper "Heuristic Induced Multimodal Risk Distribution Jailbreak Attack for Multimodal Large Language Models"☆19Aug 7, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The code for paper 'STAS: Spatial-Temporal Return Decomposition for Multi-agent Reinforcement Learning'☆17Oct 6, 2024Updated last year
- [ICLR 2025] FLAT: LLM Unlearning via Loss Adjustment with Only Forget Data☆14Feb 26, 2025Updated last year
- ☆12Mar 3, 2025Updated last year
- This repository contains data and code used for On the Risk of Misinformation Pollution with Large Language Models (EMNLP 2023 Findings).☆17Dec 14, 2023Updated 2 years ago
- [ICML'25] MELON: Provable Defense Against Indirect Prompt Injection Attacks in AI Agents☆36Jul 31, 2025Updated 11 months ago
- Repository for the paper "MALADE: Orchestration of LLM-powered Agents with Retrieval Augmented Generation for Pharmacovigilance"☆29Feb 19, 2025Updated last year
- LLMs + Persona-Plug = Personalized LLMs☆15Oct 16, 2024Updated last year
- Author's PyTorch Implementation of Deep Homomorphic Policy Gradient (DHPG) - NeurIPS 2022 and JMLR 2024☆24Apr 8, 2024Updated 2 years ago
- SkillJect: Automating Stealthy Skill-Based Prompt Injection for Coding Agents with Trace-Driven Closed-Loop Refinement☆73Jun 11, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [NeurIPS'24] RedCode: Risky Code Execution and Generation Benchmark for Code Agents☆85Apr 24, 2026Updated 2 months ago
- WSDM 2021 Tutorial on Advances in Bias-aware Recommendation on the Web☆11Mar 8, 2021Updated 5 years ago
- ☆17Sep 2, 2025Updated 10 months ago
- AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems☆16May 12, 2026Updated 2 months ago
- ☆21May 15, 2024Updated 2 years ago
- ☆35Feb 17, 2026Updated 5 months ago
- Attack-Inspired GAN - unofficial pytorch implementation☆18Jun 10, 2023Updated 3 years ago
- [CVPR 2026] LLaVAShield: Safeguarding Multimodal Multi-Turn Dialogues in Vision-Language Models☆16Jun 26, 2026Updated 3 weeks ago
- ☆18Dec 21, 2025Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [AAAI 2024] Data-Free Hard-Label Robustness Stealing Attack☆16Mar 29, 2024Updated 2 years ago
- The Official Repository for Paper "HarmfulSkillBench: How Do Harmful Skills Weaponize Your Agents?"☆15May 2, 2026Updated 2 months ago
- ☆32Oct 18, 2024Updated last year
- Skill retrieval benchmark dataset and evaluation code.☆20May 8, 2026Updated 2 months ago
- Official code for our paper "SoK: Large Language Model Copyright Auditing via Fingerprinting"☆18Dec 31, 2025Updated 6 months ago
- Knowledge graph Entity and Word Embeddings for Retrieval☆11Nov 19, 2021Updated 4 years ago
- EXPERIMENTAL PROTOTYPE code for "Bolt-on Causal Consistency" appearing in SIGMOD 2013☆12Nov 2, 2013Updated 12 years ago