Dataset and evaluation benchmark for Privacy Leakage Evaluation of Autonomous Web Agents
☆45Apr 18, 2026Updated 4 months ago
Alternatives and similar repositories for ai-agent-privacy
Users that are interested in ai-agent-privacy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A research workbench for developing and testing attacks against large language models, with a focus on prompt injection vulnerabilities a…☆61Jul 24, 2026Updated 3 weeks ago
- Official implementation of the WASP web agent security benchmark☆98Apr 13, 2026Updated 4 months ago
- A curated collection of papers and related projects on using LLMs for privacy.☆36Oct 8, 2025Updated 10 months ago
- BigOBench assesses the capacity of Large Language Models (LLMs) to comprehend time-space computational complexity of input or generated c…☆43Apr 15, 2025Updated last year
- [IEEE S&P'26] WebCloak: Characterizing and Mitigating the Threats of LLM-Driven Web Agents as Intelligent Scrapers☆29Jan 31, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- EgoToM is an egocentric theory-of-mind benchmark built on Ego4D videos, containing multi-choice questions that evaluate multimodal large …☆17Apr 1, 2025Updated last year
- Source code of "Leaky Thoughts: Large Reasoning Models Are Not Private Thinkers" EMNLP 2025☆17Jan 12, 2026Updated 7 months ago
- Code for the paper "Watermarking Makes Language Models Radioactive"☆26Oct 25, 2024Updated last year
- Code for the paper "Firewalls to Secure Dynamic LLM Agentic Networks"☆30Jun 6, 2025Updated last year
- ☆35Feb 8, 2026Updated 6 months ago
- Official release of code for the paper RL is a hammer and LLMs are nails A simple RL approach to stronger prompt injection attacks☆53May 6, 2026Updated 3 months ago
- ☆17Feb 6, 2024Updated 2 years ago
- ☆14Nov 22, 2024Updated last year
- Code repo for the paper: Attacking Vision-Language Computer Agents via Pop-ups☆52Dec 23, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A Dynamic Environment to Evaluate Attacks and Defenses for LLM Agents.☆753Jun 2, 2026Updated 2 months ago
- Patch management tool for git submodules☆16Jul 25, 2025Updated last year
- Generative Visual Code Mobile World Model☆61May 15, 2026Updated 3 months ago
- Repo for the paper "Meta SecAlign: A Secure Foundation LLM Against Prompt Injection Attacks".☆70Jun 11, 2026Updated 2 months ago
- ☆41Aug 11, 2026Updated last week
- Source code of NAACL 2025 Findings "Scaling Up Membership Inference: When and How Attacks Succeed on Large Language Models"☆16Dec 16, 2025Updated 8 months ago
- ☆16Mar 13, 2025Updated last year
- Real-time visualisation☆35Updated this week
- The code implementation of GraCeFul (Accepted in COLING 2025)☆13Jan 27, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆20Jul 22, 2025Updated last year
- Code for the paper "The Journey, Not the Destination: How Data Guides Diffusion Models"☆26Dec 12, 2023Updated 2 years ago
- ☆21May 1, 2026Updated 3 months ago
- Repo for the research paper "SecAlign: Defending Against Prompt Injection with Preference Optimization"☆100Jul 2, 2026Updated last month
- Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks☆91Aug 3, 2026Updated 2 weeks ago
- Demo web client for using matrix-files-sdk☆17Jan 31, 2025Updated last year
- ☆42May 21, 2025Updated last year
- Simulation Streams is a programming paradigm designed to efficiently control and leverage Large Language Models (LLMs) for complex, dynam…☆27Jul 2, 2026Updated last month
- ☆12Sep 8, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- DoomArena is a Framework for Testing AI Agents Against Evolving Security Threats☆62Sep 12, 2025Updated 11 months ago
- [ACL'26] Official Repository for The Paper: What If Consensus Lies? Selective-Complementary Reinforcement Learning at Test Time☆18Apr 7, 2026Updated 4 months ago
- [ICLR'21] Dataset Inference for Ownership Resolution in Machine Learning☆31Oct 10, 2022Updated 3 years ago
- A toolkit to assess data privacy in LLMs (under development)☆75Jan 2, 2025Updated last year
- The open-source materials for paper "Sparsing Law: Towards Large Language Models with Greater Activation Sparsity".☆32Nov 12, 2024Updated last year
- Official implementation of AdvPrompter https//arxiv.org/abs/2404.16873☆182May 6, 2024Updated 2 years ago
- ☆15Updated this week