☆27Nov 20, 2023Updated 2 years ago
Alternatives and similar repositories for toxic-prompt
Users that are interested in toxic-prompt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Dec 30, 2021Updated 4 years ago
- ☆16Jul 26, 2024Updated 2 years ago
- [S&P'24] Test-Time Poisoning Attacks Against Test-Time Adaptation Models☆21Feb 18, 2025Updated last year
- ☆16Mar 5, 2026Updated 6 months ago
- [CCS'22] SSLGuard: A Watermarking Scheme for Self-supervised Learning Pre-trained Encoders☆18Jul 12, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆36May 22, 2024Updated 2 years ago
- ☆15Feb 13, 2026Updated 6 months ago
- ☆13Feb 17, 2025Updated last year
- ☆71Feb 4, 2024Updated 2 years ago
- Github implementation of https://reports.chatclimate.ai/☆24Jun 16, 2025Updated last year
- [Preprint] On the Effectiveness of Mitigating Data Poisoning Attacks with Gradient Shaping☆10Feb 27, 2020Updated 6 years ago
- This repository contains the dataset and implementation details of the paper "An In-depth Analysis of Implicit and Subtle Hate Speech Mes…☆10May 9, 2024Updated 2 years ago
- ☆15Jun 4, 2024Updated 2 years ago
- ☆29Aug 21, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18Jul 1, 2021Updated 5 years ago
- ☆13Oct 20, 2022Updated 3 years ago
- Caffe code for the paper "Adversarial Manipulation of Deep Representations"☆17Nov 6, 2017Updated 8 years ago
- ☆25Jun 23, 2021Updated 5 years ago
- The implement of LLMTreeRec☆14Dec 9, 2024Updated last year
- Code for "CloudLeak: Large-Scale Deep Learning Models Stealing Through Adversarial Examples" (NDSS 2020)☆22Nov 14, 2020Updated 5 years ago
- Code for Backdoor Attacks Against Dataset Distillation☆37Apr 19, 2023Updated 3 years ago
- Synthetic data generation for TODs☆23Jul 17, 2024Updated 2 years ago
- ☆14Apr 11, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- AutoCAT: Reinforcement Learning for Automated Exploration of Cache-Timing Attacks☆48May 19, 2023Updated 3 years ago
- ☆26Nov 21, 2020Updated 5 years ago
- Insurance-RAG-Chatbot(IVA): An open-source project featuring a retrieval-augmented chatbot developed using Bedrock, LLM, LangChain, Docke…☆23May 30, 2024Updated 2 years ago
- TextHide: Tackling Data Privacy in Language Understanding Tasks☆30Apr 19, 2021Updated 5 years ago
- Modular Adversarial Robustness Toolkit☆21Jul 13, 2026Updated last month
- template for https://cnli.me☆10Feb 27, 2025Updated last year
- ☆11Oct 2, 2023Updated 2 years ago
- A project demonstrating how to improve the model accuracy by suppression the false postive using an assessor model☆10Aug 13, 2021Updated 5 years ago
- Mine conversations from novels in Project Gutenberg, to generate data for data-driven dialogue systems.☆15May 7, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆25Nov 14, 2022Updated 3 years ago
- Official implementation of ICLR'24 paper, "Curiosity-driven Red Teaming for Large Language Models" (https://openreview.net/pdf?id=4KqkizX…☆90Mar 15, 2024Updated 2 years ago
- The code and resource of "Towards Comprehensive Detection of Chinese Harmful Memes" (NeurIPS2024 D&B).☆87May 17, 2025Updated last year
- 🤫 Code and benchmark for our ICLR 2024 spotlight paper: "Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Con…☆57Dec 20, 2023Updated 2 years ago
- Local Discriminative Regions for Scene Recognition (ACMMM 2018)☆22Oct 3, 2023Updated 2 years ago
- ☆165Jan 24, 2025Updated last year
- Fortifying Toxic Speech Detectors Against Veiled Toxicity☆11Oct 21, 2020Updated 5 years ago