Composite Backdoor Attacks Against Large Language Models
☆25Apr 12, 2024Updated 2 years ago
Alternatives and similar repositories for CBA
Users that are interested in CBA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆22Sep 16, 2022Updated 3 years ago
- Implementation of the paper "Exploring the Universal Vulnerability of Prompt-based Learning Paradigm" on Findings of NAACL 2022☆32Jul 11, 2022Updated 4 years ago
- ☆26Aug 21, 2024Updated last year
- Code for paper: PoisonPrompt: Backdoor Attack on Prompt-based Large Language Models, IEEE ICASSP 2024. Demo//124.220.228.133:11107☆21Aug 10, 2024Updated last year
- 🔥🔥🔥 Detecting hidden backdoors in Large Language Models with only black-box access☆57Jun 2, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Machine Learning & Security Seminar @Purdue University☆26May 9, 2023Updated 3 years ago
- A toolbox for backdoor attacks.☆23Jan 13, 2023Updated 3 years ago
- ☆23Aug 24, 2020Updated 5 years ago
- ☆19Nov 6, 2023Updated 2 years ago
- Official Repository of Paper "Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs"☆15Sep 25, 2025Updated 9 months ago
- Code used to run experiments for the ICLR 2023 paper "Computational Language Acquisition with Theory of Mind".☆15Apr 27, 2023Updated 3 years ago
- ☆11Dec 4, 2024Updated last year
- This repository contains the source code for "Membership Inference Attacks as Privacy Tools: Reliability, Disparity and Ensemble", In Pro…☆11Jan 2, 2026Updated 6 months ago
- This is the official Gtihub repo for our paper: "BEEAR: Embedding-based Adversarial Removal of Safety Backdoors in Instruction-tuned Lang…☆23Jul 3, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆29Aug 21, 2023Updated 2 years ago
- [AAAI'21] Deep Feature Space Trojan Attack of Neural Networks by Controlled Detoxification☆30Dec 31, 2024Updated last year
- ☆16Mar 31, 2025Updated last year
- Paper list of LLM fingerprinting, based on our paper titled "SoK: Large Language Model Copyright Auditing via Fingerprinting".☆29Aug 28, 2025Updated 10 months ago
- ☆27Nov 9, 2022Updated 3 years ago
- [NeurIPS 2025] BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models☆315Mar 13, 2026Updated 4 months ago
- ☆26Dec 1, 2022Updated 3 years ago
- ☆16Jun 3, 2025Updated last year
- [ICLR 2026 Oral] Invisible Safety Threat: Malicious Finetuning for LLM via Steganography☆20Mar 22, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A library for mechanistic anomaly detection☆22Jan 9, 2025Updated last year
- Gaussian Membership Inference Privacy (NeurIPS 2023)☆12Jul 27, 2024Updated last year
- ☆13Jun 1, 2024Updated 2 years ago
- Watermarking LLM papers up-to-date☆12Dec 17, 2023Updated 2 years ago
- Code Repository for the Paper ---Revisiting the Assumption of Latent Separability for Backdoor Defenses (ICLR 2023)☆47Feb 28, 2023Updated 3 years ago
- bert蒸馏实践,包含BiLSTM蒸馏BERT和TinyBert☆13Apr 23, 2022Updated 4 years ago
- [IEEE TIP] Offical implementation for the work "BadCM: Invisible Backdoor Attack against Cross-Modal Learning".☆14Aug 30, 2024Updated last year
- ☆15Dec 24, 2025Updated 6 months ago
- some baseline attack method by pytorch☆11Oct 13, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official repository for the paper "Problem space structural adversarial attacks for Network Intrusion Detection Systems based on Graph Ne…☆15Jul 31, 2024Updated last year
- Code and data of the ACL-IJCNLP 2021 paper "Hidden Killer: Invisible Textual Backdoor Attacks with Syntactic Trigger"☆46Sep 11, 2022Updated 3 years ago
- Contains implementation of denoising algorithms.☆11Jul 16, 2020Updated 6 years ago
- ☆16Jul 23, 2022Updated 3 years ago
- Official repository for CVPR'23 paper: Detecting Backdoors in Pre-trained Encoders☆39Sep 25, 2023Updated 2 years ago
- Robust Federated Learning for Large Language Models in Adversarial Wireless Environments☆16Mar 7, 2025Updated last year
- [NeurIPS'24] RedCode: Risky Code Execution and Generation Benchmark for Code Agents☆85Apr 24, 2026Updated 2 months ago