Revisiting Character-level Adversarial Attacks for Language Models, ICML 2024
☆20Aug 28, 2026Updated 3 weeks ago
Alternatives and similar repositories for Charmer
Users that are interested in Charmer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Implementation of implicit reference attack☆11Oct 16, 2024Updated last year
- About Official PyTorch implementation of "Query-Efficient Black-Box Red Teaming via Bayesian Optimization" (ACL'23)☆15Jul 9, 2023Updated 3 years ago
- ☆24Sep 20, 2023Updated 3 years ago
- Selective Copying Task with Mamba Model. This repository contains a simple implementation for reproducing the selective copying task with…☆16Jun 3, 2024Updated 2 years ago
- ☆12Mar 7, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [USENIX Security 2025] SOFT: Selective Data Obfuscation for Protecting LLM Fine-tuning against Membership Inference Attacks☆23Sep 18, 2025Updated last year
- [USENIX Security 2026] Membership Inference Attacks on Tokenizers of Large Language Models☆23May 22, 2026Updated 4 months ago
- Code and data for the ACM CIKM 2024 paper "Adversarial Text Rewriting for Text-aware Recommender Systems"☆13Aug 1, 2024Updated 2 years ago
- [Findings of ACL 2023] Bridge the Gap Between CV and NLP! A Optimization-based Textual Adversarial Attack Framework.☆14Aug 27, 2023Updated 3 years ago
- [NeurIPS 2024] Accelerating Greedy Coordinate Gradient and General Prompt Optimization via Probe Sampling☆35Nov 8, 2024Updated last year
- ☆20Dec 26, 2022Updated 3 years ago
- ☆18Mar 30, 2024Updated 2 years ago
- The official repository for guided jailbreak benchmark☆32Aug 31, 2026Updated 3 weeks ago
- Emoji Attack [ICML 2025]☆45Jul 15, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official implementation of NOSA☆19Jun 11, 2026Updated 3 months ago
- Test LLMs against jailbreaks and unprecedented harms☆41Oct 19, 2024Updated last year
- Adversarial Item Promotion in visually-aware recommenders☆16Sep 3, 2021Updated 5 years ago
- DySCO: Dynamic Attention-Scaling Decoding for Long-Context LMs☆18May 30, 2026Updated 3 months ago
- Supplementary files for SSFT 2015 summer school☆11Sep 5, 2019Updated 7 years ago
- Intrinsic Motivation and Automatic Curricula via Asymmetric Self-Play☆14May 1, 2018Updated 8 years ago
- This repo accompanies the paper "The Complexity Trap: Simple Observation Masking Is as Efficient as LLM Summarization for Agent Context M…☆19Nov 18, 2025Updated 10 months ago
- Learning Security Classifiers with Verified Global Robustness Properties (CCS'21) https://arxiv.org/pdf/2105.11363.pdf☆28Dec 1, 2021Updated 4 years ago
- Code for the paper "Multi-Field Adaptive Retrieval," a research project on a semi-structured document retrieval☆19Feb 13, 2026Updated 7 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Tensorflow implementation of TrialAttack (Triple Adversarial Learning for Influence based Poisoning Attack in Recommender Systems. KDD 20…☆11Sep 2, 2021Updated 5 years ago
- Pytorch implementation for the pilot study on the robustness of latent diffusion models.☆13Jun 20, 2023Updated 3 years ago
- This is the code repository for a project at Ulm University. It's a fall detection system based on address-event-based cameras.☆11Sep 29, 2017Updated 8 years ago
- Set-level Guidance Attack: Boosting Adversarial Transferability of Vision-Language Pre-training Models. [ICCV 2023 Oral]☆70Sep 6, 2023Updated 3 years ago
- [EMNLP 2024] Holistic Automated Red Teaming for Large Language Models through Top-Down Test Case Generation and Multi-turn Interaction☆17Nov 9, 2024Updated last year
- [KDD'21] Official PyTorch implementation for "Data Poisoning Attack against Recommender System Using Incomplete and Perturbed Data".☆13Sep 19, 2021Updated 5 years ago
- 🥇 Amazon Nova AI Challenge Winner - ASTRA emerged victorious as the top attacking team in Amazon's global AI safety competition, defeati…☆77May 11, 2026Updated 4 months ago
- Algorithms for byte-level language modelling☆25May 14, 2026Updated 4 months ago
- HeaderGen annotates Jupyter notebooks using static analysis. Improves PyCG's call graph analysis by supporting external libraries and flo…☆15Jan 30, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A list of research towards security&privacy in AI-Generated Content☆17Jan 10, 2025Updated last year
- ☆17Sep 25, 2024Updated last year
- bert-pli应用于LeCaRD☆18Nov 14, 2021Updated 4 years ago
- ☆15Dec 10, 2024Updated last year
- code for "Generative News Recommendation"☆15May 31, 2024Updated 2 years ago
- [ICLR 2025] Code implementation of R^2-Guard: Robust Reasoning Enabled LLM Guardrail via Knowledge-Enhanced Logical Reasoning☆24Jul 8, 2024Updated 2 years ago
- Official PyTorch Implementation of Federated Learning with Positive and Unlabeled Data☆10Aug 12, 2022Updated 4 years ago