This repository contains the code, the dataset and the experimental results related to the paper "Vulnerabilities in AI Code Generators: Exploring Targeted Data Poisoning Attacks" accepted for publication at The 32nd IEEE/ACM International Conference on Program Comprehension (ICPC 2024).
☆14Aug 5, 2024Updated last year
Alternatives and similar repositories for Targeted-Data-Poisoning-Attacks
Users that are interested in Targeted-Data-Poisoning-Attacks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repository for code and presentations from the Windows Server 2016 MegaSeries Videos☆12Apr 26, 2021Updated 5 years ago
- A FaceApp like app☆10Mar 31, 2019Updated 7 years ago
- ☆12Nov 28, 2025Updated 7 months ago
- Improving Code Readability Classification using Convolutional Neural Networks☆10Apr 18, 2018Updated 8 years ago
- ReadMe++: A Multi-domain Multilingual Dataset for Readability Assessment☆13Apr 15, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Proof of concept code for poisoning code generation models.☆59Dec 6, 2023Updated 2 years ago
- R package for analyzing and manipulating multilabel datasets☆24Dec 10, 2019Updated 6 years ago
- ☆23Oct 21, 2024Updated last year
- JEMMA: An Extensible Java dataset for Many ML4Code Applications☆19Dec 12, 2022Updated 3 years ago
- ESEC/FSE'21: Prediction-Preserving Program Simplification☆10Oct 4, 2022Updated 3 years ago
- Adversarial Attack for Pre-trained Code Models☆10Jul 19, 2022Updated 4 years ago
- The project aims to evaluate the vulnerability of Federated Learning systems to targeted data poisoning attack known as Label Flipping At…☆24Jan 5, 2022Updated 4 years ago
- Code associated with ICML (2024). "Defense against Backdoor Attack on Pre-trained Language Models via Head Pruning and Attention Normaliz…☆11Feb 22, 2026Updated 5 months ago
- Exploring aspects of similarity between spoken personal narratives by disentangling them into narrative clause types -- Supplementary inf…☆12Jul 14, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- learning to protect communications with adversarial neural cryptography☆20Aug 15, 2018Updated 7 years ago
- Dataset and pre-trained model of EMNLP-IJCNLP 2019 paper "TalkDown: A Corpus for Condescension Detection in Context."☆10Jan 26, 2020Updated 6 years ago
- ☆15Apr 25, 2026Updated 3 months ago
- Data and code for "Chain-of-Thought in Neural Code Generation: From and For Lightweight Language Models", which accepted in TSE.☆15Jul 3, 2024Updated 2 years ago
- Backdooring Neural Code Search☆14Sep 8, 2023Updated 2 years ago
- Show Me Your Code! Kill Code Poisoning: A Lightweight Method Based on Code Naturalness☆19Jul 17, 2025Updated last year
- [ICLR 2025] REFINE: Inversion-Free Backdoor Defense via Model Reprogramming☆13Feb 13, 2025Updated last year
- The replication package of paper "Pre-training by Predicting Program Dependencies for Vulnerability Analysis Tasks"☆28Jan 3, 2024Updated 2 years ago
- ☆14Jul 27, 2020Updated 5 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- plots krautchan /int/ post locations on a worldmap☆16Jun 18, 2012Updated 14 years ago
- All source codes used in "얀데레 아가씨"☆11Aug 6, 2017Updated 8 years ago
- ☆11Apr 9, 2024Updated 2 years ago
- Codebase describing experiments in Truncation Sampling as Language Model Desmoothing☆13Dec 6, 2022Updated 3 years ago
- ☆15Dec 12, 2023Updated 2 years ago
- [ACL 2024] Official Repository for "Leveraging Large Language Models for Learning Complex Legal Concepts through Storytelling"☆16Aug 19, 2024Updated last year
- Code and data for "A fine-grained comparison of pragmatic language understanding in humans and language models"☆11Dec 14, 2022Updated 3 years ago
- ☆28May 9, 2026Updated 2 months ago
- ☆15Feb 26, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [AAAI 2026] This is the official implementation of the paper "ExtendAttack: Attacking Servers of LRMs via Extending Reasoning".☆25Mar 18, 2026Updated 4 months ago
- code and data associated with CoMPosT: Characterizing and Evaluating Caricature in LLM Simulations☆11Oct 13, 2023Updated 2 years ago
- A collection of small corpuses of interesting data for the creation of bots and similar stuff.☆10Sep 26, 2018Updated 7 years ago
- ☆46Jun 21, 2022Updated 4 years ago
- Code for the arXiv preprint "Answer, Assemble, Ace: Understanding How Transformers Answer Multiple Choice Questions"☆15Aug 2, 2025Updated 11 months ago
- ☆12Mar 12, 2022Updated 4 years ago
- ☆18Aug 15, 2022Updated 3 years ago