Effective Prompt Extraction from Language Models
☆43Sep 10, 2024Updated last year
Alternatives and similar repositories for prompt-extraction
Users that are interested in prompt-extraction are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Advanced GUI agents☆17Feb 3, 2026Updated 6 months ago
- ☆82Dec 19, 2024Updated last year
- 🔥🔥🔥 Detecting hidden backdoors in Large Language Models with only black-box access☆58Jun 2, 2025Updated last year
- ☆45Mar 3, 2023Updated 3 years ago
- Source code of NAACL 2025 Findings "Scaling Up Membership Inference: When and How Attacks Succeed on Large Language Models"☆16Dec 16, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code associated with ICML (2024). "Defense against Backdoor Attack on Pre-trained Language Models via Head Pruning and Attention Normaliz…☆11Feb 22, 2026Updated 6 months ago
- ☆23Updated this week
- Two-party Privacy-preserving Neural Network Training using Split Learning and Homomorphic Encryption (CKKS Scheme)☆13Sep 23, 2025Updated 11 months ago
- ☆14Feb 21, 2025Updated last year
- ☆14May 23, 2023Updated 3 years ago
- Code release for MPCViT accepted by ICCV 2023☆17Jan 6, 2025Updated last year
- Provably Secure Steganography☆22Sep 13, 2025Updated 11 months ago
- ☆23Jul 26, 2025Updated last year
- ☆16Feb 26, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆17May 23, 2025Updated last year
- Official implementation of Cross-Modal Unlearning via Influential Neuron Path Editing in Multimodal Large Language Models☆16Mar 21, 2026Updated 5 months ago
- [AAAI 2024] Data-Free Hard-Label Robustness Stealing Attack☆16Mar 29, 2024Updated 2 years ago
- Internal Consistency Regularization (CROW) for LLM Backdoor Elimination - Paper accepted to ICML 2025☆16May 6, 2025Updated last year
- ☆18Aug 6, 2025Updated last year
- Privacy-Preserving Verifiable Neural Network Inference Service (ACSAC 2024)☆17Sep 6, 2025Updated 11 months ago
- TextGuard: Provable Defense against Backdoor Attacks on Text Classification☆15Nov 7, 2023Updated 2 years ago
- The implementation for paper "UniGuardian: A Unified Defense for Detecting Prompt Injection, Backdoor Attacks and Adversarial Attacks in …☆17Jul 3, 2025Updated last year
- ☆74May 23, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆38Oct 17, 2024Updated last year
- The implementation of the IEEE S&P 2024 paper MM-BD: Post-Training Detection of Backdoor Attacks with Arbitrary Backdoor Pattern Types Us…☆16May 12, 2024Updated 2 years ago
- Official implementation of the EMNLP 2021 paper "ONION: A Simple and Effective Defense Against Textual Backdoor Attacks"☆40Nov 3, 2021Updated 4 years ago
- ☆24Apr 25, 2024Updated 2 years ago
- [USENIX Security 2025] SOFT: Selective Data Obfuscation for Protecting LLM Fine-tuning against Membership Inference Attacks☆23Sep 18, 2025Updated 11 months ago
- Official repository for PEFTGuard: Detecting Backdoor Attacks Against Parameter-Efficient Fine-Tuning, accepted at 2025 IEEE Symposium on…☆18Jul 4, 2025Updated last year
- ICDE 2025 Paper, Grounding Natural Language to SQL Translation with Data-Based Self-Explanations☆17May 24, 2025Updated last year
- TAOISM: A TEE-based Confidential Heterogeneous Deployment Framework for DNN Models☆52Apr 11, 2024Updated 2 years ago
- Code for the paper "RAP: Robustness-Aware Perturbations for Defending against Backdoor Attacks on NLP Models" (EMNLP 2021)☆25Oct 21, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Parallel Navier-Stokes solver for coupled diffusion-convection PDEs☆14Nov 17, 2022Updated 3 years ago
- Code for Findings-EMNLP 2023 paper: Multi-step Jailbreaking Privacy Attacks on ChatGPT☆37Oct 15, 2023Updated 2 years ago
- PAL: Proxy-Guided Black-Box Attack on Large Language Models☆57Aug 17, 2024Updated 2 years ago
- MMLU eval for RU/EN☆16Jul 31, 2023Updated 3 years ago
- A forkable Next.js template featuring a design canvas UI with AI integration. Build your own Canva, Figma, or tldraw alternative.☆36Updated this week
- Python wrapper for phantomjs☆15May 28, 2021Updated 5 years ago
- Code for Findings of ACL 2021 "Differential Privacy for Text Analytics via Natural Text Sanitization"☆34Mar 15, 2022Updated 4 years ago