[ACL 2025 Main] Code and data for paper "Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?"
☆23Jun 18, 2025Updated last year
Alternatives and similar repositories for Watermark-Radioactivity-Attack
Users that are interested in Watermark-Radioactivity-Attack are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆22Mar 19, 2024Updated 2 years ago
- Code and data for paper "Can Watermarked LLMs be Identified by Users via Crafted Prompts?" Accepted by ICLR 2025 (Spotlight)☆28Dec 28, 2024Updated last year
- Repository for Towards Codable Watermarking for Large Language Models☆37Sep 20, 2023Updated 2 years ago
- [ICLR 2025] Permute-and-Flip: An optimally robust and watermarkable decoder for LLMs☆19Mar 20, 2025Updated last year
- [ICML2024] Adaptive Text Watermark for Large Language Models☆26Dec 11, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [USENIX Security'24] REMARK-LLM: A robust and efficient watermarking framework for generative large language models☆28Oct 23, 2024Updated last year
- ☆40Aug 12, 2024Updated 2 years ago
- Official repository for "PostMark: A Robust Blackbox Watermark for Large Language Models"☆29Aug 30, 2024Updated 2 years ago
- [ICLR 2024] Code and data for paper "A Semantic Invariant Robust Watermark for Large Language Models".☆39Nov 13, 2024Updated last year
- [CCS-LAMPS'24] LLM IP Protection Against Model Merging☆16Oct 14, 2024Updated last year
- Code repo for "Harnessing Negative Signals: Reinforcement Distillation from Teacher Data for LLM Reasoning"☆34Jul 25, 2025Updated last year
- [ICLR 2024] Source code of paper "An Unforgeable Publicly Verifiable Watermark for Large Language Models"☆34May 23, 2024Updated 2 years ago
- Code for paper "Concrete Subspace Learning based Interference Elimination for Multi-task Model Fusion"☆14Mar 28, 2024Updated 2 years ago
- [ACL2024-Main] Data and Code for WaterBench: Towards Holistic Evaluation of LLM Watermarks☆33Nov 14, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official Inplementation of CVPR23 paper "Backdoor Defense via Deconfounded Representation Learning"☆25Mar 13, 2023Updated 3 years ago
- Official implementation for "HuRef: HUman-REadable Fingerprint for Large Language Models" (NeurIPS2024)☆16Jun 17, 2025Updated last year
- Official repository of the paper: Who Wrote this Code? Watermarking for Code Generation (ACL 2024)☆40May 28, 2024Updated 2 years ago
- Submodule of evalverse forked from [google-research/instruction_following_eval](https://github.com/google-research/google-research/tree/m…☆15May 4, 2024Updated 2 years ago
- [ACM MM 2026 Oral] Code for paper "Omni-SafetyBench: A Benchmark for Safety Evaluation of Audio-Visual Large Language Models".☆115Aug 26, 2026Updated last week
- ☆15Apr 6, 2026Updated 5 months ago
- ☆12Mar 7, 2024Updated 2 years ago
- Evaluating Durability: Benchmark Insights into Multimodal Watermarking☆12Jun 7, 2024Updated 2 years ago
- ☆17Nov 7, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A collection list for Large Language Model (LLM) Watermark☆68Mar 30, 2026Updated 5 months ago
- Code for "Goodtriever: Toxicity Mitigation with Retrieval-augmented Language Models"☆25May 30, 2024Updated 2 years ago
- LLM Program Watermarking☆19Apr 19, 2024Updated 2 years ago
- [Findings of EMNLP 2022] Code of paper Generative Prompt Tuning for Relation Classification. https://arxiv.org/abs/2210.12435☆20May 7, 2023Updated 3 years ago
- ☆19Sep 10, 2023Updated 2 years ago
- UP-TO-DATE LLM Watermark paper. 🔥🔥🔥☆379Dec 12, 2024Updated last year
- ☆17Aug 1, 2025Updated last year
- This code accompanies the paper DisentQA: Disentangling Parametric and Contextual Knowledge with Counterfactual Question Answering.☆16Mar 20, 2023Updated 3 years ago
- ☆17Aug 2, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official Repository for the CVPR 2020 paper "Universal Litmus Patterns: Revealing Backdoor Attacks in CNNs"☆45Oct 24, 2023Updated 2 years ago
- multi-bit language model watermarking (NAACL 24)☆22Sep 20, 2024Updated last year
- ☆17Jul 15, 2025Updated last year
- Repo for outstanding paper@ACL 2023 "Do PLMs Know and Understand Ontological Knowledge?"☆33Oct 16, 2023Updated 2 years ago
- ☆46Feb 27, 2026Updated 6 months ago
- Restore safety in fine-tuned language models through task arithmetic☆33Mar 28, 2024Updated 2 years ago
- [JMLR] MarkDiffusion: An Open-Source Toolkit for Generative Watermarking of Latent Diffusion Models☆331Jul 11, 2026Updated last month