Implement of Implicit Knowledge Extraction Attack.
☆24Jul 14, 2026Updated 2 weeks ago
Alternatives and similar repositories for IKEA
Users that are interested in IKEA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 本仓库整理公开披露的 OpenClaw 恶意 / 可疑 Skills 样本,并对其进行分类、分级与风险说明,供安全研究、平台治理和风险检测使用。☆20Mar 16, 2026Updated 4 months ago
- [NeurIPS 2024] "Membership Inference on Text-to-image Diffusion Models via Conditional Likelihood Discrepancy"☆12Sep 15, 2025Updated 10 months ago
- BrainWash: A Poisoning Attack to Forget in Continual Learning☆13Apr 15, 2024Updated 2 years ago
- [ACL 2025] Beyond Prompt Engineering: Robust Behavior Control in LLMs via Steering Target Atoms☆41Jun 4, 2025Updated last year
- [MM'23 Oral] "Text-to-image diffusion models can be easily backdoored through multimodal data poisoning"☆33Jun 10, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official Implementation of NIPS 2022 paper Pre-activation Distributions Expose Backdoor Neurons☆15Jan 13, 2023Updated 3 years ago
- Code and full version of the paper "Hijacking Attacks against Neural Network by Analyzing Training Data"☆14Feb 28, 2024Updated 2 years ago
- BackdoorAgent is a stage-aware framework and benchmark that instruments LLM-agent workflows (planning, memory, tools) to systematically i…☆43Mar 16, 2026Updated 4 months ago
- ☆17Mar 9, 2025Updated last year
- [CVPR 2024] Not All Prompts Are Secure: A Switchable Backdoor Attack Against Pre-trained Vision Transfomers☆16Oct 24, 2024Updated last year
- ☆22Dec 14, 2023Updated 2 years ago
- SkillJect: Automating Stealthy Skill-Based Prompt Injection for Coding Agents with Trace-Driven Closed-Loop Refinement☆73Jun 11, 2026Updated last month
- [ICLR'26 Oral] LLM DNA: Tracing Model Evolution via Functional Representations☆32Apr 8, 2026Updated 3 months ago
- Source code for the ACL'2025 paper titled "Unveiling privacy risks in llm agent memory"☆34Dec 2, 2025Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official code for our paper "SoK: Large Language Model Copyright Auditing via Fingerprinting"☆18Dec 31, 2025Updated 6 months ago
- [NeurIPS 2025] Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack☆15Nov 19, 2025Updated 8 months ago
- ☆23Apr 23, 2024Updated 2 years ago
- [S&P'24] Test-Time Poisoning Attacks Against Test-Time Adaptation Models☆21Feb 18, 2025Updated last year
- Proof-of-concept implementation for the paper "ThermalScope: A Practical Interrupt Side Channel Attack Based On Thermal Event Interrupts"…☆13Dec 17, 2024Updated last year
- [NDSS 2026] Official repo for Odysseus: Jailbreaking Commercial Multimodal LLM-integrated Systems via Dual Steganography☆59Mar 14, 2026Updated 4 months ago
- 🔥🔥🔥 Detecting hidden backdoors in Large Language Models with only black-box access☆57Jun 2, 2025Updated last year
- ☆33Feb 27, 2025Updated last year
- PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses☆22Jul 17, 2026Updated last week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code and data for the ACM CIKM 2024 paper "Adversarial Text Rewriting for Text-aware Recommender Systems"☆12Aug 1, 2024Updated last year
- ☆12Nov 2, 2024Updated last year
- PFI: Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agents☆31Mar 26, 2025Updated last year
- Proof-of-concept implementation for the paper "SegScope: Probing Fine-grained Interrupts via Architectural Footprints" (HPCA'24)☆20Apr 2, 2026Updated 3 months ago
- ☆32Mar 16, 2025Updated last year
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 3 months ago
- Official Repository for "Ten Words Only Still Help: Improving Black-Box AI-Generated Text Detection via Proxy-Guided Efficient Re-Samplin…☆23Aug 15, 2024Updated last year
- Adversarial Item Promotion in visually-aware recommenders☆17Sep 3, 2021Updated 4 years ago
- Official implementation of the USENIX Security 2024 paper ModelGuard: Information-Theoretic Defense Against Model Extraction Attacks.☆25Dec 6, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [MM '24] EvilEdit: Backdooring Text-to-Image Diffusion Models in One Second☆28Nov 19, 2024Updated last year
- ☆12Mar 5, 2024Updated 2 years ago
- Placeholder repository☆15Mar 16, 2022Updated 4 years ago
- This repository is the official implementation of StealthDiffusion: Towards Evading Diffusion Forensic Detection through Diffusion Model☆21Jul 30, 2024Updated last year
- [EMNLP 24] Official Implementation of CLEANGEN: Mitigating Backdoor Attacks for Generation Tasks in Large Language Models☆19Mar 9, 2025Updated last year
- ☆29Jun 13, 2026Updated last month
- 无线感知相关论文☆11Oct 26, 2021Updated 4 years ago