Code for "TrustRAG: Enhancing Robustness and Trustworthiness in RAG" AAAI 2026 Workshop on Trust and Control in Agentic AI (TrustAgent)
☆61Mar 24, 2025Updated last year
Alternatives and similar repositories for TrustRAG
Users that are interested in TrustRAG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆32Sep 15, 2024Updated last year
- [USENIX Security 2025] PoisonedRAG: Knowledge Corruption Attacks to Retrieval-Augmented Generation of Large Language Models☆298Jan 27, 2026Updated 7 months ago
- ☆64Mar 11, 2025Updated last year
- Official PyTorch implementation of "MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks"☆18Dec 4, 2025Updated 8 months ago
- ☆48Dec 9, 2025Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official repository for "Conflict-Aware Soft Prompting for Retrieval-Augmented Generation" (EMNLP 2025)☆20Nov 13, 2025Updated 9 months ago
- ☆114Jun 1, 2026Updated 2 months ago
- [ACL'26 Main Oral] Official code for "LogicPoison: Logical Attacks on Graph Retrieval-Augmented Generation".☆37Updated this week
- [EMNLP 2023] Poisoning Retrieval Corpora by Injecting Adversarial Passages https://arxiv.org/abs/2310.19156☆51Dec 14, 2023Updated 2 years ago
- This is the code repo for the paper "RAG-DDR: Optimizing Retrieval-Augmented Generation Using Differentiable Data Rewards".☆23Oct 28, 2024Updated last year
- ☆19Nov 3, 2025Updated 9 months ago
- [NeurIPS 2024] Official implementation for "AgentPoison: Red-teaming LLM Agents via Memory or Knowledge Base Backdoor Poisoning"☆240Aug 16, 2026Updated 2 weeks ago
- Official Code for "Baseline Defenses for Adversarial Attacks Against Aligned Language Models"☆34Oct 26, 2023Updated 2 years ago
- ☆16May 18, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [CCS 2026] The official implementation of our CCS 2026 paper "ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathological…☆17Aug 5, 2026Updated 3 weeks ago
- ☆25Dec 18, 2025Updated 8 months ago
- A Benchmark Study on Machine Learning Methods for Fake News Detection☆16Jun 8, 2021Updated 5 years ago
- ☆15Jan 24, 2025Updated last year
- Repo for paper "CODIS: Benchmarking Context-Dependent Visual Comprehension for Multimodal Large Language Models".☆13Oct 14, 2024Updated last year
- HackerGPT Lite Docs☆20Jun 18, 2026Updated 2 months ago
- BackdoorAgent is a stage-aware framework and benchmark that instruments LLM-agent workflows (planning, memory, tools) to systematically i…☆46Mar 16, 2026Updated 5 months ago
- [ACL 2025] Removal of Hallucination on Hallucination: Debate-Augmented RAG☆43Aug 4, 2025Updated last year
- ☆32Jun 13, 2026Updated 2 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ACL 2024] PyTorch implementation for "Stealthy Attack on Large Language Model based Recommendation"☆22Jun 19, 2024Updated 2 years ago
- Paper Reading list for Shilling Attack and Defense on Recommender Systems☆24Feb 26, 2025Updated last year
- ☆56Jun 23, 2026Updated 2 months ago
- ☆15Jun 28, 2025Updated last year
- Code and Data for ACL 2025 Paper "Aristotle: Mastering Logical Reasoning with A Logic-Complete Decompose-Search-Resolve Framework".☆28Oct 3, 2025Updated 10 months ago
- Exchange-of-Thought: Enhancing Large Language Model Capabilities through Cross-Model Communication☆21Mar 21, 2024Updated 2 years ago
- Code and data for the ACM CIKM 2024 paper "Adversarial Text Rewriting for Text-aware Recommender Systems"☆13Aug 1, 2024Updated 2 years ago
- A Framework for Evaluating AI Agent Safety in Realistic Environments☆38Aug 10, 2026Updated 3 weeks ago
- (WWW'25 + Netflix) The first CRS that retrieves collaborative filtering knowledge with two-step context-aware reflection.☆21Sep 10, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Enabling Ultra-Fast Cardiovascular Imaging Across Heterogeneous Clinical Environments with A Generalist Foundation Model and Multimodal D…☆15Jun 20, 2026Updated 2 months ago
- ☆42Mar 24, 2026Updated 5 months ago
- Code for reproducing our paper "Low Rank Adapting Models for Sparse Autoencoder Features"☆17Mar 31, 2025Updated last year
- [ICML 2026] Official implementation for paper "Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Ag…☆35Jul 31, 2026Updated last month
- official repository for paper:First Multi-Dimensional Evaluation of Flowchart Comprehension for Multimodal Large Language Models☆14Jul 29, 2024Updated 2 years ago
- We propose a novel adversarial example generation technique (i.e., CODA) for testing deep code models. Its key idea is to use code differ…☆20Oct 8, 2023Updated 2 years ago
- ☆67May 21, 2025Updated last year