A certifier for bias in LLMs
☆25Apr 11, 2025Updated last year
Alternatives and similar repositories for LLMCert-B
Users that are interested in LLMCert-B are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Benchmarking Open-Ended Inference Optimization by AI Agents☆35Jul 6, 2026Updated last month
- Corresponding code to "Improving Robustness of ML Classifiers against Realizable Evasion Attacks Using Conserved Features" @ USENIX Secur…☆11Aug 5, 2019Updated 7 years ago
- ☆12Apr 25, 2025Updated last year
- ☆14Oct 17, 2024Updated last year
- ☆13Oct 14, 2020Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A very limited implementation of arXiv:1904.00759☆13Dec 2, 2019Updated 6 years ago
- Code for the paper "Evading Black-box Classifiers Without Breaking Eggs" [SaTML 2024]☆21Apr 15, 2024Updated 2 years ago
- ☆19Jan 22, 2026Updated 6 months ago
- Private Adaptive Optimization with Side Information (ICML '22)☆16Jun 23, 2022Updated 4 years ago
- ☆15Jun 25, 2025Updated last year
- ☆21Mar 19, 2023Updated 3 years ago
- Awesome Agent Environments☆17Apr 10, 2026Updated 4 months ago
- Official codebase for permutation self-consistency.☆19Feb 11, 2024Updated 2 years ago
- ☆17Aug 2, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- UnQovering Stereotyping Biases via Underspecified Questions - EMNLP 2020 (Findings)☆20Jul 6, 2021Updated 5 years ago
- Test equality between a black-box LLM API and a reference distribution☆20Oct 29, 2024Updated last year
- A new algorithm that formulates jailbreaking as a reasoning problem.☆26Jul 2, 2025Updated last year
- Code for NeurIPS 2024 Paper - Superposed Decoding: Multiple Generations from a Single Autoregressive Inference Pass☆21Aug 22, 2024Updated last year
- Code to generate NeuralExecs (prompt injection for LLMs)☆27Oct 5, 2025Updated 10 months ago
- ☆21Oct 23, 2024Updated last year
- Application of CollaGAN (Collaborative GAN) for MRI Image Imputation☆28Dec 8, 2019Updated 6 years ago
- A Recipe for Building LLM Reasoners to Solve Complex Instructions☆32Oct 9, 2025Updated 10 months ago
- Fast Memorization of Prompt Improves Context Awareness of Large Language Models (Findings of EMNLP 2024)☆22Oct 22, 2024Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Fine-tuning base models to build robust task-specific models☆36Apr 11, 2024Updated 2 years ago
- ☆38Feb 12, 2025Updated last year
- Library for training globally-robust neural networks.☆31Aug 7, 2025Updated last year
- ☆11Jan 3, 2024Updated 2 years ago
- Code for the paper "Spectral Editing of Activations for Large Language Model Alignments"☆31Dec 20, 2024Updated last year
- LongProc: Benchmarking Long-Context Language Models on Long Procedural Generation☆36Feb 26, 2026Updated 5 months ago
- This repository contains a simple implementation of Interval Bound Propagation (IBP) using TensorFlow: https://arxiv.org/abs/1810.12715☆163Dec 20, 2019Updated 6 years ago
- Accompanying code for "Boosted Prompt Ensembles for Large Language Models"☆31Apr 13, 2023Updated 3 years ago
- EMNLP 2022: "MABEL: Attenuating Gender Bias using Textual Entailment Data" https://arxiv.org/abs/2210.14975☆38Dec 14, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code and data to go with the Zhu et al. paper "An Objective for Nuanced LLM Jailbreaks"☆37Jul 2, 2026Updated last month
- An open-source non-official community implementation of the model from the paper: Surgical Robot Transformer (SRT): Imitation Learning fo…☆13Aug 3, 2026Updated last week
- Learning Security Classifiers with Verified Global Robustness Properties (CCS'21) https://arxiv.org/pdf/2105.11363.pdf☆28Dec 1, 2021Updated 4 years ago
- Interpretating the latent space representations of attention head outputs for LLMs☆39Aug 13, 2024Updated last year
- Multilingual Pre-training with Language and Task Adaptation for Multilingual Text Style Transfer (ACL 2022)☆10Sep 22, 2022Updated 3 years ago
- ☆35Nov 12, 2024Updated last year
- ☆38Jan 17, 2025Updated last year