Repository for the Bias Benchmark for QA dataset.
☆152Jan 8, 2024Updated 2 years ago
Alternatives and similar repositories for BBQ
Users that are interested in BBQ are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Dataset associated with "BOLD: Dataset and Metrics for Measuring Biases in Open-Ended Language Generation" paper☆90Mar 2, 2021Updated 5 years ago
- This repository contains the data and code introduced in the paper "CrowS-Pairs: A Challenge Dataset for Measuring Social Biases in Maske…☆140Mar 1, 2024Updated 2 years ago
- ☆163Sep 12, 2023Updated 3 years ago
- ACL 2022: An Empirical Survey of the Effectiveness of Debiasing Techniques for Pre-trained Language Models.☆156Aug 18, 2025Updated last year
- StereoSet: Measuring stereotypical bias in pretrained language models☆206Dec 8, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official Implementation of "Learning to Refuse: Towards Mitigating Privacy Risks in LLMs"☆10Dec 13, 2024Updated last year
- ☆15Jun 25, 2025Updated last year
- Code and test data for "On Measuring Bias in Sentence Encoders", to appear at NAACL 2019.☆58May 23, 2021Updated 5 years ago
- Repository for research in the field of Responsible NLP at Meta.☆213Apr 18, 2026Updated 5 months ago
- The Codebase for Causal Distillation for Language Models (NAACL '22)☆26May 1, 2022Updated 4 years ago
- ☆31Aug 9, 2023Updated 3 years ago
- [NeurIPS 2024 D&B] Evaluating Copyright Takedown Methods for Language Models☆17Jul 17, 2024Updated 2 years ago
- ☆34Aug 9, 2024Updated 2 years ago
- Butler 是一个用于自动化服务管理和任务调度的工具项目。☆18Oct 1, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Framework for controlling demographic biases in NLG (using adversarial prompts)☆21Jun 12, 2023Updated 3 years ago
- ☆235Feb 23, 2021Updated 5 years ago
- ☆20Jun 21, 2025Updated last year
- Official repo for EMNLP'24 paper "SOUL: Unlocking the Power of Second-Order Optimization for LLM Unlearning"☆31Oct 1, 2024Updated 2 years ago
- ☆29Oct 6, 2024Updated 2 years ago
- ☆45Mar 3, 2023Updated 3 years ago
- This repository contains the code for "Self-Diagnosis and Self-Debiasing: A Proposal for Reducing Corpus-Based Bias in NLP".☆89Aug 20, 2021Updated 5 years ago
- A list of ethics related resources for researchers and practitioners of Natural Language Processing and Computational Linguistics☆34Oct 20, 2025Updated 11 months ago
- Data for evaluating gender bias in coreference resolution systems.☆83May 14, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Source code for the paper "Automatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled Data"☆20Feb 24, 2024Updated 2 years ago
- Official repo for NeurIPS'24 paper "WAGLE: Strategic Weight Attribution for Effective and Modular Unlearning in Large Language Models"☆19Dec 16, 2024Updated last year
- Dataset + classifier tools to study social perception biases in natural language generation☆71Jun 12, 2023Updated 3 years ago
- Bias Benchmark for Natural Language Inference. Code repo for the Findings of NAACL 2022 paper "On Measuring Social Biases in Prompt-Based…☆14Apr 28, 2022Updated 4 years ago
- Sensitive-rs is a Rust library for finding, validating, filtering, and replacing sensitive words. It provides efficient algorithms to han…☆27Sep 29, 2026Updated last week
- ☆10Jul 6, 2023Updated 3 years ago
- 🌏 UI component library for the future, based on WebComponent.☆23Nov 12, 2024Updated last year
- The benchmark proposed in paper: GraphInstruct: Empowering Large Language Models with Graph Understanding and Reasoning Capability☆25Aug 12, 2025Updated last year
- ☆13Mar 7, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆55Apr 26, 2022Updated 4 years ago
- WMDP is a LLM proxy benchmark for hazardous knowledge in bio, cyber, and chemical security. We also release code for RMU, an unlearning m…☆186May 29, 2025Updated last year
- Augmenting Statistical Models with Natural Language Parameters☆28Sep 17, 2024Updated 2 years ago
- To analyze and remove gender bias in coreference resolution systems☆78May 6, 2025Updated last year
- ☆58Jun 30, 2023Updated 3 years ago
- Code and data of the EMNLP 2022 paper "Why Should Adversarial Perturbations be Imperceptible? Rethink the Research Paradigm in Adversaria…☆86Feb 19, 2023Updated 3 years ago
- ☆42Oct 29, 2024Updated last year