☆31Aug 9, 2023Updated 3 years ago
Alternatives and similar repositories for CBBQ
Users that are interested in CBBQ are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21Mar 17, 2025Updated last year
- ☆11Oct 12, 2023Updated 2 years ago
- A new release of Chinese sexism dataset and lexicon☆14May 23, 2023Updated 3 years ago
- Flames is a highly adversarial benchmark in Chinese for LLM's harmlessness evaluation developed by Shanghai AI Lab and Fudan NLP Group.☆68May 21, 2024Updated 2 years ago
- Code for "Goodtriever: Toxicity Mitigation with Retrieval-augmented Language Models"☆25May 30, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Chinese corpus for gender bIas probing and mitigation, which contains 32.9k sentences with high-quality labels.☆21Aug 15, 2024Updated 2 years ago
- Code for EMNLP 2023 findings paper "A Closer Look into Using Large Language Models for Automatic Evaluation"☆19Oct 9, 2023Updated 2 years ago
- A Bilingual Role Evaluation Benchmark for Large Language Models☆43Jan 9, 2024Updated 2 years ago
- 🚲 Code and benchmark for our COLM 2025 paper - "Thought Tracing: Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models"☆15Aug 8, 2025Updated last year
- PRODIGy is a collection of dialogues in which each conversation is aligned with speaker profile representations.☆20Jan 8, 2025Updated last year
- ☆16Apr 23, 2025Updated last year
- ☆15Oct 24, 2022Updated 3 years ago
- This repository contains the code for the paper "Can Transformers Learn Full Bayesian Inference In Context?"☆16Apr 6, 2025Updated last year
- An R package implementing computational models of Eriksen flanker task performance.☆10Sep 19, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆17Nov 7, 2023Updated 2 years ago
- ☆11Apr 28, 2024Updated 2 years ago
- ☆15Sep 8, 2023Updated 2 years ago
- Chinese safety prompts for evaluating and improving the safety of LLMs. 中文安全prompts,用于评估和提升大模型的安全性。☆1,205Feb 27, 2024Updated 2 years ago
- ☆15Jun 8, 2024Updated 2 years ago
- รวบรวมทุกแอคเค้าท์ Twitter, Facebook และ Instagram ของสลิ่มที่น่าติดตาม คุยด้วยเหตุและผล☆12Nov 5, 2020Updated 5 years ago
- Membership inference against Federated learning.☆10May 30, 2021Updated 5 years ago
- ☆10Sep 17, 2022Updated 3 years ago
- A Docker workflow to work reproducibly with papaja in RStudio☆10Nov 16, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- มุขแป๊ก ๆ ที่เพื่อนคุณยังจะเอามาเล่น...☆13Sep 18, 2025Updated 11 months ago
- ☆11Apr 8, 2022Updated 4 years ago
- This repository is for the paper Incorporating External POS Tagger for Punctuation Restoration. Proc. Interspeech 2021, 1987-1991, doi: 1…☆11May 24, 2026Updated 2 months ago
- ☆20Oct 28, 2025Updated 9 months ago
- Search engine results page scraper☆13Dec 19, 2018Updated 7 years ago
- Fortifying Toxic Speech Detectors Against Veiled Toxicity☆11Oct 21, 2020Updated 5 years ago
- ☆25Aug 21, 2024Updated last year
- อาจารย์จ้องจะช่วยคุณ☆11Oct 31, 2022Updated 3 years ago
- MetricEval: A framework that conceptualizes and operationalizes four main components of metric evaluation, in terms of reliability and va…☆12Nov 6, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Bias Benchmark for Natural Language Inference. Code repo for the Findings of NAACL 2022 paper "On Measuring Social Biases in Prompt-Based…☆15Apr 28, 2022Updated 4 years ago
- Official github repo for SafetyBench, a comprehensive benchmark to evaluate LLMs' safety. [ACL 2024]☆297Jul 28, 2025Updated last year
- 南开计算机学院本科生毕设模板 根据硕士/博士模板修改而来☆16Jun 11, 2021Updated 5 years ago
- R package on Nigeria and for Nigeria☆12Aug 7, 2026Updated last week
- ☆10Nov 6, 2021Updated 4 years ago
- The implementation for "Comprehensive Knowledge Distillation with Causal Intervention".☆15Mar 12, 2022Updated 4 years ago
- Tools and examples for fitting (Hierarchical) Drift Diffusion Models in R☆11Jul 11, 2023Updated 3 years ago