A certifier for bias in LLMs
☆25Apr 11, 2025Updated last year
Alternatives and similar repositories for LLMCert-B
Users that are interested in LLMCert-B are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Formalize "Logic Notes" by Lou van den Dries in Lean☆12May 31, 2025Updated last year
- ☆12Apr 25, 2025Updated last year
- ☆14Oct 17, 2024Updated last year
- ☆15Jul 24, 2022Updated 4 years ago
- ☆15Jun 6, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆13Oct 14, 2020Updated 5 years ago
- A very limited implementation of arXiv:1904.00759☆13Dec 2, 2019Updated 6 years ago
- Two-Level Collaborative Fuzzing for Python Runtimes☆18Nov 25, 2023Updated 2 years ago
- MAP: Low-compute Model Merging with Amortized Pareto Fronts via Quadratic Approximation☆18Sep 2, 2024Updated 2 years ago
- ☆20Jan 22, 2026Updated 8 months ago
- The code repository for "MINGLE: Mixture of Null-Space Gated Low-Rank Experts for Test-Time Continual Model Merging"(NeurIPS25) in PyTorc…☆15Jun 2, 2026Updated 4 months ago
- Private Adaptive Optimization with Side Information (ICML '22)☆16Jun 23, 2022Updated 4 years ago
- ☆21Mar 19, 2023Updated 3 years ago
- Awesome Agent Environments☆19Apr 10, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official codebase for permutation self-consistency.☆20Feb 11, 2024Updated 2 years ago
- ☆17Aug 2, 2023Updated 3 years ago
- [CVPR 2026]The official code and datasets for "UniVBench: Towards Unified Evaluation for Video Foundation Models"☆25May 27, 2026Updated 4 months ago
- Test equality between a black-box LLM API and a reference distribution☆23Oct 29, 2024Updated last year
- A new algorithm that formulates jailbreaking as a reasoning problem.☆26Jul 2, 2025Updated last year
- This repo contains the source code for reproducing the experimental results in semantic density paper (Neurips 2024)☆22Sep 28, 2025Updated last year
- TDD-Bench-Verified is a new benchmark for generating test cases for test-driven development (TDD)☆35Jul 21, 2026Updated 2 months ago
- Code for NeurIPS 2024 Paper - Superposed Decoding: Multiple Generations from a Single Autoregressive Inference Pass☆21Aug 22, 2024Updated 2 years ago
- Code to generate NeuralExecs (prompt injection for LLMs)☆27Oct 5, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆21Oct 23, 2024Updated last year
- Artifact repository for the paper "AlphaTrans: A Neuro-Symbolic Compositional Approach for Repository-Level Code Translation and Validati…☆38Dec 20, 2025Updated 9 months ago
- A Recipe for Building LLM Reasoners to Solve Complex Instructions☆32Oct 9, 2025Updated last year
- Fine-tuning base models to build robust task-specific models☆36Apr 11, 2024Updated 2 years ago
- Multi-Agent Verification: Scaling Test-Time Compute with Multiple Verifiers☆33Mar 1, 2025Updated last year
- ☆38Feb 12, 2025Updated last year
- Library for training globally-robust neural networks.☆31Aug 7, 2025Updated last year
- ☆11Jan 3, 2024Updated 2 years ago
- Code for the paper "Spectral Editing of Activations for Large Language Model Alignments"☆31Dec 20, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- LongProc: Benchmarking Long-Context Language Models on Long Procedural Generation☆36Feb 26, 2026Updated 7 months ago
- Accompanying code for "Boosted Prompt Ensembles for Large Language Models"☆31Apr 13, 2023Updated 3 years ago
- Dataset associated with "BOLD: Dataset and Metrics for Measuring Biases in Open-Ended Language Generation" paper☆90Mar 2, 2021Updated 5 years ago
- Code and data to go with the Zhu et al. paper "An Objective for Nuanced LLM Jailbreaks"☆37Jul 2, 2026Updated 3 months ago
- An open-source non-official community implementation of the model from the paper: Surgical Robot Transformer (SRT): Imitation Learning fo…☆13Updated this week
- Interpretating the latent space representations of attention head outputs for LLMs☆39Aug 13, 2024Updated 2 years ago
- Multilingual Pre-training with Language and Task Adaptation for Multilingual Text Style Transfer (ACL 2022)☆10Sep 22, 2022Updated 4 years ago