Benchmarking Dark Patterns in LLMs (ICLR 2025)
β18Mar 29, 2025Updated last year
Alternatives and similar repositories for DarkBench
Users that are interested in DarkBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pin files for contextual, codebase-level AI assistance.β16Jul 11, 2024Updated 2 years ago
- πππππππππ Reading everythingβ17Mar 11, 2026Updated 6 months ago
- Bayesian scaling laws for in-context learning.β16Mar 12, 2025Updated last year
- [ICML 2023] "NeRFool: Uncovering the Vulnerability of Generalizable Neural Radiance Fields against Adversarial Perturbations" by Yonggan β¦β19Mar 10, 2024Updated 2 years ago
- Methods 2: The General Linear Modelβ15May 5, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β22Sep 9, 2021Updated 5 years ago
- Code for experiments on self-prediction as a way to measure introspection in LLMsβ17Dec 10, 2024Updated last year
- β14Mar 31, 2024Updated 2 years ago
- AIR-Bench 2024 is a safety benchmark that aligns with emerging government regulations and company policiesβ31Aug 14, 2024Updated 2 years ago
- Multiplayer JS game platformβ16Oct 16, 2017Updated 8 years ago
- Can Large Language Models Solve Security Challenges? We test LLMs' ability to interact and break out of shell environments using the Overβ¦β13Aug 21, 2023Updated 3 years ago
- Example fNIRS BIDS datasetβ15Nov 4, 2022Updated 3 years ago
- π₯ A repository for collecting cyberdefense thoughts, books, and documents about AI cyberdefenseβ13Jul 2, 2023Updated 3 years ago
- β56Oct 23, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- β18Updated this week
- Chain-of-thought λ°©μμ νμ©νμ¬ llama2λ₯Ό fine-tuningβ10Nov 18, 2023Updated 2 years ago
- β19Feb 18, 2026Updated 7 months ago
- Blind Justice Code for the paper "Blind Justice: Fairness with Encrypted Sensitive Attributes", ICML 2018β14Mar 20, 2019Updated 7 years ago
- CR-LT KGQA Dataset Repositoryβ10Jun 1, 2025Updated last year
- tools and benchmarks for verified codingβ33Jun 5, 2026Updated 3 months ago
- Fine-tuning of transformers for Sentiment Analysisβ18May 25, 2021Updated 5 years ago
- An ultimate pdf file disintegration toolβ11Jun 12, 2020Updated 6 years ago
- A web service in PHP that "translates" HackNPlan webhook messages to Discord webhook messages.β16Feb 23, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- β15Aug 7, 2025Updated last year
- 3cb: Catastrophic Cyber Capabilities Benchmarking of Large Language Modelsβ17Oct 30, 2024Updated last year
- β51Sep 28, 2025Updated 11 months ago
- ARC gym: a data generation framework for the Abstraction & Reasoning Corpusβ25Mar 25, 2026Updated 5 months ago
- Decomposing and measuring evaluation awareness in existing benchmarks and our proposed EvalAwareBench.β20Jun 1, 2026Updated 3 months ago
- The Happy Faces Benchmarkβ15Jul 20, 2023Updated 3 years ago
- Tools for exploring Transformer neuron behaviour, including input pruning and diversification.β24Sep 28, 2023Updated 2 years ago
- open Source code for propensity evaluationβ20Apr 25, 2026Updated 4 months ago
- R-package: Methods for dividing data into groups. Create balanced partitions and cross-validation folds. Perform time series windowing anβ¦β25Dec 18, 2024Updated last year
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Rearrrange data by a set of methodsβ23Mar 6, 2025Updated last year
- Meetup theme for Slidevβ25Updated this week
- Applications for OpenCL testing on Toradex Apalis iMX6Qβ13Dec 2, 2022Updated 3 years ago
- β11Jul 3, 2023Updated 3 years ago
- (Model-written) LLM evals libraryβ19Dec 13, 2024Updated last year
- Benchmarks for the Evaluation of LLM Supervisionβ35Jan 19, 2026Updated 8 months ago
- β18Dec 11, 2024Updated last year