LLM Divergent Thinking Creativity Benchmark. LLMs generate 25 unique words that start with a given letter with no connections to each other or to 50 initial random words.
☆36Mar 20, 2025Updated last year
Alternatives and similar repositories for divergent
Users that are interested in divergent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Systemic, uninstructed collusion among frontier LLMs in a simulated bidding environment☆20Jul 15, 2025Updated last year
- A benchmark for testing whether LLM judges keep the same preference when two lightly edited versions of the same story are shown in oppos…☆17Jun 11, 2026Updated 3 months ago
- Benchmark evaluating LLMs on their ability to create and resist disinformation. Includes comprehensive testing across major models (Claud…☆35Mar 20, 2025Updated last year
- Adversarial multi-turn benchmark for LLM debate quality, using side-swapped matchups and multi-model judging to rank models by judged deb…☆37Sep 25, 2026Updated 2 weeks ago
- Thematic Generalization Benchmark: measures how effectively various LLMs can infer a narrow or specific "theme" (category/rule) from a sm…☆74Apr 16, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Multi-Agent Step Race Benchmark: Assessing LLM Collaboration and Deception Under Pressure. A multi-player “step-race” that challenges LLM…☆91Dec 9, 2025Updated 10 months ago
- LLM Persuasion Benchmark tests whether one language model can change another model’s stated position over the course of a multi-turn conv…☆34Mar 27, 2026Updated 6 months ago
- Documents the style side of the short-story Creative Writing LLM benchmark: we generated many short stories with a range of LLMs, then an…☆28Dec 18, 2025Updated 9 months ago
- Public Goods Game (PGG) Benchmark: Contribute & Punish is a multi-agent benchmark that tests cooperative and self-interested strategies a…☆42Apr 10, 2025Updated last year
- LLM benchmark and leaderboard for narrator-bias sycophancy, opposite-narrator contradictions, and judgment consistency.☆64Aug 6, 2026Updated 2 months ago
- Hallucinations (Confabulations) Document-Based Benchmark for RAG. Includes human-verified questions and answers.☆249Aug 7, 2025Updated last year
- This benchmark tests how well LLMs incorporate a set of 10 mandatory story elements (characters, objects, core concepts, attributes, moti…☆453Updated this week
- A multi-player tournament benchmark that tests LLMs in social reasoning, strategy, and deception. Players engage in public and private co…☆299Jan 7, 2026Updated 9 months ago
- Groq-powered MAD: The first work to explore Multi-Agent Debate with Large Language Models :D☆12Jul 5, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Benchmark that evaluates LLMs using 759 NYT Connections puzzles extended with extra trick words☆243Sep 22, 2026Updated 2 weeks ago
- Attend - to what matters.☆16Feb 22, 2025Updated last year
- An OBSOLETE yt-dlp extractor plugin that used to solve YouTube JS challenges using Deno☆36Nov 12, 2025Updated 10 months ago
- UnReal World self-sufficiency mod☆16Oct 15, 2021Updated 4 years ago
- Automated terminal emulator benchmarks☆23Aug 31, 2026Updated last month
- ☆24Jan 22, 2025Updated last year
- Yet another frontend for LLM, written using .NET and WinUI 3☆11Sep 14, 2025Updated last year
- Create text chunks which end at natural stopping points without using a tokenizer☆26Nov 26, 2025Updated 10 months ago
- Teaching a humanoid to walk(ish), then displaying in your browser (using tensorflow.js and reinforcement learning)☆10Sep 7, 2020Updated 6 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- CoMakery Ethereum helps you create and administer collaborative token based projects.☆12Dec 16, 2017Updated 8 years ago
- A conversational UI for chatbots using the llama.cpp server☆15May 26, 2025Updated last year
- Bonded Token Buy/Sell/Paramaterize Component for React☆11Nov 8, 2018Updated 7 years ago
- UI that handle ESOP contracts☆11Dec 7, 2022Updated 3 years ago
- A cli app for experimenting with kokoro voice creating and mixing using the available voices to interpolate new ones☆40Feb 5, 2025Updated last year
- Learning as you go☆14Oct 25, 2015Updated 10 years ago
- ☆13Sep 2, 2024Updated 2 years ago
- Yet Another (LLM) Web UI, made with Gemini☆12Dec 25, 2024Updated last year
- 💼 Browser extension - Update your bookmarks with site descriptions☆12Sep 4, 2017Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Vulnerable Windows Driver with exploits which were used for demonstration purposes on Hunting and exploiting bugs in kernel drivers prese…☆13Jan 29, 2013Updated 13 years ago
- very tunnel☆33Oct 3, 2026Updated last week
- ☆12Dec 26, 2018Updated 7 years ago
- Angle Project (https://code.google.com/p/angleproject/) with support for Windows Store Apps (WinRT)☆44May 2, 2014Updated 12 years ago
- Compiler for a C/C++/C#-like language targeting the kOS scripting language used in the Kerbal Operating System mod (https://ksp-kos.githu…☆19Apr 7, 2017Updated 9 years ago
- Our new NixOS infrastructure flake 👩🏻💻 (pull/push mirror with GitLab)☆25Updated this week
- A playable 3D voxel game built in Lean 4.☆23Jul 12, 2026Updated 2 months ago