Clue inspired puzzles for testing LLM deduction abilities
☆47Mar 19, 2026Updated 4 months ago
Alternatives and similar repositories for temporal-clue
Users that are interested in temporal-clue are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OpenPipe Reinforcement Learning Experiments☆34Mar 14, 2025Updated last year
- ☆13Mar 23, 2025Updated last year
- ☆25Dec 13, 2024Updated last year
- Various LLM Benchmarks☆26Feb 20, 2026Updated 5 months ago
- ☆17Nov 23, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- All-in-one environment to use Dria, the collective knowledge for AI.☆14Mar 15, 2024Updated 2 years ago
- Python SDK for FirstBatch: Real-time personalization using vectorDBs☆17Nov 26, 2023Updated 2 years ago
- slowly building a set of infinite riddle generators for data-hungry methods☆14Nov 15, 2022Updated 3 years ago
- ☆15Jun 12, 2024Updated 2 years ago
- Samples and demos with LangGraph and LangChain frameworks.☆16Aug 22, 2025Updated 11 months ago
- Lego for GRPO☆30May 27, 2025Updated last year
- ☆39Aug 1, 2025Updated 11 months ago
- Codebase from our first release.☆58Feb 17, 2026Updated 5 months ago
- An example implementation of RLHF (or, more accurately, RLAIF) built on MLX and HuggingFace.☆37Jun 21, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Crossword puzzles in your terminal.☆22Feb 4, 2026Updated 5 months ago
- Clean RL implementation using MLX☆34Mar 8, 2024Updated 2 years ago
- MetaLadder: Ascending Mathematical Solution Quality via Analogical-Problem Reasoning Transfer (EMNLP 2025)☆12Apr 18, 2025Updated last year
- ☆17Mar 28, 2025Updated last year
- Official code for "How2Everything: Mining the Web for How-To Procedures to Evaluate and Improve LLMs"☆24Feb 10, 2026Updated 5 months ago
- Agent Engineering course files☆72Jul 12, 2025Updated last year
- Visualize any repo or codebase into diagram or animation☆24Oct 14, 2024Updated last year
- ☆16Feb 22, 2026Updated 5 months ago
- This library supports evaluating disparities in generated image quality, diversity, and consistency between geographic regions.☆20Jun 3, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Skill to annotate and create ai judges from agent logs☆17Oct 28, 2025Updated 9 months ago
- Simple repository for training small reasoning models☆51Feb 17, 2026Updated 5 months ago
- run deepseek v3 on a single node. Drops unused experts from memory.☆16Jan 26, 2025Updated last year
- Library for creating card games in general.☆10Apr 3, 2026Updated 3 months ago
- Memory Agent monorepo☆89Oct 9, 2025Updated 9 months ago
- This repo consists of the code as discussed in the Medium blog.☆17Sep 10, 2023Updated 2 years ago
- The DPAB-α Benchmark☆32Jan 15, 2025Updated last year
- A reading list of relevant papers and projects on foundation model annotation☆29Feb 27, 2025Updated last year
- A Datasette instance for searching WebVid-10M☆15Sep 30, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 🅱️ 基于termui实现的命令行版b站 | 支持MAC&WIN #正在开发中☆15Mar 20, 2023Updated 3 years ago
- Pretraining codebase for Apertus models, based on Megatron-LM☆21Sep 25, 2025Updated 10 months ago
- ☆15Jan 27, 2025Updated last year
- Simple node proxy for llama-server that enables MCP use☆19May 10, 2025Updated last year
- Alice in Wonderland code base for experiments and raw experiments data☆129Feb 4, 2026Updated 5 months ago
- seq-JEPA: Autoregressive Predictive Learning of Invariant-Equivariant World Models☆16Jan 26, 2026Updated 6 months ago
- Benchmarking execution environments ability to prevent reward hacking in agent evals.☆15Jun 18, 2026Updated last month