☆102Jun 10, 2026Updated last month
Alternatives and similar repositories for SnakeBench
Users that are interested in SnakeBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ARLC, a probabilistic abductive reasoner for solving Raven's progressive matrices.☆25Sep 18, 2025Updated 10 months ago
- ☆57Nov 22, 2024Updated last year
- Repo for solving arc problems with an Neural Cellular Automata☆27Mar 9, 2026Updated 4 months ago
- Kaggle AIMO2 solution with token-efficient reasoning LLM recipes☆50Aug 7, 2025Updated 11 months ago
- ☆226Jan 5, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Abstraction and Reasoning Corpus☆15Nov 22, 2022Updated 3 years ago
- ☆15Jun 13, 2025Updated last year
- ☆27Aug 16, 2025Updated 11 months ago
- Unofficial Implementation of Selective Attention Transformer☆20Oct 31, 2024Updated last year
- MMLU-Pro eval results☆15Aug 21, 2025Updated 11 months ago
- ☆18Nov 30, 2025Updated 7 months ago
- Awesome-RL-Reasoning☆17Updated this week
- Exploring Applications of GRPO☆252Aug 25, 2025Updated 11 months ago
- ☆15Jun 17, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Simple repository for training small reasoning models☆51Feb 17, 2026Updated 5 months ago
- A framework for pitting LLMs against each other in an evolving library of games ⚔☆35Apr 17, 2025Updated last year
- ☆16Jun 15, 2026Updated last month
- ☆132Jun 1, 2026Updated last month
- Official implementation of Paper "System-Aware 4-Bit KV-Cache Quantization for Real-World LLM Serving"☆30Apr 17, 2026Updated 3 months ago
- Leverage the power of the Google Natural Language API NLP to retrieve entity relationships from Wikipedia URLs or topics! Get interactive…☆16Jun 23, 2021Updated 5 years ago
- Model Context Protocol (MCP) server to capture images from an OpenCV-compatible webcam or video source☆17Mar 28, 2025Updated last year
- A data processing module implemented with numpy☆10Aug 16, 2022Updated 3 years ago
- AI for Mathematics Paper List☆17Jan 14, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SutroYaro — Sutro Group research workspace for energy-efficient AI training. Point any coding agent at the repo and it becomes a research…☆15May 29, 2026Updated last month
- Give langchain access to the terminal☆33Apr 10, 2023Updated 3 years ago
- TUI conversation explorer for Claude Code & OpenCode☆20Aug 21, 2025Updated 11 months ago
- Lighter than feather, Harder than steel, RL Framework☆24Mar 26, 2026Updated 3 months ago
- Domain Specific Language for the Abstraction and Reasoning Corpus☆343Oct 11, 2024Updated last year
- a Python library that uses Reinforcement Learning (RL) to train LLMs.☆43Jul 12, 2026Updated last week
- Rethinking the Trust Region in LLM Reinforcement Learning☆62Mar 2, 2026Updated 4 months ago
- ☆11Jul 21, 2024Updated 2 years ago
- Code for 1st place solution to Kaggle's Abstraction and Reasoning Challenge☆164Jul 10, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Reverse Engineering the Abstraction and Reasoning Corpus☆355Feb 24, 2025Updated last year
- Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), p…☆16Jun 6, 2026Updated last month
- A Mimetic Procedural Benchmark Generator for the Abstraction and Reasoning Corpus☆51Apr 18, 2026Updated 3 months ago
- A simple MLX implementation for pretraining LLMs on Apple Silicon.☆85Aug 20, 2025Updated 11 months ago
- ☆59Jan 28, 2025Updated last year
- Evals meant to evaluate language models' ability to reason over long contexts.☆10Sep 12, 2024Updated last year
- ☆287May 28, 2026Updated last month