☆180Aug 15, 2025Updated last year
Alternatives and similar repositories for hierarchical-reasoning-model-analysis
Users that are interested in hierarchical-reasoning-model-analysis are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆6,567Apr 1, 2026Updated 5 months ago
- Universal Reasoning Model☆134Jan 15, 2026Updated 7 months ago
- [Tech Report] Expanded Hyper-Connections☆62Jul 21, 2026Updated last month
- Our solution for the arc challenge 2024☆188Jun 17, 2025Updated last year
- Core Library of Discrete Distribution Networks (ICLR 2025)☆16Oct 12, 2025Updated 10 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [preprint] Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning☆19Feb 18, 2026Updated 6 months ago
- ☆15Jun 19, 2025Updated last year
- ☆228Jan 5, 2026Updated 7 months ago
- ☆37Aug 7, 2025Updated last year
- Simple Python Game Engine☆29Jan 29, 2026Updated 7 months ago
- The official github repo for "Diffusion Language Models are Super Data Learners".☆228Nov 6, 2025Updated 9 months ago
- ☆21Feb 22, 2025Updated last year
- [ICML 2026] An Evaluation Suite for Chain-of-Thought Controllability☆53Mar 10, 2026Updated 5 months ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆17Apr 30, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Mixture-of-Recursions: Learning Dynamic Recursive Depths for Adaptive Token-Level Computation (NeurIPS 2025)☆581Sep 26, 2025Updated 11 months ago
- [ICML 2026] Code for Equilibrium Reasoners: learning attractor dynamics for scalable reasoning☆46Aug 4, 2026Updated 3 weeks ago
- Code for the Fractured Entangled Representation Hypothesis position paper!☆228Nov 6, 2025Updated 9 months ago
- Implementation of the proposed Adam-atan2 from Google Deepmind in Pytorch☆143Jul 17, 2026Updated last month
- Minimal open-source implementation of AlphaProof and HyperTree Proof Search.☆88May 13, 2026Updated 3 months ago
- Implementation of SOAR☆56Sep 17, 2025Updated 11 months ago
- ☆25Dec 18, 2024Updated last year
- Bootstrapping ARC☆165Nov 20, 2024Updated last year
- ☆23Apr 4, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Custom triton kernels for training Karpathy's nanoGPT.☆19Oct 21, 2024Updated last year
- Measuring the Signal to Noise Ratio in Language Model Evaluation☆31Aug 19, 2025Updated last year
- A Python implementation of an agent swarm system that works with local LLM servers. The system allows you to create multiple agents that …☆14Nov 20, 2024Updated last year
- Landing repository for the paper "Predicting the Order of Upcoming Tokens Improves Language Modeling"☆48May 13, 2026Updated 3 months ago
- Minimal and highly hackable implementation of Looped Transformers with GPT☆25Mar 8, 2026Updated 5 months ago
- Physics of Language Models: Part 4.2, Canon Layers at Scale where Synthetic Pretraining Resonates in Reality☆360Aug 21, 2026Updated last week
- An implementation is provided here for the NeurIPS2024 paper "MemoryFormer : Minimize Transformer Computation by Removing Fully-Connected…☆16Mar 24, 2026Updated 5 months ago
- Pretraining and inference code for a large-scale depth-recurrent language model☆911Dec 29, 2025Updated 8 months ago
- Generalized Optimal Transport Attention with Trainable Priors☆72Aug 23, 2026Updated last week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- minimal Energy-based transformer☆44Dec 11, 2025Updated 8 months ago
- Official Implementation for Inference-time Scaling of Diffusion Models through Classical Search☆37Oct 8, 2025Updated 10 months ago
- [ICML 2025] Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction☆91May 26, 2025Updated last year
- rl from zero pretrain, can it be done? yes.☆296Sep 28, 2025Updated 11 months ago
- DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation☆306Feb 18, 2026Updated 6 months ago
- VC-FB and MC-FB algorithms from "Zero-Shot Reinforcement Learning from Low Quality Data" (NeurIPS 2024)☆29Jan 14, 2025Updated last year
- ☆34Mar 3, 2025Updated last year