☆179Aug 15, 2025Updated 11 months ago
Alternatives and similar repositories for hierarchical-reasoning-model-analysis
Users that are interested in hierarchical-reasoning-model-analysis are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Hierarchical Reasoning Model Official Release☆12,598Mar 31, 2026Updated 3 months ago
- ☆6,573Apr 1, 2026Updated 3 months ago
- Universal Reasoning Model☆134Jan 15, 2026Updated 6 months ago
- [Tech Report] Expanded Hyper-Connections☆42Updated this week
- Our solution for the arc challenge 2024☆189Jun 17, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Core Library of Discrete Distribution Networks (ICLR 2025)☆15Oct 12, 2025Updated 9 months ago
- [preprint] Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning☆19Feb 18, 2026Updated 5 months ago
- ☆15Jun 19, 2025Updated last year
- ☆226Jan 5, 2026Updated 6 months ago
- ☆37Aug 7, 2025Updated 11 months ago
- The official github repo for "Diffusion Language Models are Super Data Learners".☆227Nov 6, 2025Updated 8 months ago
- ☆21Feb 22, 2025Updated last year
- [ICML 2026] An Evaluation Suite for Chain-of-Thought Controllability☆50Mar 10, 2026Updated 4 months ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆16Apr 30, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Mixture-of-Recursions: Learning Dynamic Recursive Depths for Adaptive Token-Level Computation (NeurIPS 2025)☆577Sep 26, 2025Updated 9 months ago
- [ICML 2026] Code for Equilibrium Reasoners: learning attractor dynamics for scalable reasoning☆44Jun 1, 2026Updated last month
- Code for the Fractured Entangled Representation Hypothesis position paper!☆227Nov 6, 2025Updated 8 months ago
- Minimal open-source implementation of AlphaProof and HyperTree Proof Search.☆87May 13, 2026Updated 2 months ago
- Implementation of SOAR☆55Sep 17, 2025Updated 10 months ago
- Bootstrapping ARC☆162Nov 20, 2024Updated last year
- ☆23Apr 4, 2024Updated 2 years ago
- Custom triton kernels for training Karpathy's nanoGPT.☆19Oct 21, 2024Updated last year
- Measuring the Signal to Noise Ratio in Language Model Evaluation☆31Aug 19, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Landing repository for the paper "Predicting the Order of Upcoming Tokens Improves Language Modeling"☆48May 13, 2026Updated 2 months ago
- A Python implementation of an agent swarm system that works with local LLM servers. The system allows you to create multiple agents that …☆14Nov 20, 2024Updated last year
- Minimal and highly hackable implementation of Looped Transformers with GPT☆25Mar 8, 2026Updated 4 months ago
- Physics of Language Models: Part 4.2, Canon Layers at Scale where Synthetic Pretraining Resonates in Reality☆356May 20, 2026Updated 2 months ago
- Unofficial implementation of Tiny Recursive Model (TRM), improvement to HRM from Sapient AI, by Alexia Jolicoeur-Martineau☆191Dec 23, 2025Updated 6 months ago
- An implementation is provided here for the NeurIPS2024 paper "MemoryFormer : Minimize Transformer Computation by Removing Fully-Connected…☆16Mar 24, 2026Updated 3 months ago
- Pretraining and inference code for a large-scale depth-recurrent language model☆900Dec 29, 2025Updated 6 months ago
- minimal Energy-based transformer☆44Dec 11, 2025Updated 7 months ago
- Generalized Optimal Transport Attention with Trainable Priors☆70Jan 25, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence☆68Nov 11, 2025Updated 8 months ago
- [ICML 2025] Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction☆89May 26, 2025Updated last year
- Official Implementation for Inference-time Scaling of Diffusion Models through Classical Search☆33Oct 8, 2025Updated 9 months ago
- A memory efficient implementation of custom SWISH and MISH activation functions in Pytorch☆12Jun 29, 2020Updated 6 years ago
- rl from zero pretrain, can it be done? yes.☆295Sep 28, 2025Updated 9 months ago
- DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation☆240Feb 18, 2026Updated 5 months ago
- Official Implementation of `An Optimisation Framework for Unsupervised Environment Design` from RLC 2025☆17Nov 24, 2025Updated 7 months ago