An implementation of Tiny Recursive Models (TRM)
☆128Mar 30, 2026Updated 6 months ago
Alternatives and similar repositories for nano-trm
Users that are interested in nano-trm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MLX Implementation of Recursive Reasoning with Tiny Networks☆79Oct 11, 2025Updated 11 months ago
- Unofficial implementation of Tiny Recursive Model (TRM), improvement to HRM from Sapient AI, by Alexia Jolicoeur-Martineau☆192Dec 23, 2025Updated 9 months ago
- Generic building-block toolbox for training neural networks with adaptive and recursive execution. It provides reusable components to con…☆27Jun 29, 2026Updated 3 months ago
- ☆35Nov 11, 2025Updated 10 months ago
- Grokking on modular arithmetic in less than 150 epochs in MLX☆15Oct 24, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆6,556Apr 1, 2026Updated 6 months ago
- Explorations into the proposed SDFT, Self-Distillation Enables Continual Learning, from Shenfeld et al. of MIT☆33Feb 6, 2026Updated 8 months ago
- Universal Reasoning Model☆138Jan 15, 2026Updated 8 months ago
- Implementation of 2-simplicial attention proposed by Clift et al. (2019) and the recent attempt to make practical in Fast and Simplex, Ro…☆49Sep 2, 2025Updated last year
- Implementation of Fast Weight Attention☆35Sep 17, 2026Updated 3 weeks ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆17Apr 30, 2025Updated last year
- ☆23Jul 10, 2026Updated 3 months ago
- ☆82Sep 11, 2026Updated 3 weeks ago
- Official PyTorch implementation of "Latent Reasoning in TRMs is Secretly a Policy Improvement Operator" (ICML 2026)☆26May 29, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for Fast-weight Product Key Memory (FwPKM)☆26Mar 18, 2026Updated 6 months ago
- MLX implementation of Hierarchical Reasoning Model (HRM) - Adaptive computation for complex reasoning tasks☆29Aug 27, 2025Updated last year
- Classifier for CIFAR-10. Grayscaling, HOG, PCA, and RBF SVM. 62% test accuracy. Walkthrough on YouTube: https://youtu.be/gmTweV0eHhk☆14Nov 24, 2024Updated last year
- ☆13Jun 3, 2024Updated 2 years ago
- 🤖 Complete reproduction of 'AlphaGo Moment for Model Architecture Discovery' using MLX-LM instead of GPT-4. Autonomous neural architectu…☆30Jul 27, 2025Updated last year
- Model-Based RL Demo for Pendulum-v0☆12Jun 16, 2020Updated 6 years ago
- Implementation of the work Variational multiple shooting for Bayesian ODEs with Gaussian processes☆13Aug 5, 2022Updated 4 years ago
- [WWW 2026 Oral] MoE-CL:Self-Evolving LLMs via Continual Instruction Tuning☆22Dec 1, 2025Updated 10 months ago
- ☆71Dec 12, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation of Multiscreen proposed by Ken Nakanishi for "Screening is Enough"☆19May 13, 2026Updated 4 months ago
- ☆154Sep 29, 2025Updated last year
- 100M tokens. Infinite compute. Lowest val loss wins.☆560Sep 15, 2026Updated 3 weeks ago
- Source code for the paper "Positional Attention: Expressivity and Learnability of Algorithmic Computation"☆14May 26, 2025Updated last year
- A collection of optimizers for MLX☆59Dec 12, 2025Updated 9 months ago
- ☆51Jul 3, 2026Updated 3 months ago
- Implementation of the proposed Adam-atan2 from Google Deepmind in Pytorch☆143Sep 3, 2026Updated last month
- Language modeling with linear-cost context☆122Sep 25, 2025Updated last year
- supporting pytorch FSDP for optimizers☆85Dec 8, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of a holodeck, written in Pytorch☆19Nov 1, 2023Updated 2 years ago
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated last year
- [EMNLP 2025🔥] UNComp: Can Matrix Entropy Uncover Sparsity? -- A Compressor Design from an Uncertainty-Aware Perspective☆20Jan 7, 2026Updated 9 months ago
- Synthetic Alphabet Dataset☆19Mar 27, 2025Updated last year
- RAG application to answer questions about PDF documents using LLMs.☆16Dec 1, 2023Updated 2 years ago
- An annotated implementation of the Hyena Hierarchy paper☆34May 28, 2023Updated 3 years ago
- train entropix like a champ!☆20Oct 10, 2024Updated last year