A transformer that executes a one-instruction Turing-complete computer — two approaches: hand-coded weights (no training) and learned from data
☆41Mar 3, 2026Updated 5 months ago
Alternatives and similar repositories for subleq-transformer
Users that are interested in subleq-transformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of the transformer (TF) architecture suggested in a paper entitled "Looped Transformers as Programmable Computers…☆43Apr 8, 2023Updated 3 years ago
- Repository for "Training Language Models To Explain Their Own Computations"☆28Jul 7, 2026Updated 3 weeks ago
- Cited 83-model x 49-benchmark LLM evaluation matrix with 18 matrix completion methods☆40Feb 25, 2026Updated 5 months ago
- [ICLR 2026] Official code for BézierFlow: Learning Bézier Stochastic Interpolant Schedulers for Few-Step Generation☆23Apr 13, 2026Updated 3 months ago
- Technical Appendices☆17Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 6,080-param transformer achieving 100% accuracy on 10-digit addition. Trained from scratch in 10 minutes.☆22Feb 19, 2026Updated 5 months ago
- [ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.☆39Mar 5, 2026Updated 5 months ago
- Implementation and explorations into PopuLoRA, Co-Evolving LLM Populations for Reasoning Self-Play☆16Updated this week
- Implementation of SelfExtend from the paper "LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning" from Pytorch and Zeta☆13Nov 11, 2024Updated last year
- Code for Fast-weight Product Key Memory (FwPKM)☆20Mar 18, 2026Updated 4 months ago
- Official code, models, and dataset for "Evolution Fine-Tuning (EFT): Learning to Discover Across 371 Optimization Tasks"☆26Jun 30, 2026Updated last month
- Agent Digivolve Harness is built around a simple observation: for many agent workflows, the first draft is not the hard part. The hard pa…☆31Mar 28, 2026Updated 4 months ago
- ☆94Jun 8, 2026Updated last month
- ☆21Aug 26, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Minimal and scalable research codebase in JAX, designed for rapid iteration on frontier research in LLM and other autoregressive models.☆552Jul 24, 2026Updated last week
- 100M tokens. Infinite compute. Lowest val loss wins.☆521Jul 3, 2026Updated last month
- DOrder -- Automatically Learning Shape Specifications☆20Jun 19, 2017Updated 9 years ago
- Implementation and datasets for "Training Language Models to Generate Quality Code with Program Analysis Feedback"☆42Jul 21, 2025Updated last year
- Code repository for the ICML 2026 Oral paper "Characterizing, Evaluating, and Optimizing Complex Reasoning".☆17Jun 21, 2026Updated last month
- Code to generate figures of paper "When do spectral gradient updates help in deep learning?"☆16Dec 3, 2025Updated 8 months ago
- FrontierSmith, a new system that uses AI to synthesize open-ended coding problems at scale☆50May 30, 2026Updated 2 months ago
- Numerical Optimisation Library☆17Jul 9, 2023Updated 3 years ago
- Kolu is an identification friend or foe (IFF) system for drones - European Defense Tech Hackathon 2025 LDN☆13Sep 28, 2025Updated 10 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The official implemention of "Depth-Breadth Synergy in RLVR: Unlocking LLM Reasoning Gains with Adaptive Exploration" (ICML 2026)☆24Feb 4, 2026Updated 6 months ago
- Agent History Protocol — tamper-evident recording for AI agents☆24Apr 16, 2026Updated 3 months ago
- Video Diffusion Model. Autoregressive, long context, efficient training and inference. WIP☆36Feb 17, 2026Updated 5 months ago
- The official repo for “Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem” [EMNLP25]☆33Sep 1, 2025Updated 11 months ago
- MLX implementation of Hierarchical Reasoning Model (HRM) - Adaptive computation for complex reasoning tasks☆29Aug 27, 2025Updated 11 months ago
- ⚙️ A library for proving PLONKish circuits (halo2) in the EVM.☆11Aug 14, 2023Updated 2 years ago
- A simple and efficient wrapper around the OpenAI API☆29Aug 5, 2024Updated 2 years ago
- bindings to gnuplot (fork of https://bitbucket.org/ogu/gnuplot-ocaml/)☆13May 6, 2024Updated 2 years ago
- ☆12Nov 21, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- autonomous nanogpt optimizer speedrun☆109May 14, 2026Updated 2 months ago
- Adapting Karpathy's autoresearch to train SAEs on language models.☆19Mar 12, 2026Updated 4 months ago
- ∂B nets: learning discrete, boolean-valued functions by gradient descent☆21Jan 30, 2024Updated 2 years ago
- ☆13Jun 2, 2024Updated 2 years ago
- Efficient Finetuning for OpenAI GPT-OSS☆24Oct 2, 2025Updated 10 months ago
- Code for the paper "Searching Privacy Risks in Multi-Agent Systems via Simulation"☆24Oct 13, 2025Updated 9 months ago
- Repo for Paper: Discovering Interpretable Algorithms by Decompiling Transformers to RASP☆15May 25, 2026Updated 2 months ago