Official implementation of the transformer (TF) architecture suggested in a paper entitled "Looped Transformers as Programmable Computers"
☆42Apr 8, 2023Updated 3 years ago
Alternatives and similar repositories for Looped-Transformer
Users that are interested in Looped-Transformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆84Aug 31, 2023Updated 2 years ago
- A transformer that executes a one-instruction Turing-complete computer — two approaches: hand-coded weights (no training) and learned fro…☆41Mar 3, 2026Updated 4 months ago
- ☆20Oct 25, 2022Updated 3 years ago
- ☆21Mar 1, 2023Updated 3 years ago
- ☆10Oct 28, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The git repository of Modular Prompted Chatbot paper☆35May 24, 2023Updated 3 years ago
- Code for Language-Interfaced FineTuning for Non-Language Machine Learning Tasks.☆135Nov 11, 2024Updated last year
- Embroid: Unsupervised Prediction Smoothing Can Improve Few-Shot Classification☆11Aug 12, 2023Updated 2 years ago
- Generative Equilibrium Transformer☆28Nov 11, 2023Updated 2 years ago
- PyTorch implementation for the Deep Symbolic Simplification Without Human Knowledge☆14Feb 25, 2021Updated 5 years ago
- An implementation is provided here for the NeurIPS2024 paper "MemoryFormer : Minimize Transformer Computation by Removing Fully-Connected…☆16Mar 24, 2026Updated 4 months ago
- ☆13Jul 9, 2018Updated 8 years ago
- Sampling-Based Minimum Bayes-Risk Decoding for Neural Machine Translation☆16Oct 14, 2022Updated 3 years ago
- ☆16Jul 7, 2026Updated 2 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆13Jun 26, 2024Updated 2 years ago
- Official implementation of ICLR 2025 'LORO: Parameter and Memory Efficient Pretraining via Low-rank Riemannian Optimization'☆17Apr 24, 2025Updated last year
- FR-Train: A Mutual Information-Based Approach to Fair and Robust Training (ICML 2020)☆13Jun 3, 2021Updated 5 years ago
- Representation Learning in RL☆13Jun 1, 2022Updated 4 years ago
- Position Coupling: Improving Length Generalization of Arithmetic Transformers Using Task Structure (NeurIPS 2024) + Arithmetic Transfor…☆14Oct 26, 2025Updated 9 months ago
- TDMS 2.0 support for F# and C#☆13Dec 26, 2022Updated 3 years ago
- Code for the arXiv preprint "Answer, Assemble, Ace: Understanding How Transformers Answer Multiple Choice Questions"☆15Aug 2, 2025Updated 11 months ago
- Claude Code skill for KAI presentation design in HTML☆16Mar 20, 2026Updated 4 months ago
- Code for the paper "Decomposing the Enigma: Subgoal-based Demonstration Learning for Formal Theorem Proving"☆20May 25, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Legacy Code of ZJU Campus App for iOS☆11Jan 31, 2024Updated 2 years ago
- Implementation of "LM-Infinite: Simple On-the-Fly Length Generalization for Large Language Models"☆40Nov 11, 2024Updated last year
- ☆13May 23, 2021Updated 5 years ago
- 这是一个大学四年的cs基础课部分专业课的复习笔记的扫描版备份仓库☆12Jun 29, 2019Updated 7 years ago
- ☆18Jul 10, 2022Updated 4 years ago
- PyTorch implementation of StableMask (ICML'24)☆15Jun 27, 2024Updated 2 years ago
- ☆93Aug 18, 2024Updated last year
- [ICLR 2025] "Training LMs on Synthetic Edit Sequences Improves Code Synthesis" (Piterbarg, Pinto, Fergus)☆19Feb 11, 2025Updated last year
- ☆16Feb 21, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for the paper "Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns"☆18Mar 15, 2024Updated 2 years ago
- [ICLR 2025] Permute-and-Flip: An optimally robust and watermarkable decoder for LLMs☆19Mar 20, 2025Updated last year
- ☆20May 5, 2023Updated 3 years ago
- Equivalent Linear Mappings of Large Language Models☆35Nov 7, 2025Updated 8 months ago
- My coding assignment for UIUC-CS441-Applied Machine Learning☆10Mar 24, 2022Updated 4 years ago
- Collect papers related to personalized text generation☆18Sep 6, 2021Updated 4 years ago
- [CoLM 24] Official Repository of MambaByte: Token-free Selective State Space Model☆27Oct 12, 2024Updated last year