Unofficial implementation of Linear Recurrent Units, by Deepmind, in Pytorch
☆78Apr 22, 2025Updated last year
Alternatives and similar repositories for LRU-pytorch
Users that are interested in LRU-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Non official implementation of the Linear Recurrent Unit (LRU, Orvieto et al. 2023)☆62Sep 3, 2025Updated 10 months ago
- Implementations of various linear RNN layers using pytorch and triton☆55Aug 4, 2023Updated 2 years ago
- This repository is pytorch version implement of LRU from the paper "Resurrecting Recurrent Neural Networks for Long Sequences" (https://a…☆12May 22, 2023Updated 3 years ago
- ☆36Nov 22, 2024Updated last year
- ☆30Feb 27, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆325Jan 8, 2025Updated last year
- Code for "Theoretical Foundations of Deep Selective State-Space Models" (NeurIPS 2024)☆17Jan 7, 2025Updated last year
- Curse-of-memory phenomenon of RNNs in sequence modelling☆19May 8, 2025Updated last year
- ☆16Mar 3, 2023Updated 3 years ago
- Dreamer on JAX☆16Jan 19, 2022Updated 4 years ago
- ☆45Apr 30, 2018Updated 8 years ago
- Closed-loop simulator of complex behavior and learning based on reinforcement learning and deep neural networks☆15Mar 20, 2026Updated 4 months ago
- Accelerated First Order Parallel Associative Scan☆198Jan 7, 2026Updated 6 months ago
- Modular and Hierachical RL baseline solution for the IGLU RL track @ NeurIPS 2022☆20Sep 16, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The accompanying code for "Simplifying and Understanding State Space Models with Diagonal Linear RNNs" (Ankit Gupta, Harsh Mehta, Jonatha…☆23Dec 30, 2022Updated 3 years ago
- Codebase for the paper "How Crucial is Transformer in Decision Transformer?". Containing experiments on different pendulum tasks and code…☆28Mar 24, 2023Updated 3 years ago
- Implementation of Spectral State Space Models☆16Feb 23, 2024Updated 2 years ago
- ☆65Jul 11, 2023Updated 3 years ago
- JAX/Flax implementation of the Hyena Hierarchy☆35Apr 27, 2023Updated 3 years ago
- Training Recurrent Neural Networks via Forward Propagation Through Time☆42Jun 11, 2021Updated 5 years ago
- Docker containers of baseline agents for the Crafter environment☆30Dec 14, 2021Updated 4 years ago
- Embroid: Unsupervised Prediction Smoothing Can Improve Few-Shot Classification☆11Aug 12, 2023Updated 2 years ago
- ☆14Sep 2, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Unofficial implementation of paper : Exploring the Space of Key-Value-Query Models with Intention☆12May 24, 2023Updated 3 years ago
- Unofficial but Efficient Implementation of "Mamba: Linear-Time Sequence Modeling with Selective State Spaces" in JAX☆94Jan 25, 2024Updated 2 years ago
- An implementation of DreamerV2 written in JAX, with support for running multiple random seeds of an experiment on a single GPU.☆18Jan 16, 2023Updated 3 years ago
- ☆29Jul 9, 2024Updated 2 years ago
- Engineering the state of RNN language models (Mamba, RWKV, etc.)☆33May 25, 2024Updated 2 years ago
- STABILIZING GRADIENTS FOR DEEP NEURAL NETWORKS VIA EFFICIENT SVD PARAMETERIZATION☆16Jun 5, 2018Updated 8 years ago
- PyTorch implementation for PaLM: A Hybrid Parser and Language Model.☆10Jan 7, 2020Updated 6 years ago
- ☆15Jun 8, 2026Updated last month
- Self-attentive Associative Memory & SAM-based Two-Memory Model☆61May 4, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official Code Repository for the paper "Key-value memory in the brain"☆32Feb 25, 2025Updated last year
- A repository for DenseSSMs☆90Apr 11, 2024Updated 2 years ago
- Codes for paper : "A Stroke-based RNN for Writer-Independent Online Signature Verification"☆11May 6, 2019Updated 7 years ago
- Benchmarking RL for POMDPs in Pure JAX [Code for "Structured State Space Models for In-Context Reinforcement Learning" (NeurIPS 2023)]☆116Dec 5, 2023Updated 2 years ago
- A short conceptual replication of "Prefrontal cortex as a meta-reinforcement learning system" in Jax.☆19Feb 27, 2023Updated 3 years ago
- ☆35Apr 12, 2024Updated 2 years ago
- Jax implementation of "Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models"☆15May 10, 2024Updated 2 years ago