Non-official implementation of "Attention as an RNN" from https://arxiv.org/pdf/2405.13956, efficient associative parallel prefix scan and recurrent version implemented.
☆28Jul 27, 2024Updated last year
Alternatives and similar repositories for Attention-as-RNN
Users that are interested in Attention-as-RNN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reference implementation of "Softmax Attention with Constant Cost per Token" (Heinsen, 2024)☆25Jun 6, 2024Updated 2 years ago
- Zicx's Notebook.☆11Nov 7, 2025Updated 8 months ago
- [IEEE TETCI 2025] Official repository of the PETformer paper: "PETformer: Long-term Time Series Forecasting via Placeholder-enhanced Tran…☆14Dec 6, 2024Updated last year
- [ICLR 2025] This repo is the official implementation of "The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs".☆13Jan 25, 2025Updated last year
- ☆16Jul 24, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICML 2024] Official repository of ICML 2024 - RoboMP2: A Robotic Multimodal Perception-Planning Framework with Multimodal Large Language…☆12Apr 4, 2026Updated 3 months ago
- This is my codes that can visualize the psnr image in testing videos.☆15Apr 4, 2021Updated 5 years ago
- The official Languini Kitchen repository☆14May 6, 2024Updated 2 years ago
- Unofficial PyTorch implementation of 'Fast and High-Quality Image Denoising via Malleable Convolutions'.☆12Mar 7, 2026Updated 4 months ago
- [ECCV 2024] Official repository of ECCV 2024 paper: Object-Conditioned Energy-Based Attention Map Alignment in Text-to-Image Diffusion M…☆16May 24, 2025Updated last year
- Implementation for paper "Link Prediction on Heterophilic Graphs via Disentangled Representation Learning"☆13Aug 26, 2022Updated 3 years ago
- Torch MinGRU implementation based on "Were RNNs All We Needed?"☆23Dec 5, 2024Updated last year
- Official implementation of RMoE (Layerwise Recurrent Router for Mixture-of-Experts)☆33Aug 4, 2024Updated last year
- ☆24Sep 25, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Train toy models using multi-token prediction objective☆14Apr 18, 2026Updated 3 months ago
- High order Moment Models☆42Nov 13, 2025Updated 8 months ago
- [NeurIPS 2024] Official implementation of NeurIPS 2024 paepr "Flow Priors for Linear Inverse Problems via Iterative Corrupted Trajectory …☆26Feb 24, 2025Updated last year
- [ICLR 2025] "Understanding Constraint Inference in Safety-Critical Inverse Reinforcement Learning"☆16Nov 30, 2025Updated 7 months ago
- ☆19Jul 21, 2019Updated 7 years ago
- ☆20Aug 6, 2024Updated last year
- Parallel Prefix Sum (Scan) with CUDA.☆15Jul 17, 2020Updated 6 years ago
- ☆36Oct 20, 2020Updated 5 years ago
- Official implementation for "Let Offline RL Flow: Training Conservative Agents in the Latent Space of Normalizing Flows", NeurIPS 2022, O…☆12Jan 31, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆33Jul 1, 2024Updated 2 years ago
- Planning Algorithm for All-Wheel-Steering Platforms☆21Jan 19, 2026Updated 6 months ago
- A Fast, Portable Deep Reinforcement Learning Library for Continuous Control☆13Jul 26, 2023Updated 2 years ago
- Official implementation of the UMDQN algorithm presented in the scientific research paper entitled "Distributional Reinforcement Learning…☆11Jun 3, 2022Updated 4 years ago
- Wasserstein GAN with gradient penalty (WGAN-GP) applied to financial time series.☆17Oct 17, 2018Updated 7 years ago
- ☆10Sep 5, 2024Updated last year
- ☆17Dec 19, 2024Updated last year
- A Data Science pipeline for Algorithmic Trading: A comparative study in applications to Finance and cryptoeconomics☆14Jul 1, 2022Updated 4 years ago
- Code for ICLR 2022 Paper (HyperDQN: A Randomized Exploration Method for Deep Reinforcement Learning)☆12Nov 28, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Alpha mining with DEAP-based genetic programming.☆11Jul 7, 2023Updated 3 years ago
- ROS packages for facilitating text-to-speech and the use of Amazon Polly.☆18Feb 8, 2022Updated 4 years ago
- Source code for Interpretable Reward Redistribution in Reinforcement Learning: A Causal Approach (NeurIPS 2023)☆10Dec 12, 2023Updated 2 years ago
- ☆11Nov 13, 2025Updated 8 months ago
- OpenAI gym environments for goal-conditioned and language-conditioned reinforcement learning☆14Jan 27, 2026Updated 5 months ago
- Code for the paper "Minimum-Delay Adaptation in Non-Stationary Reinforcement Learning via Online High-Confidence Change-Point Detection"☆11Aug 7, 2023Updated 2 years ago
- Reference implementation of models from Nyonic Model Factory☆12May 13, 2024Updated 2 years ago