Scalable and Stable Parallelization of Nonlinear RNNS
☆33Jun 28, 2026Updated 3 months ago
Alternatives and similar repositories for elk
Users that are interested in elk are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Parallelizing non-linear sequential models over the sequence length☆56Jun 23, 2025Updated last year
- Training Recurrent Neural Networks via Forward Propagation Through Time☆42Jun 11, 2021Updated 5 years ago
- Implementation of the "Online learning of long-range dependencies" paper, NeurIPS 2023☆21Nov 4, 2024Updated last year
- HGRN2: Gated Linear RNNs with State Expansion☆59Aug 20, 2024Updated 2 years ago
- ☆212Sep 11, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Flax Implementation of DreamerV3 on Crafter☆19Nov 29, 2025Updated 10 months ago
- speed-running solving robot manipulation tasks☆24Oct 31, 2024Updated last year
- ☆21Jan 4, 2023Updated 3 years ago
- ☆12May 16, 2024Updated 2 years ago
- Code for "Structured Linear CDEs: Maximally Expressive and Parallel-in-Time Sequence Models" (NeurIPS 2025, Spotlight)☆19Feb 17, 2026Updated 7 months ago
- ☆16Sep 24, 2026Updated 2 weeks ago
- Non official implementation of the Linear Recurrent Unit (LRU, Orvieto et al. 2023)☆65Sep 3, 2025Updated last year
- Reference implementation of "Softmax Attention with Constant Cost per Token" (Heinsen, 2024)☆25Jun 6, 2024Updated 2 years ago
- Deep memory and sequence models in JAX☆36Sep 25, 2026Updated 2 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Aug 21, 2026Updated last month
- Attention Kernels for Symmetric Power Transformers☆131Sep 25, 2025Updated last year
- ☆34Oct 22, 2024Updated last year
- Repository to reproduce the results of the paper "Holomorphic Equilibrium Propagation Computes Exact Gradients Through Finite Size Oscill…☆11Oct 20, 2024Updated last year
- Comparison between GFlowNets & Maximum Entropy RL☆19Feb 19, 2024Updated 2 years ago
- ☆14Oct 11, 2022Updated 3 years ago
- GP Sinkhorn Implementation, paper: https://www.mdpi.com/1099-4300/23/9/1134☆23May 1, 2022Updated 4 years ago
- Implementation of Hyena Hierarchy in JAX☆10Apr 30, 2023Updated 3 years ago
- Hrrformer: A Neuro-symbolic Self-attention Model (ICML23)☆66Oct 8, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Experiments on the impact of depth in transformers and SSMs.☆47Oct 23, 2025Updated 11 months ago
- A State-Space Model with Rational Transfer Function Representation.☆85May 17, 2024Updated 2 years ago
- FlashRNN - Fast RNN Kernels with I/O Awareness☆189Sep 14, 2026Updated 3 weeks ago
- [ICCV 2025] Code for "Flow Stochastic Segmentation Networks"☆19Jun 9, 2026Updated 4 months ago
- nanoGPT using Equinox☆15Mar 3, 2023Updated 3 years ago
- DeltaProduct is a new linear recurrent neural network architecture that uses products of generalized Householder matrices as state-transi…☆18Oct 13, 2025Updated 11 months ago
- Subset-Norm and Subset-Momentum. This repo is built on top of https://github.com/jiaweizzhao/GaLore.☆19Jul 9, 2025Updated last year
- ☆36Apr 12, 2024Updated 2 years ago
- AGaLiTe: Approximate Gated Linear Transformers for Online Reinforcement Learning (Published in TMLR)☆24Oct 15, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- MonarchAttention: Zero-Shot Conversion to Fast, Hardware-Aware Structured Attention (NeurIPS'25 Spotlight)☆27Sep 25, 2026Updated 2 weeks ago
- nanoGPT-like codebase for LLM training☆120Nov 7, 2025Updated 11 months ago
- A simple hypernetwork implementation in jax using haiku.☆24Aug 20, 2022Updated 4 years ago
- Code for lin-RFM used for sparse recovery tasks☆17Mar 13, 2025Updated last year
- NanoGPT speedrun in JAX. Originally at https://nor-git.pages.dev/modded-nanogpt-jax/☆17Aug 28, 2025Updated last year
- krazy grid world☆26Mar 2, 2020Updated 6 years ago
- ☆25Oct 21, 2024Updated last year