This repositories contains the reference implementation for the Sparse Delta Memory paper.More precisely, it contains the model definition as well as triton and cuda kernels for the Sparse Delta Memory layer.
☆38Jul 9, 2026Updated 2 months ago
Alternatives and similar repositories for sparse-delta-memory
Users that are interested in sparse-delta-memory are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DeltaProduct is a new linear recurrent neural network architecture that uses products of generalized Householder matrices as state-transi…☆17Oct 13, 2025Updated 10 months ago
- Code and data to explore neural scaling laws of xLSTM and Transformer models.☆24Apr 8, 2026Updated 5 months ago
- Code for Fast-weight Product Key Memory (FwPKM)☆24Mar 18, 2026Updated 5 months ago
- ☆86Aug 19, 2026Updated 3 weeks ago
- ☆19Dec 12, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [Tech Report] Expanded Hyper-Connections☆62Jul 21, 2026Updated last month
- Attention variant with per-channel multiplicative decay☆50Jun 3, 2026Updated 3 months ago
- Implementation of Fast Weight Attention☆34Aug 20, 2026Updated 3 weeks ago
- Official repository for Parallax (Parameterized Local Linear Attention)☆69Jul 30, 2026Updated last month
- Flash-Linear-Attention models beyond language☆21Aug 28, 2025Updated last year
- Flash Attention in 300-500 lines of CUDA/C++☆39Aug 22, 2025Updated last year
- Multi-modal, multi-task modeling of the mouse visual cortex☆19May 5, 2026Updated 4 months ago
- High-performance GPU kernels for LLM inference in OpenAI Triton. Fused RMSNorm, SwiGLU, INT8 GEMM with benchmarks and roofline analysis.☆43Jul 22, 2026Updated last month
- ☆27Jul 13, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 更纯粹、更高压缩率的Tokenizer in Rust☆14Updated this week
- Implementation and explorations into DiscoRL, Discovering state-of-the-art reinforcement learning algorithms, David Silver's last work at…☆21Jun 13, 2026Updated 3 months ago
- Official Project Page for HLA: Higher-order Linear Attention (https://arxiv.org/abs/2510.27258)☆102Jun 15, 2026Updated 2 months ago
- Expanding linear RNN state-transition matrix eigenvalues to include negatives improves state-tracking tasks and language modeling without…☆22Mar 15, 2025Updated last year
- Official repository for "SSA: Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space"☆30May 7, 2026Updated 4 months ago
- Experiments on the impact of depth in transformers and SSMs.☆47Oct 23, 2025Updated 10 months ago
- XWikisCorpus, cross-lingual summarisation, multi-lingual summarisation, pre-trained language models, zero-shot and few-shot summarisation…☆10Nov 4, 2022Updated 3 years ago
- A Multi-Policy, Multi-Agent RL Training Framework☆32Jun 16, 2026Updated 2 months ago
- STABILIZING GRADIENTS FOR DEEP NEURAL NETWORKS VIA EFFICIENT SVD PARAMETERIZATION☆16Jun 5, 2018Updated 8 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆48Dec 13, 2025Updated 8 months ago
- Code with CliqueFlowmer model for Optimal Computational Materials Discovery☆17Sep 6, 2026Updated last week
- ☆31Aug 10, 2026Updated last month
- Delta Attention Residuals - supplementary code and pretrained models☆43May 20, 2026Updated 3 months ago
- 🔥 A minimal training framework for scaling FLA models☆416Apr 22, 2026Updated 4 months ago
- ☆13Sep 27, 2022Updated 3 years ago
- Efficient PScan implementation in PyTorch☆17Jan 2, 2024Updated 2 years ago
- ☆13Aug 19, 2024Updated 2 years ago
- mHC-lite: You Don’t Need 20 Sinkhorn-Knopp Iterations☆94Jan 12, 2026Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- An official implementation for the EMNLP 2023 Findings paper "Prompt-Based Editing for Text Style Transfer"☆13Dec 9, 2023Updated 2 years ago
- Fast Punctuation Restoration using Transformer Models for Vietnamese☆11Jun 10, 2022Updated 4 years ago
- Assist Non-native Viewers: Multimodal Crosslingual Summarization for How2 Videos☆10Sep 2, 2024Updated 2 years ago
- ☆20Dec 24, 2024Updated last year
- ☆12Sep 26, 2019Updated 6 years ago
- Julia implementation of flash-attention operation for neural networks.☆11May 31, 2023Updated 3 years ago
- Compositional Muon release☆25Jun 5, 2026Updated 3 months ago