This repositories contains the reference implementation for the Sparse Delta Memory paper.More precisely, it contains the model definition as well as triton and cuda kernels for the Sparse Delta Memory layer.
☆36Jul 9, 2026Updated last month
Alternatives and similar repositories for sparse-delta-memory
Users that are interested in sparse-delta-memory are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DeltaProduct is a new linear recurrent neural network architecture that uses products of generalized Householder matrices as state-transi…☆17Oct 13, 2025Updated 10 months ago
- Code and data to explore neural scaling laws of xLSTM and Transformer models.☆24Apr 8, 2026Updated 4 months ago
- Code for Fast-weight Product Key Memory (FwPKM)☆21Mar 18, 2026Updated 5 months ago
- ☆85Updated this week
- ☆19Dec 12, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [Tech Report] Expanded Hyper-Connections☆61Jul 21, 2026Updated last month
- The Newton-Muon optimizer☆30Jun 5, 2026Updated 2 months ago
- Attention variant with per-channel multiplicative decay☆50Jun 3, 2026Updated 2 months ago
- Official repository for Parallax (Parameterized Local Linear Attention)☆68Jul 30, 2026Updated 3 weeks ago
- Flash-Linear-Attention models beyond language☆21Aug 28, 2025Updated 11 months ago
- Flash Attention in 300-500 lines of CUDA/C++☆39Aug 22, 2025Updated last year
- Multi-modal, multi-task modeling of the mouse visual cortex☆19May 5, 2026Updated 3 months ago
- High-performance GPU kernels for LLM inference in OpenAI Triton. Fused RMSNorm, SwiGLU, INT8 GEMM with benchmarks and roofline analysis.☆40Jul 22, 2026Updated last month
- ☆27Jul 13, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 更纯粹、更高压缩率的Tokenizer in Rust☆14Dec 21, 2024Updated last year
- Aurora optimizer release☆156Jul 18, 2026Updated last month
- Implementation and explorations into DiscoRL, Discovering state-of-the-art reinforcement learning algorithms, David Silver's last work at…☆21Jun 13, 2026Updated 2 months ago
- Official Project Page for HLA: Higher-order Linear Attention (https://arxiv.org/abs/2510.27258)☆103Jun 15, 2026Updated 2 months ago
- Training-free KV cache compression via E8 lattice VQ. 2-bit KV that preserves retrieval (30/30 NIAH vs TurboQuant 0/30). Calibration-free…☆25Aug 4, 2026Updated 2 weeks ago
- Expanding linear RNN state-transition matrix eigenvalues to include negatives improves state-tracking tasks and language modeling without…☆22Mar 15, 2025Updated last year
- Official repository for "SSA: Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space"☆28May 7, 2026Updated 3 months ago
- Experiments on the impact of depth in transformers and SSMs.☆46Oct 23, 2025Updated 10 months ago
- XWikisCorpus, cross-lingual summarisation, multi-lingual summarisation, pre-trained language models, zero-shot and few-shot summarisation…☆10Nov 4, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A Multi-Policy, Multi-Agent RL Training Framework☆32Jun 16, 2026Updated 2 months ago
- Lean formalizations for the paper "On the paucity of lattice triangles"☆19Mar 26, 2026Updated 4 months ago
- ☆30Aug 10, 2026Updated last week
- 🔥 A minimal training framework for scaling FLA models☆411Apr 22, 2026Updated 4 months ago
- ☆13Sep 27, 2022Updated 3 years ago
- Efficient PScan implementation in PyTorch☆17Jan 2, 2024Updated 2 years ago
- This is the official code implementation of Bongard-OpenWorld (ICLR 2024).☆14Jan 6, 2025Updated last year
- ☆13Aug 19, 2024Updated 2 years ago
- Fast Punctuation Restoration using Transformer Models for Vietnamese☆11Jun 10, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Assist Non-native Viewers: Multimodal Crosslingual Summarization for How2 Videos☆10Sep 2, 2024Updated last year
- [EMNLP'22] Textual Manifold-based Defense Against Natural Language Adversarial Examples☆11Apr 6, 2023Updated 3 years ago
- A framework to train language models to learn invariant representations.☆14Jan 24, 2022Updated 4 years ago
- Transformers components but in Triton☆34May 9, 2025Updated last year
- ☆20Dec 24, 2024Updated last year
- Julia implementation of flash-attention operation for neural networks.☆11May 31, 2023Updated 3 years ago
- Compositional Muon release☆25Jun 5, 2026Updated 2 months ago