An implementation is provided here for the NeurIPS2024 paper "MemoryFormer : Minimize Transformer Computation by Removing Fully-Connected Layers"
☆16Mar 24, 2026Updated 5 months ago
Alternatives and similar repositories for MemoryFormer
Users that are interested in MemoryFormer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Homepage for paper “MeKi : Memory-based Expert Knowledge Injection for Efficient LLM Scaling”☆29Mar 5, 2026Updated 5 months ago
- An open-sourced PyTorch library for developing energy efficient multiplication-less models and applications.☆14Feb 3, 2025Updated last year
- Spartan is an algorithm for training sparse neural network models. This repository accompanies the paper "Spartan Differentiable Sparsity…☆26Oct 31, 2022Updated 3 years ago
- ☆14Sep 7, 2024Updated last year
- Code for ICML 2024 paper☆34Sep 18, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆16Jun 11, 2025Updated last year
- ☆39Jun 2, 2026Updated 2 months ago
- This is the project of my undergraduate thesis, an enhanced Capsule Network that integrates EEG and ECG signals for improved emotion reco…☆13Jan 17, 2024Updated 2 years ago
- The main code of our proposed multi-resolution interactive transformer☆11May 16, 2025Updated last year
- ☆25Oct 31, 2024Updated last year
- Using Capsule Network for EEG emotion classification using Digit-Caps model☆10Apr 30, 2023Updated 3 years ago
- [EMBC-2025] PyTorch implementation of EEG-PatchFormer☆15Apr 9, 2025Updated last year
- Official implementation of [NeurIPS 2024] Con4m: Context-aware Consistency Learning Framework for Segmented Time Series Classification☆15May 14, 2025Updated last year
- Binary neural networks developed by Huawei Noah's Ark Lab☆29Feb 19, 2021Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- TensorFlow implementation of GhostNet: More Features from Cheap Operations.☆10Feb 6, 2020Updated 6 years ago
- ☆18Sep 23, 2025Updated 11 months ago
- [NeurIPS 2024] VeLoRA : Memory Efficient Training using Rank-1 Sub-Token Projections☆22Oct 15, 2024Updated last year
- [NeurIPS 2023] Code release for "Going Beyond Linear Mode Connectivity: The Layerwise Linear Feature Connectivity"☆19Oct 19, 2023Updated 2 years ago
- Adaptive Sparse Attention and Robust Learning for Multimodal Dynamic Time Series☆19Mar 23, 2026Updated 5 months ago
- ICML2025: Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning☆55May 1, 2025Updated last year
- Official repo for BWLer: Barycentric Weight Layer☆31Aug 11, 2026Updated 2 weeks ago
- Generative Modeling via Drifting in MLX☆43Feb 6, 2026Updated 6 months ago
- Mixture-of-Basis-Experts for Compressing MoE-based LLMs☆37Dec 24, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This repository is the implementation of "A Lightweight Spiking Neural Network for EEG-Based Motor Imagery Classification".☆16Jun 7, 2025Updated last year
- Self Reproduction Code of Paper "Reducing Transformer Key-Value Cache Size with Cross-Layer Attention (MIT CSAIL)☆17May 24, 2024Updated 2 years ago
- [IEEE GRSL 2025] PHDMamba: Progressive Hybrid Mamba for Hyperspectral Image Classification☆17Jan 18, 2026Updated 7 months ago
- [ICLR 2023] "Sparse MoE as the New Dropout: Scaling Dense and Self-Slimmable Transformers" by Tianlong Chen*, Zhenyu Zhang*, Ajay Jaiswal…☆56Feb 28, 2023Updated 3 years ago
- This repository contains scripts to generate weather- and climate-driven power supply and demand time series for power and energy system …☆21Oct 9, 2025Updated 10 months ago
- ☆60Nov 18, 2025Updated 9 months ago
- [ICML 2025 Oral] Mixture of Lookup Experts☆79Dec 3, 2025Updated 8 months ago
- ☆15Sep 1, 2025Updated 11 months ago
- [ACL 2025] Squeezed Attention: Accelerating Long Prompt LLM Inference☆58Nov 20, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- EDA toolchain for processing-in-memory architectures, including an architecture synthesizer, a compiler, and a simulator☆27Jun 12, 2025Updated last year
- [ICCV 2023] Source code of "Fcaformer: Forward Cross Attention in Hybrid Vision Transformer"☆25Aug 23, 2023Updated 3 years ago
- ☆32Mar 31, 2025Updated last year
- Spatial-Temporal Graph-Enhanced Transformer for EEG Based Major Depressive Disorder Detection☆23Feb 8, 2026Updated 6 months ago
- MicroMix: Efficient Mixed-Precision Quantization with Microscaling Formats for Large Language Models☆30Apr 2, 2026Updated 4 months ago
- Code for reproducing "AC/DC: Alternating Compressed/DeCompressed Training of Deep Neural Networks" (NeurIPS 2021)☆23Nov 9, 2021Updated 4 years ago
- ☆19Nov 21, 2025Updated 9 months ago