An implementation is provided here for the NeurIPS2024 paper "MemoryFormer : Minimize Transformer Computation by Removing Fully-Connected Layers"
☆16Mar 24, 2026Updated 6 months ago
Alternatives and similar repositories for MemoryFormer
Users that are interested in MemoryFormer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An open-sourced PyTorch library for developing energy efficient multiplication-less models and applications.☆14Feb 3, 2025Updated last year
- Code for ICML 2024 paper☆34Sep 18, 2025Updated last year
- ☆17Jun 11, 2025Updated last year
- ☆23Oct 26, 2022Updated 3 years ago
- ☆25Oct 31, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An unofficial wrapper of Baidu Baike☆12Feb 20, 2014Updated 12 years ago
- Binary neural networks developed by Huawei Noah's Ark Lab☆30Feb 19, 2021Updated 5 years ago
- 机器学习部分算法实现,分类、聚类、回归(LR、Kmeans、GMM、PCA)☆10Mar 12, 2019Updated 7 years ago
- ☆18Sep 23, 2025Updated last year
- [NeurIPS 2024] VeLoRA : Memory Efficient Training using Rank-1 Sub-Token Projections☆22Oct 15, 2024Updated last year
- 魔镜魔镜,无所不知的魔镜[-_-](并不是)☆13Jun 10, 2021Updated 5 years ago
- ☆13Nov 9, 2014Updated 11 years ago
- Mixture-of-Basis-Experts for Compressing MoE-based LLMs☆40Dec 24, 2025Updated 9 months ago
- MLX implementation of Hierarchical Reasoning Model (HRM) - Adaptive computation for complex reasoning tasks☆29Aug 27, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- dynamic planning, hybrid models, hierarchical active inference, tool use☆15Jun 13, 2025Updated last year
- ☆60Nov 18, 2025Updated 10 months ago
- ☆16Sep 1, 2025Updated last year
- [ICLR 2024] Official pytorch implementation of "Denoising Task Routing for Diffusion Models"☆25Feb 19, 2024Updated 2 years ago
- [ACL 2025] Squeezed Attention: Accelerating Long Prompt LLM Inference☆58Nov 20, 2024Updated last year
- EDA toolchain for processing-in-memory architectures, including an architecture synthesizer, a compiler, and a simulator☆28Jun 12, 2025Updated last year
- Hyperledger Indy/Sovrin/DID Comprehensive Architecture Reference Model (INDY ARM) - Draft document for discussion purposes☆14Jan 25, 2021Updated 5 years ago
- ☆32Mar 31, 2025Updated last year
- ☆19Nov 20, 2025Updated 10 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆23Sep 9, 2026Updated last month
- MicroMix: Efficient Mixed-Precision Quantization with Microscaling Formats for Large Language Models☆32Apr 2, 2026Updated 6 months ago
- Code for reproducing "AC/DC: Alternating Compressed/DeCompressed Training of Deep Neural Networks" (NeurIPS 2021)☆23Nov 9, 2021Updated 4 years ago
- ☆19Nov 21, 2025Updated 10 months ago
- Evaluation code for the EMNLP 2024 Findings paper introducing LongGenBench, a benchmark for long-context generation by large language mod…☆25Oct 8, 2024Updated 2 years ago
- Reference implementation of models from Nyonic Model Factory☆12May 13, 2024Updated 2 years ago
- Official implementation of the transformer (TF) architecture suggested in a paper entitled "Looped Transformers as Programmable Computers…☆49Apr 8, 2023Updated 3 years ago
- LLaDA implementation☆19Jul 24, 2025Updated last year
- A general framework for optimizing DNN dataflow on systolic array☆40Jan 2, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Low-rank sparse attention decomposition for LLM interpretability; active development continues in Llamascopium☆30Nov 9, 2025Updated 11 months ago
- Official PyTorch implementation of CD-MOE☆12Mar 18, 2026Updated 6 months ago
- SAMO: Streaming Architecture Mapping Optimisation☆36Oct 4, 2023Updated 3 years ago
- Residual vector quantization for KV cache compression in large language model☆12Oct 22, 2024Updated last year
- Clustered Compositional Embeddings☆13Oct 25, 2023Updated 2 years ago
- Based on BrainTransformers, BrainGPTForCausalLM is a Large Language Model (LLM) implemented using Spiking Neural Networks (SNN). We are e…☆39Oct 22, 2024Updated last year
- Simple and Ideal Circuit Simulation☆13Dec 4, 2017Updated 8 years ago