☆19Sep 15, 2022Updated 4 years ago
Alternatives and similar repositories for MPOE
Users that are interested in MPOE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13May 21, 2023Updated 3 years ago
- The official repository for the experiments included in the paper titled "Patch-level Routing in Mixture-of-Experts is Provably Sample-ef…☆15Feb 12, 2026Updated 7 months ago
- Spatial Mixture-of-Experts☆21Nov 29, 2022Updated 3 years ago
- ☆21Oct 31, 2022Updated 3 years ago
- BESA is a differentiable weight pruning technique for large language models.☆17Mar 4, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆92Apr 2, 2022Updated 4 years ago
- [ICLR 2023] "Sparse MoE as the New Dropout: Scaling Dense and Self-Slimmable Transformers" by Tianlong Chen*, Zhenyu Zhang*, Ajay Jaiswal…☆56Feb 28, 2023Updated 3 years ago
- ☆10Dec 26, 2023Updated 2 years ago
- A PYNQ overlay demonstrating the Xilinx RFSoC SD-FEC☆13Updated this week
- Nsolv - A front-end that allows multiple SMTLIBv2 compliant solvers to executed in parallel.☆11Dec 7, 2012Updated 13 years ago
- Framework based on Partial Reconfiguration for chip characterization utilizing ring-oscillator PUFs☆14Apr 1, 2020Updated 6 years ago
- Implementation of SelfExtend from the paper "LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning" from Pytorch and Zeta☆13Nov 11, 2024Updated last year
- A fork of llama3.c used to do some R&D on inferencing☆24Dec 20, 2024Updated last year
- Source code of ACL 2023 Main Conference Paper "PAD-Net: An Efficient Framework for Dynamic Networks".☆14Aug 22, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICCV23] Robust Mixture-of-Expert Training for Convolutional Neural Networks by Yihua Zhang, Ruisi Cai, Tianlong Chen, Guanhua Zhang, Hua…☆69Aug 27, 2023Updated 3 years ago
- ☆146Jul 21, 2024Updated 2 years ago
- The source code of "Merging Experts into One: Improving Computational Efficiency of Mixture of Experts (EMNLP 2023)":☆47Aug 22, 2026Updated last month
- [ICLR‘24 Spotlight] Code for the paper "Merge, Then Compress: Demystify Efficient SMoE with Hints from Its Routing Policy"☆110Jun 20, 2025Updated last year
- Codes of the paper Deformable Butterfly: A Highly Structured and Sparse Linear Transform.☆16Nov 1, 2021Updated 4 years ago
- Elana: A Simple Energy & Latency Analyzer for LLMs☆16Apr 3, 2026Updated 5 months ago
- Code for "Memory Efficient Meta-Learning with Large Images"☆11Nov 24, 2021Updated 4 years ago
- Implementation of "Gradients without backpropagation" paper (https://arxiv.org/abs/2202.08587) using functorch☆114Jun 14, 2023Updated 3 years ago
- /g/ programming language☆13Nov 9, 2011Updated 14 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- This is the official implementation of the ICML 2023 paper - Can Forward Gradient Match Backpropagation ?☆13May 31, 2023Updated 3 years ago
- Real-Time Hardware Sorter, Using A Multi-Dimensional Sorting Algorithm☆21Jan 4, 2026Updated 8 months ago
- Machine Learning System☆14May 11, 2020Updated 6 years ago
- ☆12Oct 14, 2025Updated 11 months ago
- "Stochasticity in Neural ODEs: An Empirical Study". Experiments from the paper☆13Apr 27, 2020Updated 6 years ago
- FPGA routing delay sensors for effective remote power analysis attacks☆14Aug 13, 2024Updated 2 years ago
- Codebase for ACL 2023 paper "Mixture-of-Domain-Adapters: Decoupling and Injecting Domain Knowledge to Pre-trained Language Models' Memori…☆51Oct 8, 2023Updated 2 years ago
- Portable SIMD specification☆19Apr 24, 2017Updated 9 years ago
- This is the implementation of the paper AdaMix: Mixture-of-Adaptations for Parameter-efficient Model Tuning (https://arxiv.org/abs/2205.1…☆139Aug 14, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 100G Udp Link For axi Stream☆17Jun 27, 2023Updated 3 years ago
- Python implementation of STATIS for analysis of several data tables☆12Aug 14, 2017Updated 9 years ago
- The Déjà Vu programming language☆16Aug 12, 2014Updated 12 years ago
- Repository to store all design and testbench files for Senior Design☆23Apr 16, 2020Updated 6 years ago
- Code for "MetaFun: Meta-Learning with Iterative Functional Updates"☆14Aug 27, 2020Updated 6 years ago
- Web application for the simulation of day-ahead energy markets☆11Jul 9, 2017Updated 9 years ago
- ☆30Sep 28, 2023Updated 2 years ago