☆21Oct 31, 2022Updated 3 years ago
Alternatives and similar repositories for EvoMoE
Users that are interested in EvoMoE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official repository for the experiments included in the paper titled "Patch-level Routing in Mixture-of-Experts is Provably Sample-ef…☆14Feb 12, 2026Updated 5 months ago
- The source code of "Merging Experts into One: Improving Computational Efficiency of Mixture of Experts (EMNLP 2023)":☆47Updated this week
- ☆19Sep 15, 2022Updated 3 years ago
- This package implements THOR: Transformer with Stochastic Experts.☆64Oct 7, 2021Updated 4 years ago
- [ACL 2023 Findings] Emergent Modularity in Pre-trained Transformers☆26Jun 7, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Pytorch implementation of the paper "Debiasing the Cloze Task in Sequential Recommendation with Bidirectional Transformers".☆12Jan 22, 2023Updated 3 years ago
- ☆17Dec 9, 2022Updated 3 years ago
- ☆37Nov 13, 2020Updated 5 years ago
- [MIDL 2023] Official Imeplementation of "Making Your First Choice: To Address Cold Start Problem in Vision Active Learning"☆37Aug 3, 2023Updated 3 years ago
- A high-performance distributed deep learning system targeting large-scale and automated distributed training.☆340Dec 13, 2025Updated 7 months ago
- ☆16Aug 28, 2022Updated 3 years ago
- ☆28Feb 26, 2023Updated 3 years ago
- ☆15Mar 5, 2024Updated 2 years ago
- This PyTorch package implements MoEBERT: from BERT to Mixture-of-Experts via Importance-Guided Adaptation (NAACL 2022).☆114May 2, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 本项目使用python对影响共享单车使用量的因素进行可视化分析,并使用lightGBM算法对已知条件下的共享单车使用量进行预测。其中为了选择最优模型,使用了k折交叉验证和网格搜索选择最优参数。☆10Jul 15, 2020Updated 6 years ago
- Official PyTorch Implementation of EMoE: Unlocking Emergent Modularity in Large Language Models [main conference @ NAACL2024]☆39May 28, 2024Updated 2 years ago
- A comprehensive overview of Data Distillation and Condensation (DDC). DDC is a data-centric task where a representative (i.e., small but …☆13Dec 1, 2022Updated 3 years ago
- ☆18May 26, 2020Updated 6 years ago
- TurboServe: Serving Streaming Video Generation Efficiently and Economically☆40Jul 12, 2026Updated 3 weeks ago
- ☆11Nov 14, 2021Updated 4 years ago
- Examples for MS-AMP package.☆30Jul 17, 2025Updated last year
- 很久以前写的人生倒计时工具,由于博客内无法运行,拿出来☆11Jun 9, 2022Updated 4 years ago
- ☆10Apr 2, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- pytorch implementation for "Variational Autoencoder with Implicit Optimal Priors".☆11Oct 12, 2020Updated 5 years ago
- ☆13Nov 19, 2020Updated 5 years ago
- code repo for EMNLP'21 Finding Counter-Interference Adapter for Multilingual Machine Translation☆18Oct 19, 2022Updated 3 years ago
- Sentiment analysis meets music☆11Nov 23, 2014Updated 11 years ago
- PyTorch implementation of Soft MoE by Google Brain in "From Sparse to Soft Mixtures of Experts" (https://arxiv.org/pdf/2308.00951.pdf)☆83Oct 5, 2023Updated 2 years ago
- 京东薅羊毛脚本,自动签到,做任务等docker一键启动。有使用上的问题可以加qq群644989387交流。【以上内容为原作者说明】☆10Feb 8, 2022Updated 4 years ago
- ☆11Sep 26, 2022Updated 3 years ago
- ☆18Dec 9, 2025Updated 8 months ago
- ☆27Aug 23, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2026] SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based Benchmark☆22May 6, 2026Updated 3 months ago
- [WWW 2023] Official code of "Adap-$\tau$: Adaptively Modulating Embedding Magnitude for Recommendation"☆29Jan 4, 2024Updated 2 years ago
- PyTorch implementation of "A Simple Baseline for Low-Budget Active Learning".☆14Dec 22, 2021Updated 4 years ago
- ☆19May 13, 2022Updated 4 years ago
- Compression for Foundation Models☆37Jul 21, 2025Updated last year
- Efficient Expert Pruning for Sparse Mixture-of-Experts Language Models: Enhancing Performance and Reducing Inference Costs☆25Nov 11, 2025Updated 8 months ago
- AutoPEFT: Automatic Configuration Search for Parameter-Efficient Fine-Tuning (Zhou et al.; TACL 2024)☆51Mar 17, 2024Updated 2 years ago