☆30Aug 7, 2026Updated this week
Alternatives and similar repositories for X-MoE
Users that are interested in X-MoE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code repo for efficient quantized MoE inference with mixture of low-rank compensators☆39Apr 14, 2025Updated last year
- ☆37May 31, 2026Updated 2 months ago
- This is the implementation for paper: AdaTune: Adaptive Tensor Program CompilationMade Efficient (NeurIPS 2020).☆14May 16, 2021Updated 5 years ago
- ☆21Jul 7, 2026Updated last month
- PyTorch Code for the Paper: "Exploiting Uncertainty of Loss Landscape for Stochastic Optimization [Bhaskara et al. (2019)]☆16Apr 30, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆16May 14, 2025Updated last year
- ☆15Jan 24, 2025Updated last year
- ☆14Aug 28, 2019Updated 6 years ago
- [ICLR 2025] SWIFT: On-the-Fly Self-Speculative Decoding for LLM Inference Acceleration☆70Feb 21, 2025Updated last year
- ☆24Sep 10, 2025Updated 11 months ago
- An recognition oriented deep learning framework for biometric sample quality assessment☆12Aug 24, 2023Updated 2 years ago
- [NeurIPS 2025] L-MTP: Leap Multi-Token Prediction Beyond Adjacent Context for Large Language Models☆32May 8, 2026Updated 3 months ago
- Accelerating MoE with IO and Tile-aware Optimizations☆737Jul 4, 2026Updated last month
- ☆67Jun 25, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICLR 2026] AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size☆15Jan 28, 2026Updated 6 months ago
- ☆15Apr 26, 2022Updated 4 years ago
- Jax implementation of the AdaHessian optimizer☆19Mar 11, 2021Updated 5 years ago
- Starlight: A Kernel Optimizer for GPU Processing☆16Jan 10, 2024Updated 2 years ago
- A logical, reasonably standardized, but flexible project structure for conducting ml research 🍪☆19Apr 9, 2026Updated 4 months ago
- A place to store my knowledge base☆12Apr 27, 2026Updated 3 months ago
- Manifold-Mixup implementation for fastai V1☆19Oct 1, 2020Updated 5 years ago
- Lightweight Adapting for Black-Box Large Language Models☆26Feb 15, 2024Updated 2 years ago
- Official repository for the paper "SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention"☆101Sep 30, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆17Mar 18, 2025Updated last year
- Enhancing the convergence speed by 2x and improving the training success of Physics-Informed Neural Networks (PINNs).☆13Oct 14, 2024Updated last year
- [ICLR 2026] SERE: Similarity-Based Expert Re-routing for Efficient Batch Decoding in MoE Models☆19Feb 4, 2026Updated 6 months ago
- TACOS: [T]opology-[A]ware [Co]llective Algorithm [S]ynthesizer for Distributed Machine Learning☆37Jun 13, 2025Updated last year
- COCCL: Compression and precision co-aware collective communication library☆38Jul 20, 2026Updated 3 weeks ago
- Acceleration codes for the Ozaki-scheme on integer matrix multiplication units.☆27Dec 10, 2025Updated 8 months ago
- Incremental bundler and build system optimized for JavaScript and TypeScript, written in Rust – including Turborepo and Turbopack.☆10Nov 2, 2022Updated 3 years ago
- Deep Learning CNN using FastAI for the Stanford MRNet Knee MRI diagnosis challenge☆16May 18, 2019Updated 7 years ago
- ☆34Mar 13, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official implementation of CrossPipe: Towards Optimal Pipeline Schedules for Cross-Datacenter Training (ATC '25), built on top of Megatro…☆17Jul 6, 2025Updated last year
- ☆15Oct 2, 2025Updated 10 months ago
- A C/C++ header file that converts Intel SSE intrinsics to MIPS/MIPS64 MSA intrinsics.☆10Nov 16, 2021Updated 4 years ago
- SRSA: Skill Retrieval and Adaptation for Robotic Assembly Tasks☆18Mar 25, 2026Updated 4 months ago
- Cross-GPU KV Cache Marketplace☆26Nov 12, 2025Updated 8 months ago
- Advancing the frontier of efficient AI☆68Jul 10, 2026Updated last month
- ☆23Jan 23, 2024Updated 2 years ago