Samoyeds: Accelerating MoE Models with Structured Sparsity Leveraging Sparse Tensor Cores (EuroSys'25)
☆16Jul 17, 2025Updated last year
Alternatives and similar repositories for Samoyeds
Users that are interested in Samoyeds are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆46Jun 19, 2024Updated 2 years ago
- Code for High Performance Unstructured SpMM Computation Using Tensor Cores☆35Nov 3, 2024Updated last year
- Horizontal Fusion☆24Jan 7, 2022Updated 4 years ago
- ☆12May 19, 2025Updated last year
- SpInfer: Leveraging Low-Level Sparsity for Efficient Large Language Model Inference on GPUs☆70Mar 25, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A Vectorized N:M Format for Unleashing the Power of Sparse Tensor Cores☆62Nov 24, 2023Updated 2 years ago
- ☆17Dec 5, 2024Updated last year
- Source code of the SC '23 paper: "DASP: Specific Dense Matrix Multiply-Accumulate Units Accelerated General Sparse Matrix-Vector Multipli…☆29Jun 18, 2024Updated 2 years ago
- ☆27Apr 13, 2025Updated last year
- Cleanlab Vizzy: illustrating the core ideas behind the Cleanlab algorithm☆16Apr 19, 2023Updated 3 years ago
- 关于AI,ML,DA,DV等的几个经典案例,包括堵车模拟(NagelSchreckenberg)、蒙特卡洛排队问题(Monte Carlo Queuing Problem)、人脸识别(RecognitionFace)、遗传算法推断图像(IconGenetic)☆10Oct 14, 2018Updated 7 years ago
- The goal of this design is to use the PYNQ-Z2 development board to design a general convolution neural network accelerator. And through r…☆12Sep 30, 2020Updated 5 years ago
- ☆13Dec 19, 2025Updated 9 months ago
- This repository will soon contain all scripts and links to the annotated corpora of Tibetan.☆14Feb 4, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [NeurIPS 2024] Search for Efficient LLMs☆16Jan 16, 2025Updated last year
- ☆10Sep 10, 2026Updated 2 weeks ago
- ☆13Sep 19, 2024Updated 2 years ago
- This is the official repo for "Differentiable Model Scaling using Differentiable Topk"☆12May 16, 2024Updated 2 years ago
- [ASPLOS'25] Towards End-to-End Optimization of LLM-based Applications with Ayo☆77Mar 11, 2026Updated 6 months ago
- ☆16Apr 11, 2025Updated last year
- A LaTeX template provides a beautiful design of class schedule with colorful course blocks.☆14Jul 9, 2026Updated 2 months ago
- vortex particles for simulating smoke in 2d☆17Dec 13, 2021Updated 4 years ago
- Source code repository for ASPLOS '25 paper "Syno: Structured Synthesis for Neural Operators"☆15Aug 31, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- 中科大郑启龙2021年并行程序设计课程实验☆11Jan 15, 2022Updated 4 years ago
- ☆32Mar 24, 2025Updated last year
- Medusa: Accelerating Serverless LLM Inference with Materialization [ASPLOS'25]☆12Nov 8, 2024Updated last year
- ☆13Dec 20, 2025Updated 9 months ago
- Examples showing how to utilize the NVML library for GPU monitoring☆29May 31, 2022Updated 4 years ago
- 上海交通大学软件学院课程《应用系统体系架构》(SE3353)笔记☆11Feb 2, 2024Updated 2 years ago
- ☆14Dec 20, 2023Updated 2 years ago
- ☆14Apr 24, 2024Updated 2 years ago
- [ICML 2025] Adaptive Self-improvement LLM Agentic System for ML Library Development☆17Jan 6, 2026Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation of Acc-SpMM: Accelerating General-purpose Sparse Matrix-Matrix Multiplication with GPU Tensor Cores.☆38Nov 13, 2025Updated 10 months ago
- ☆86Jun 23, 2025Updated last year
- 水源社区 API client☆16Dec 11, 2023Updated 2 years ago
- ☆43Apr 25, 2024Updated 2 years ago
- ☆35Apr 2, 2025Updated last year
- Artifact for USENIX ATC'23: TC-GNN: Bridging Sparse GNN Computation and Dense Tensor Cores on GPUs.☆59Oct 16, 2023Updated 2 years ago
- Flash-LLM: Enabling Cost-Effective and Highly-Efficient Large Generative Model Inference with Unstructured Sparsity☆247Sep 24, 2023Updated 3 years ago