Benchmark PyTorch Custom Operators
☆14Jul 6, 2023Updated 3 years ago
Alternatives and similar repositories for epoi
Users that are interested in epoi are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆143Jan 30, 2025Updated last year
- ☆23Aug 21, 2025Updated 11 months ago
- Artifacts for SOSP'19 paper Optimizing Deep Learning Computation with Automatic Generation of Graph Substitutions☆21Apr 15, 2022Updated 4 years ago
- A schedule language for large model training☆153Aug 21, 2025Updated 11 months ago
- DietCode Code Release☆65Jul 21, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- SparseTIR: Sparse Tensor Compiler for Deep Learning☆145Mar 31, 2023Updated 3 years ago
- PyTorch compilation tutorial covering TorchScript, torch.fx, and Slapo☆17Mar 13, 2023Updated 3 years ago
- A language and compiler for irregular tensor programs.☆152Updated this week
- Code for Incorporating Relevance Feedback for Information-Seeking Retrieval using Few-Shot Document Re-Ranking, EMNLP 2022, https://aclan…☆14Mar 30, 2026Updated 4 months ago
- HeteroHalide: From Image Processing DSL to Efficient FPGA Acceleration☆15Sep 14, 2020Updated 5 years ago
- An experimental ahead of time compiler for Relay.☆49Apr 21, 2020Updated 6 years ago
- ☆11Apr 5, 2021Updated 5 years ago
- ☆13Jan 7, 2025Updated last year
- A self-contained version of the tutorial which can be easily cloned and viewed by others.☆24Jun 24, 2019Updated 7 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- HeteroCL-MLIR dialect for accelerator design☆42Sep 18, 2024Updated last year
- ACM Class 2017 Computer Architecture☆10Jan 11, 2018Updated 8 years ago
- ☆21Dec 27, 2019Updated 6 years ago
- PET: Optimizing Tensor Programs with Partially Equivalent Transformations and Automated Corrections☆126Jun 23, 2022Updated 4 years ago
- ☆20Sep 28, 2020Updated 5 years ago
- ☆193Mar 28, 2023Updated 3 years ago
- Compact and Agent-Native MoE Training System☆328Jul 31, 2026Updated last week
- APPy (Annotated Parallelism for Python) enables users to annotate loops and tensor expressions in Python with compiler directives akin to…☆28Mar 22, 2026Updated 4 months ago
- Automatic Mapping Generation, Verification, and Exploration for ISA-based Spatial Accelerators☆125Oct 26, 2022Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- DOSA: Differentiable Model-Based One-Loop Search for DNN Accelerators☆20Oct 10, 2024Updated last year
- A home for the final text of all TVM RFCs.☆111Sep 24, 2024Updated last year
- ☆249Jul 27, 2025Updated last year
- ☆23Dec 11, 2024Updated last year
- 🔮 Execution time predictions for deep neural network training iterations across different GPUs.☆14Dec 16, 2024Updated last year
- C++ "borrowing" smart pointer.☆10May 13, 2022Updated 4 years ago
- Chameleon: Adaptive Code Optimization for Expedited Deep Neural Network Compilation☆26Nov 7, 2019Updated 6 years ago
- DASS HLS Compiler☆31Oct 4, 2023Updated 2 years ago
- Re-implementation of the TASO compiler using equality saturation☆142Jun 28, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆28Jun 18, 2026Updated last month
- A repository to perform self-instruct with a model on HF Hub☆32Sep 29, 2023Updated 2 years ago
- ☆15Oct 26, 2022Updated 3 years ago
- Polyhedral High-Level Synthesis in MLIR☆35Mar 17, 2023Updated 3 years ago
- ☆11Sep 14, 2020Updated 5 years ago
- Supplemental materials for The ASPLOS 2025 / EuroSys 2025 Contest on Intra-Operator Parallelism for Distributed Deep Learning☆25May 12, 2025Updated last year
- ☆10Sep 16, 2021Updated 4 years ago