The quantitative performance comparison among DL compilers on CNN models.
☆73Aug 27, 2020Updated 6 years ago
Alternatives and similar repositories for dlcompiler-comparison
Users that are interested in dlcompiler-comparison are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SparseTIR: Sparse Tensor Compiler for Deep Learning☆145Mar 31, 2023Updated 3 years ago
- [MLSys 2021] IOS: Inter-Operator Scheduler for CNN Acceleration☆201Apr 27, 2022Updated 4 years ago
- examples for tvm schedule API☆101Jun 12, 2023Updated 3 years ago
- This is the implementation for paper: AdaTune: Adaptive Tensor Program CompilationMade Efficient (NeurIPS 2020).☆14May 16, 2021Updated 5 years ago
- Neural Network Acceleration such as ASIC, FPGA, GPU, and PIM☆54Apr 13, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Study Group of Deep Learning Compiler☆173Jan 15, 2023Updated 3 years ago
- TVM learning and research☆13Jan 8, 2021Updated 5 years ago
- Dive into Deep Learning Compiler☆650Jun 19, 2022Updated 4 years ago
- tophub autotvm log collections☆68Dec 30, 2022Updated 3 years ago
- An MLIR-based compiler framework bridges DSLs (domain-specific languages) to DSAs (domain-specific architectures).☆762Updated this week
- A list of awesome compiler projects and papers for tensor computation and deep learning.☆2,787Oct 19, 2024Updated last year
- A flexible and efficient deep neural network (DNN) compiler that generates high-performance executable from a DNN model description.☆1,002Sep 19, 2024Updated 2 years ago
- An open-sourced PyTorch library for developing energy efficient multiplication-less models and applications.☆14Feb 3, 2025Updated last year
- ☆23Dec 8, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Provide Docker build sequences of PyTorch for various environments.☆16May 26, 2021Updated 5 years ago
- [ICML 2022] ShiftAddNAS: Hardware-Inspired Search for More Accurate and Efficient Neural Networks☆15May 18, 2022Updated 4 years ago
- Automated DNN generation for fuzz testing and more☆151Jan 14, 2025Updated last year
- ParaDnn: A systematic performance analysis methodology for deep learning.☆40Mar 30, 2020Updated 6 years ago
- Representation and Reference Lowering of ONNX Models in MLIR Compiler Infrastructure☆1,057Updated this week
- ☆16Mar 10, 2024Updated 2 years ago
- Must read research papers and links to tools and datasets that are related to using machine learning for compilers and systems optimisati…☆1,700Aug 26, 2026Updated last month
- compiler learning resources collect.☆2,779May 20, 2026Updated 4 months ago
- The Tensor Algebra SuperOptimizer for Deep Learning☆744Jan 26, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- TVM stack: exploring the incredible explosion of deep-learning frameworks and how to bring them together☆65May 22, 2018Updated 8 years ago
- code reading for tvm☆75Jan 20, 2022Updated 4 years ago
- notes on reading tensorflow source code☆13Aug 18, 2018Updated 8 years ago
- [NeurIPS 2021] "Drawing Robust Scratch Tickets: Subnetworks with Inborn Robustness Are Found within Randomly Initialized Networks" by Yon…☆13Feb 13, 2022Updated 4 years ago
- PET: Optimizing Tensor Programs with Partially Equivalent Transformations and Automated Corrections☆126Jun 23, 2022Updated 4 years ago
- This repository is a meta package to provide Samsung OneMCC (Memory-Centric Computing) infrastructure.☆35Nov 26, 2025Updated 10 months ago
- Artifacts of EVT ASPLOS'24☆29Mar 6, 2024Updated 2 years ago
- CNN Accelerator in Frequency Domain☆12Feb 22, 2020Updated 6 years ago
- BladeDISC is an end-to-end DynamIc Shape Compiler project for machine learning workloads.☆932Dec 30, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ECCV 2022] SuperTickets: Drawing Task-Agnostic Lottery Tickets from Supernets via Jointly Architecture Searching and Parameter Pruning☆20Jul 7, 2022Updated 4 years ago
- Binary neural networks developed by Huawei Noah's Ark Lab☆29Feb 19, 2021Updated 5 years ago
- Social Disatancing Monitor using yolov3 and DPU HW acceleration for Xilinx adaptive computing challenge 2020☆12Feb 17, 2023Updated 3 years ago
- research, experimentation and implementation of hardware-agnostic accelerated DL framework☆41Aug 7, 2026Updated last month
- ☆193Mar 28, 2023Updated 3 years ago
- LLM inference in C/C++☆22Oct 22, 2025Updated 11 months ago
- GSDMM: Short text clustering (Rust implementation)☆24Apr 26, 2023Updated 3 years ago