☆43Apr 25, 2024Updated 2 years ago
Alternatives and similar repositories for tlp
Users that are interested in tlp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆101Nov 4, 2022Updated 3 years ago
- ☆49Jul 13, 2024Updated 2 years ago
- Chameleon: Adaptive Code Optimization for Expedited Deep Neural Network Compilation☆26Nov 7, 2019Updated 6 years ago
- Automatic Mapping Generation, Verification, and Exploration for ISA-based Spatial Accelerators☆125Oct 26, 2022Updated 3 years ago
- Automatic Schedule Exploration and Optimization Framework for Tensor Computations☆182Apr 25, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Optimize tensor program fast with Felix, a gradient descent autotuner.☆33Mar 5, 2026Updated 6 months ago
- ☆54Dec 13, 2022Updated 3 years ago
- ☆13Jan 7, 2025Updated last year
- ☆17Dec 8, 2023Updated 2 years ago
- ☆11Sep 14, 2020Updated 5 years ago
- Heron: Automatically Constrained High-Performance Library Generation for Deep Learning Accelerators☆24Jan 30, 2024Updated 2 years ago
- Official implementation of Acc-SpMM: Accelerating General-purpose Sparse Matrix-Matrix Multiplication with GPU Tensor Cores.☆38Nov 13, 2025Updated 9 months ago
- [NeurIPS 2024] Search for Efficient LLMs☆16Jan 16, 2025Updated last year
- A home for the final text of all TVM RFCs.☆111Sep 24, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆227Nov 22, 2024Updated last year
- ASPLOS'24: Optimal Kernel Orchestration for Tensor Programs with Korch☆41Mar 27, 2025Updated last year
- SparseTIR: Sparse Tensor Compiler for Deep Learning☆145Mar 31, 2023Updated 3 years ago
- PIM-ML is a benchmark for training machine learning algorithms on the UPMEM architecture, which is the first publicly-available real-worl…☆30Jan 7, 2025Updated last year
- A shader system built using staged metaprogramming☆15Jul 9, 2022Updated 4 years ago
- ☆42Sep 8, 2023Updated 2 years ago
- Mille Crepe Bench: layer-wise performance analysis for deep learning frameworks.☆18Oct 22, 2019Updated 6 years ago
- examples for tvm schedule API☆101Jun 12, 2023Updated 3 years ago
- 微信Ipad协议golang版本,基于grpc的实现策略。这套代码需要通过gprc服务端组包解包才可以正常使用☆13Jul 8, 2019Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This is a list of awesome edgeAI inference related papers.☆98Dec 21, 2023Updated 2 years ago
- An open-source efficient deep learning framework/compiler, written in python.☆743Sep 4, 2025Updated last year
- A distributed in-memory store for temporal knowledge graphs☆10Mar 20, 2024Updated 2 years ago
- A Cluster-Wide Model Manager to Accelerate DNN Training via Automated Training Warmup☆36Jan 9, 2023Updated 3 years ago
- Generative Models for Image Captioning☆10Jun 7, 2017Updated 9 years ago
- Draw emoji on USTC logo.☆10Sep 15, 2017Updated 8 years ago
- Multi-branch model for concurrent execution☆18Jun 27, 2023Updated 3 years ago
- Automatic Differentiation for Tensor Algebras☆28May 8, 2018Updated 8 years ago
- Skeletonide is a parallel implementation of Zhang-Suen morphological thinning algorithm written in Halide-lang. Use it for fast skeletoni…☆14Oct 21, 2020Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Horizontal Fusion☆24Jan 7, 2022Updated 4 years ago
- A list of awesome compiler projects and papers for tensor computation and deep learning.☆2,778Oct 19, 2024Updated last year
- A self-contained version of the tutorial which can be easily cloned and viewed by others.☆24Jun 24, 2019Updated 7 years ago
- DeepSeek-V3.2-Exp DSA Warmup Lightning Indexer training operator based on tilelang☆52Nov 19, 2025Updated 9 months ago
- Samoyeds: Accelerating MoE Models with Structured Sparsity Leveraging Sparse Tensor Cores (EuroSys'25)☆16Jul 17, 2025Updated last year
- Alex Graves' Adaptive Computation Time in PyTorch☆14Jan 9, 2018Updated 8 years ago
- Code released to accompany the ISCA paper: "T4: Compiling Sequential Code for Effective Speculative Parallelization in Hardware"☆29Feb 18, 2022Updated 4 years ago