AMD-SHARK Inference Modeling and Serving
☆70Jul 9, 2026Updated 2 months ago
Alternatives and similar repositories for amd-shark-ai
Users that are interested in amd-shark-ai are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- IREE's PyTorch Frontend, based on Torch Dynamo.☆110Jul 20, 2026Updated last month
- Unified compiler/runtime for interfacing with PyTorch Dynamo.☆100Jul 2, 2026Updated 2 months ago
- C++ Graph API and JIT Engine powered by IREE☆25Jul 1, 2026Updated 2 months ago
- Row-wise block scaling for fp8 quantization matrix multiplication. Solution to GPU mode AMD challenge.☆19Feb 9, 2026Updated 7 months ago
- IREE plugin repository for the AMD AIE accelerator☆134Sep 7, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Generator for MLIR files from known front-ends☆17Oct 31, 2023Updated 2 years ago
- MLIR-based toolkit targeting intel heterogeneous hardware☆54Jun 26, 2026Updated 2 months ago
- MLIR based Tiny Graph Compiler [dev-stage]☆20Nov 22, 2024Updated last year
- MAD (Model Automation and Dashboarding)☆43Updated this week
- Tensor Compute Primitives: Mid-level Intermediate Representation for Machine Learning Programs☆14Jun 10, 2025Updated last year
- ☆19Jun 6, 2025Updated last year
- ☆17Updated this week
- ARIES: An Agile MLIR-Based Compilation Flow for Reconfigurable Devices with AI Engines (FPGA 2025 Best Paper Nominee)☆67Mar 8, 2026Updated 6 months ago
- ☆13Apr 15, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The goal of the OSSCI Fleet is to provide a central mechanism to enable test automation, batch job scheduling, and developer access to a …☆13Apr 28, 2026Updated 4 months ago
- HRX: Hip Runtime Extended☆45Updated this week
- ☆15Nov 24, 2025Updated 9 months ago
- ☆18May 28, 2024Updated 2 years ago
- C/C++ Guide☆23Jan 3, 2022Updated 4 years ago
- Simply log all kernel durations☆18Aug 26, 2026Updated 3 weeks ago
- ☆149Updated this week
- Torq compiler sources☆56Jul 28, 2026Updated last month
- Regular expression derivatives and DFAs☆12Dec 26, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Artifact evaluation of PLDI'24 paper "Allo: A Programming Model for Composable Accelerator Design"☆35Apr 11, 2024Updated 2 years ago
- A lightweight MLIR Python frontend with support for PyTorch☆28Sep 3, 2024Updated 2 years ago
- console version of StarDict formerly used by koreader; fork of https://github.com/Dushistov/sdcv☆10Aug 29, 2017Updated 9 years ago
- examples of using ip rules and multiple routing tables to solve real-world problems☆14Aug 20, 2010Updated 16 years ago
- Discord Bot that moderates Harassment along with Toxicity and Depressive Behavior☆12May 3, 2025Updated last year
- Tenstorrent MLIR compiler☆306Updated this week
- MLIR-based partitioning system☆210Updated this week
- Triton to TVM transpiler.☆24Oct 14, 2024Updated last year
- AMD-SHARK Studio -- Web UI for SHARK+IREE High Performance Machine Learning Distribution☆1,451Dec 15, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- The missing pieces (as far as boilerplate reduction goes) of the upstream MLIR python bindings.☆118Mar 4, 2026Updated 6 months ago
- Allo Accelerator Design and Programming Framework (PLDI'24)☆411Sep 8, 2026Updated last week
- Examples and training code for Machine Learning samples that can be run on various Edge devices☆10Jan 8, 2025Updated last year
- IREE C++ Template☆17Jul 30, 2024Updated 2 years ago
- WaferLLM: Large Language Model Inference at Wafer Scale☆121Jun 12, 2026Updated 3 months ago
- The Torch-MLIR project aims to provide first class support from the PyTorch ecosystem to the MLIR ecosystem.☆1,919Updated this week
- Fork of LLVM to support AMD AIEngine processors☆211Updated this week