AMD related optimizations for transformer models
☆101Oct 6, 2026Updated this week
Alternatives and similar repositories for optimum-amd
Users that are interested in optimum-amd are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AMD Ryzen™ AI Software includes the tools and runtime libraries for optimizing and deploying AI inference on AMD Ryzen™ AI powered PCs.☆889Aug 18, 2026Updated last month
- 🏋️ A unified multi-backend utility for benchmarking Transformers, Timm, PEFT, Diffusers and Sentence-Transformers with full support of O…☆340Sep 29, 2026Updated last week
- Github action to connect to tailscale☆22Aug 28, 2026Updated last month
- Accelerated inference of 🤗 models using FuriosaAI NPU chips.☆27Sep 24, 2026Updated 2 weeks ago
- Google TPU optimizations for transformers models☆136Jan 23, 2026Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Fast and memory-efficient exact attention☆240Aug 12, 2026Updated last month
- 🐉 Revolutionary NPU framework for Linux | 24,988 FPS face recognition | AMD XDNA support | World's first complete NPU stack☆43Aug 7, 2025Updated last year
- A pytorch quantization backend for optimum☆1,052Sep 29, 2026Updated last week
- Development repository for the Triton language and compiler☆146Oct 1, 2026Updated last week
- The Riallto Open Source Project from AMD☆87Apr 10, 2025Updated last year
- Pixel Parsing. A reproduction of OCR-free end-to-end document understanding models with open data☆25Jul 30, 2024Updated 2 years ago
- AMD iGPU AI Setup and Speed Test - GPD Pocket 4 - Linux + ROCm + Vulkan + AgentMake AI☆28Jun 17, 2026Updated 3 months ago
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆27Sep 28, 2026Updated last week
- Super fast FP32 matrix multiplication on RDNA3☆92Mar 30, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Easy and lightning fast training of 🤗 Transformers on Habana Gaudi processor (HPU)☆213Updated this week
- Torch-MIGraphX integrates AMD's graph inference engine with the PyTorch ecosystem.☆22Sep 29, 2026Updated last week
- Binary Neural Network-based COVID-19 Face-Mask Wear and Positioning Predictor on Edge Devices☆12Jul 1, 2021Updated 5 years ago
- MAD (Model Automation and Dashboarding)☆43Sep 24, 2026Updated 2 weeks ago
- Ghidra's development plugins, scripts, contributing. Presentation☆13Aug 10, 2020Updated 6 years ago
- 🚀 Accelerate inference and training of 🤗 Transformers, Diffusers, TIMM and Sentence Transformers with easy to use hardware optimization…☆3,498Updated this week
- No-code CLI designed for accelerating ONNX workflows☆247Jul 1, 2026Updated 3 months ago
- a quick and dirty little program to convert Apple CoreML model to ANE hwx file☆51Oct 2, 2026Updated last week
- ☆23Mar 6, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Manage scalable open LLM inference endpoints in Slurm clusters☆296Jul 11, 2024Updated 2 years ago
- Training code and resources for AMD-135M language models on AMD GPUs.☆195Jun 23, 2026Updated 3 months ago
- 👷 Build compute kernels☆213Apr 6, 2026Updated 6 months ago
- ☆13May 26, 2021Updated 5 years ago
- Fork of LLVM to support AMD AIEngine processors☆212Oct 2, 2026Updated last week
- Easy and Efficient Quantization for Transformers☆205Mar 25, 2026Updated 6 months ago
- ☆24Sep 17, 2026Updated 3 weeks ago
- The HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm☆1,370Updated this week
- ☆181Sep 28, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- CMake modules used within the ROCm libraries☆78Oct 2, 2026Updated last week
- Efficient implementation of DeepSeek Ops (Blockwise FP8 GEMM, MoE, and MLA) and Distributed Ops (All2All, GEMM-RS, AG-GEMM) for AMD Insti…☆80Feb 11, 2026Updated 7 months ago
- ☆24Oct 2, 2025Updated last year
- ☆209May 5, 2025Updated last year
- React-Autosuggest for Plotly Dash with Elasticsearch integration.☆12Dec 3, 2022Updated 3 years ago
- Run LLMs on AMD Ryzen™ AI NPUs in minutes; purpose-built and deeply optimized for the AMD NPUs.☆1,944Updated this week
- ANE (Apple Neural Engine) CostModel profiler for CoreML models☆35Apr 9, 2026Updated 6 months ago