☆75Sep 18, 2026Updated this week
Alternatives and similar repositories for rocprof-compute-viewer
Users that are interested in rocprof-compute-viewer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Apr 10, 2026Updated 5 months ago
- FlyDSL is the Python front‑end of the project: a Flexible Layout Python DSL for expressing tiling, partitioning, data movement, and kerne…☆283Updated this week
- A visualizer for the ROCm Profiler Tools☆30Updated this week
- Super fast FP32 matrix multiplication on RDNA3☆92Mar 30, 2025Updated last year
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆30May 28, 2026Updated 3 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- super repo for rocm libraries☆434Updated this week
- super repo for rocm systems projects☆508Updated this week
- HRX: Hip Runtime Extended☆47Updated this week
- amdgpu example code in hip/asm☆69Aug 10, 2026Updated last month
- A practical guide to high-performance gluon kernel development on AMD GFX9 GPUs.☆51Sep 14, 2026Updated last week
- ☆76Updated this week
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆114Updated this week
- Automating analysis from trace files☆91Updated this week
- A tool for generating information about the matrix multiplication instructions in AMD Radeon™ and AMD Instinct™ accelerators☆144Apr 10, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Generating Efficient AI-Centric Kernels☆180Updated this week
- ☆19Jun 6, 2025Updated last year
- Automated bottleneck detection and solution orchestration☆23Feb 24, 2026Updated 6 months ago
- QuickReduce is a performant all-reduce library designed for AMD ROCm that supports inline compression.☆38Aug 29, 2025Updated last year
- A lightweight triton-based General Matrix Multiplication (GEMM) library.☆68Jul 21, 2026Updated 2 months ago
- AiTer Optimized Model☆184Updated this week
- Wave: Python Domain-Specific Language for High Performance Machine Learning☆60Jun 29, 2026Updated 2 months ago
- HIP backend patch for Numba, the NumPy aware dynamic Python compiler using LLVM.☆22Jul 10, 2026Updated 2 months ago
- Development repository for the Triton language and compiler☆146Sep 14, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Fast and Furious AMD Kernels☆470Updated this week
- AMD lab notes with code examples to demonstrate use of AMD GPUs☆117Jun 28, 2024Updated 2 years ago
- AMD RAD's multi-GPU Triton-based framework for seamless multi-GPU programming☆202Updated this week
- ☆27Mar 5, 2026Updated 6 months ago
- Test suite for C/C++/Fortran compilers developed by Fujitsu☆53Updated this week
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆83Updated this week
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆153May 28, 2026Updated 3 months ago
- ☆23Mar 16, 2026Updated 6 months ago
- A ROCm library for GPU-Initiated IO. This provides support for initiating IO from a ROCm-capable GPU against a range of targets including…☆58Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆30Sep 10, 2026Updated last week
- HERACLES++ is a hydrodynamic code with a moving grid, designed to run on different architectures using Kokkos.☆11Sep 7, 2026Updated 2 weeks ago
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆142Sep 15, 2026Updated last week
- Tools for MPI programmers☆14Sep 21, 2020Updated 6 years ago
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆167May 28, 2026Updated 3 months ago
- Modular RDMA Interface☆181Updated this week
- AMD’s C++ library for accelerating tensor primitives☆49Updated this week