Enabling on-the-fly manipulations with LLVM IR code of CUDA sources
☆125Apr 18, 2025Updated last year
Alternatives and similar repositories for nvcc-llvm-ir
Users that are interested in nvcc-llvm-ir are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆74Jun 29, 2023Updated 3 years ago
- Flexible GPGPU instrumentation☆91Oct 10, 2019Updated 6 years ago
- Scalable GPU Kernel Fission/Fusion Transformation for Memory-Bound Kernels☆14Aug 26, 2015Updated 11 years ago
- Fault injector for GPUs based on the LLFI Fault Injection Tool☆19May 4, 2018Updated 8 years ago
- The translator that supports translating NVPTX to SPIR-V. This translator is modified from LLVM-SPIR-V Translator.☆45Oct 25, 2021Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A repository where GPU applications are aggregated using a common build flow that supports multiple CUDA versions.☆95Apr 14, 2026Updated 4 months ago
- ☆44Apr 3, 2022Updated 4 years ago
- A GPU FP32 computation method with Tensor Cores.☆27Dec 8, 2025Updated 8 months ago
- Python bindings for libNVVM☆39Apr 3, 2014Updated 12 years ago
- ngAP's artifact for ASPLOS'24☆25Jul 29, 2025Updated last year
- An llvm pass for counting global uncoalesced acceses for cuda code via dynamic analysis.☆14Nov 17, 2018Updated 7 years ago
- A framework for pipelined computing on GPU☆30Jul 17, 2019Updated 7 years ago
- A GPU benchmark suite for assessing on-chip GPU memory bandwidth☆113Aug 12, 2017Updated 9 years ago
- HeteroSync is a benchmark suite for performing fine-grained synchronization on tightly coupled GPUs☆32Sep 19, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A set of tools for visualizing and inspecting LLVM bitcode modules☆31Jan 4, 2015Updated 11 years ago
- GPGPU-Sim provides a detailed simulation model of contemporary NVIDIA GPUs running CUDA and/or OpenCL workloads. It includes support for…☆1,704Feb 15, 2025Updated last year
- A intelligent matrix format designer for SpMV☆10Oct 10, 2023Updated 2 years ago
- Fast GPU based tensor core reductions☆12Jan 13, 2023Updated 3 years ago
- ☆25Jun 24, 2022Updated 4 years ago
- ☆10May 12, 2022Updated 4 years ago
- Assembler for NVIDIA Volta and Turing GPUs☆248Jan 13, 2022Updated 4 years ago
- outline and links for PLDI 2022 tutorial☆17Jun 13, 2022Updated 4 years ago
- LLVM IR CMake utils for bitcode file manipulation by opt and friends☆75Dec 7, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆13Dec 9, 2024Updated last year
- A tool for examining GPU scheduling behavior.☆97Aug 3, 2026Updated 3 weeks ago
- GPUOCelot: A dynamic compilation framework for PTX☆289Jul 31, 2023Updated 3 years ago
- ☆19Oct 3, 2022Updated 3 years ago
- ☆83Nov 16, 2020Updated 5 years ago
- PTX-EMU is a simple emulator for CUDA program.☆40Apr 25, 2025Updated last year
- VASim is a virtual homogeneous non-deterministic finite automata automata simulator and transformation tool. VASim can parse, transform, …☆36May 17, 2024Updated 2 years ago
- GPUOcelot: A dynamic compilation framework for PTX☆236Feb 9, 2025Updated last year
- CUDAAdvisor: a GPU profiling tool☆53Aug 24, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Third party assembler and GEMM library for NVIDIA Kepler GPU☆86Oct 8, 2019Updated 6 years ago
- Automata Benchmark Suite☆23Oct 23, 2023Updated 2 years ago
- ☆160Dec 26, 2024Updated last year
- Assembler and Decompiler for NVIDIA (Maxwell Pascal Volta Turing Ampere) GPUs.☆98Feb 23, 2023Updated 3 years ago
- ☆15Nov 14, 2023Updated 2 years ago
- An implementation of HPL-AI Mixed-Precision Benchmark based on hpl-2.3☆30May 30, 2021Updated 5 years ago
- ☆77May 29, 2019Updated 7 years ago