Enabling on-the-fly manipulations with LLVM IR code of CUDA sources
☆125Apr 18, 2025Updated last year
Alternatives and similar repositories for nvcc-llvm-ir
Users that are interested in nvcc-llvm-ir are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆74Jun 29, 2023Updated 3 years ago
- Flexible GPGPU instrumentation☆92Oct 10, 2019Updated 6 years ago
- Scalable GPU Kernel Fission/Fusion Transformation for Memory-Bound Kernels☆14Aug 26, 2015Updated 11 years ago
- Fault injector for GPUs based on the LLFI Fault Injection Tool☆19May 4, 2018Updated 8 years ago
- The translator that supports translating NVPTX to SPIR-V. This translator is modified from LLVM-SPIR-V Translator.☆45Oct 25, 2021Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A repository where GPU applications are aggregated using a common build flow that supports multiple CUDA versions.☆96Sep 21, 2026Updated 2 weeks ago
- ☆45Apr 3, 2022Updated 4 years ago
- A GPU FP32 computation method with Tensor Cores.☆27Dec 8, 2025Updated 10 months ago
- Python bindings for libNVVM☆39Apr 3, 2014Updated 12 years ago
- ngAP's artifact for ASPLOS'24☆25Jul 29, 2025Updated last year
- An llvm pass for counting global uncoalesced acceses for cuda code via dynamic analysis.☆14Nov 17, 2018Updated 7 years ago
- A framework for pipelined computing on GPU☆30Jul 17, 2019Updated 7 years ago
- A GPU benchmark suite for assessing on-chip GPU memory bandwidth☆113Aug 12, 2017Updated 9 years ago
- HeteroSync is a benchmark suite for performing fine-grained synchronization on tightly coupled GPUs☆32Sep 19, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A set of tools for visualizing and inspecting LLVM bitcode modules☆31Jan 4, 2015Updated 11 years ago
- GPGPU-Sim provides a detailed simulation model of contemporary NVIDIA GPUs running CUDA and/or OpenCL workloads. It includes support for…☆1,741Oct 3, 2026Updated last week
- A intelligent matrix format designer for SpMV☆10Oct 10, 2023Updated 3 years ago
- Fast GPU based tensor core reductions☆12Jan 13, 2023Updated 3 years ago
- ☆10May 12, 2022Updated 4 years ago
- Assembler for NVIDIA Volta and Turing GPUs☆248Jan 13, 2022Updated 4 years ago
- outline and links for PLDI 2022 tutorial☆17Jun 13, 2022Updated 4 years ago
- LLVM IR CMake utils for bitcode file manipulation by opt and friends☆75Dec 7, 2024Updated last year
- ☆13Dec 9, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This is the top-level repository for the Accel-Sim framework.☆705Aug 26, 2026Updated last month
- A tool for examining GPU scheduling behavior.☆97Aug 3, 2026Updated 2 months ago
- GPUOCelot: A dynamic compilation framework for PTX☆289Jul 31, 2023Updated 3 years ago
- ☆19Oct 3, 2022Updated 4 years ago
- ☆85Nov 16, 2020Updated 5 years ago
- Rebuild YatSenOS On RISC-V 64.☆23Jan 6, 2022Updated 4 years ago
- Official BOLT Repository☆34Aug 16, 2024Updated 2 years ago
- VASim is a virtual homogeneous non-deterministic finite automata automata simulator and transformation tool. VASim can parse, transform, …☆36May 17, 2024Updated 2 years ago
- GPUOcelot: A dynamic compilation framework for PTX☆238Feb 9, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- CUDAAdvisor: a GPU profiling tool☆53Aug 24, 2018Updated 8 years ago
- Third party assembler and GEMM library for NVIDIA Kepler GPU☆87Oct 8, 2019Updated 7 years ago
- Automata Benchmark Suite☆23Oct 23, 2023Updated 2 years ago
- ☆160Dec 26, 2024Updated last year
- Assembler and Decompiler for NVIDIA (Maxwell Pascal Volta Turing Ampere) GPUs.☆98Feb 23, 2023Updated 3 years ago
- ☆16Nov 14, 2023Updated 2 years ago
- An implementation of HPL-AI Mixed-Precision Benchmark based on hpl-2.3☆30May 30, 2021Updated 5 years ago