Experiments evaluating preemption on the NVIDIA Pascal architecture
☆16Nov 10, 2016Updated 9 years ago
Alternatives and similar repositories for CUDA-preemption
Users that are interested in CUDA-preemption are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Aug 9, 2022Updated 4 years ago
- An Open Source Kepler GPU Assembler☆23Jan 23, 2017Updated 9 years ago
- Efficient CUDA Stream Compaction Library☆34Jun 9, 2023Updated 3 years ago
- ☆28Oct 26, 2019Updated 6 years ago
- Spack package repository maintained by Student Cluster Competition Team @ Sun Yat-sen University.☆16Aug 20, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Cinder port of https://github.com/gangliao/Order-Independent-Transparency-GPU☆15Sep 22, 2018Updated 7 years ago
- The most complete C/C++ snippets extension for VS Code☆19Jun 6, 2021Updated 5 years ago
- Use tensor core to calculate back-to-back HGEMM (half-precision general matrix multiplication) with MMA PTX instruction.☆13Nov 3, 2023Updated 2 years ago
- Third party assembler and GEMM library for NVIDIA Kepler GPU☆86Oct 8, 2019Updated 6 years ago
- Artifacts for SOSP'19 paper Optimizing Deep Learning Computation with Automatic Generation of Graph Substitutions☆21Apr 15, 2022Updated 4 years ago
- ☆11Jan 26, 2016Updated 10 years ago
- Convert CUDA programs from float data type to half or half2 with SIMDization☆19May 28, 2019Updated 7 years ago
- assembler for NVIDIA FERMI. Imported from Google Code☆78Mar 22, 2015Updated 11 years ago
- Torch Distributed Experimental☆117Aug 5, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A header-only C++17 library implementing a simple concurent lock-free memory pool☆27Feb 9, 2021Updated 5 years ago
- ☆44Apr 3, 2022Updated 4 years ago
- Pure Rust implementation of LZ4 compression and decompression as a library☆17Jun 9, 2020Updated 6 years ago
- Several common methods of matrix multiplication are implemented on CPU and Nvidia GPU using C++11 and CUDA.☆14Feb 8, 2023Updated 3 years ago
- Performance of the C++ interface of flash attention and flash attention v2 in large language model (LLM) inference scenarios.☆45Feb 27, 2025Updated last year
- CUPTI GPU Profiler☆39Feb 26, 2019Updated 7 years ago
- ☆18Mar 12, 2025Updated last year
- Polyhedral Extraction Tool (source repository: http://repo.or.cz/w/pet.git)☆42Jul 22, 2022Updated 4 years ago
- ☆20Aug 26, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆25Mar 31, 2022Updated 4 years ago
- experimental port of nervana neon kernels in OpenCL☆11Jul 24, 2016Updated 10 years ago
- Regal for OpenGL☆11Dec 2, 2019Updated 6 years ago
- ☆23Feb 18, 2025Updated last year
- eRPC library for Rust☆14Jan 16, 2020Updated 6 years ago
- A tool for examining GPU scheduling behavior.☆97Aug 3, 2026Updated last month
- Protecting Real-Time GPU Kernels on Integrated CPU-GPU SoC Platforms☆12Apr 9, 2018Updated 8 years ago
- Density Constrained Reinforcement Learning☆12Mar 24, 2023Updated 3 years ago
- Software-based rasterization library☆11Jan 30, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆15Sep 19, 2024Updated last year
- ☆85Dec 2, 2022Updated 3 years ago
- Jaguar port of Doom. Not mine☆11Jan 31, 2021Updated 5 years ago
- pwning challenge with a minimal hypervisor on apple hypervisor framework☆13May 13, 2019Updated 7 years ago
- Gave a talk on Vectorized emulation at Recon Montreal 2019, here are the slides☆19Jun 28, 2019Updated 7 years ago
- ☆11Mar 28, 2023Updated 3 years ago
- ☆10Oct 30, 2021Updated 4 years ago