☆238Aug 2, 2024Updated last year
Alternatives and similar repositories for Programming-Massively-Parallel-Processors
Users that are interested in Programming-Massively-Parallel-Processors are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Examples and exercises from the book Programming Massively Parallel Processors - A Hands-on Approach. David B. Kirk and Wen-mei W. Hwu (T…☆79Jan 21, 2021Updated 5 years ago
- Solution of Programming Massively Parallel Processors☆51Jan 15, 2024Updated 2 years ago
- Complete solutions to the Programming Massively Parallel Processors Edition 4☆816Jun 18, 2025Updated last year
- A parser for PTX 6.5☆13Jun 19, 2023Updated 3 years ago
- Material for gpu-mode lectures☆6,355Jun 15, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Repository for compilation and cycle-accurate simulator for scale-out systolic arrays☆16Jan 4, 2023Updated 3 years ago
- Create cohorts from databases utilizing the OMOP CDM☆10May 19, 2025Updated last year
- CUDA 6大并行计算模式 代码与笔记☆63Jul 30, 2020Updated 5 years ago
- Open source RTL implementation of Tensor Core, Sparse Tensor Core, BitWave and SparSynergy in the article: "SparSynergy: Unlocking Flexib…☆26Mar 29, 2025Updated last year
- USB-to-PS2 mouse controller for FPGAs written in Verilog. Performs clock division, signal sampling, processing, error checking, and valid…☆17Feb 26, 2022Updated 4 years ago
- Allen-Cahn Equation☆16Feb 20, 2023Updated 3 years ago
- Code base and slides for ECE408:Applied Parallel Programming On GPU.☆147Jul 2, 2021Updated 5 years ago
- Dissecting NVIDIA GPU Architecture☆126Jul 11, 2022Updated 4 years ago
- Fast CUDA matrix multiplication from scratch☆1,262Sep 2, 2025Updated 10 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Step-by-step optimization of CUDA SGEMM☆486Mar 30, 2022Updated 4 years ago
- ezDPS: An Efficient and Zero-Knowledge Machine Learning Inference Pipeline☆21Jul 14, 2023Updated 3 years ago
- Optimized Parallel Tiled Approach to perform Matrix Multiplication by taking advantage of the lower latency, higher bandwidth shared memo…☆17Sep 24, 2017Updated 8 years ago
- ☆14Mar 8, 2025Updated last year
- ☆49Apr 15, 2024Updated 2 years ago
- learn TensorRT from scratch🥰☆18Sep 29, 2024Updated last year
- CUDA solutions for the lab assignments in the UIUC-ECE408 Applied Parallel Programming course.☆22Apr 18, 2023Updated 3 years ago
- Awesome code, projects, books, etc. related to CUDA☆38Jun 2, 2026Updated last month
- Learn LLVM 17, published by Packt☆217Apr 22, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆19Jul 9, 2026Updated 2 weeks ago
- Acceleration codes for the Ozaki-scheme on integer matrix multiplication units.☆26Dec 10, 2025Updated 7 months ago
- 📚LeetCUDA: Modern CUDA Learn Notes with PyTorch for Beginners🐑, 200+ CUDA Kernels, Tensor Cores, HGEMM, FA-2 MMA.🎉☆11,631Updated this week
- GPU programming related news and material links☆2,239Jun 15, 2026Updated last month
- GPT2 in handwritten PTX☆15Jun 29, 2025Updated last year
- ☆17Mar 26, 2025Updated last year
- Single-header C++ implementation of a Z-order octree data structure☆11May 3, 2021Updated 5 years ago
- ☆10Dec 15, 2023Updated 2 years ago
- Tracking acceptance rates at global CS/AI conferences. This repository contains metadata for building the OpenAccept main site.☆21Jul 11, 2026Updated 2 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆193May 7, 2025Updated last year
- RISCV C and Triton AI-Benchmark☆26Jan 28, 2026Updated 5 months ago
- Jumpstart your custom DNN accelerator today. This project holds scripts to build and start containers that can compile binaries to the ze…☆10Jun 17, 2020Updated 6 years ago
- ☆15May 8, 2025Updated last year
- LeetGPU Solutions☆123Oct 9, 2025Updated 9 months ago
- Samples for CUDA Developers which demonstrates features in CUDA Toolkit☆9,421May 27, 2026Updated last month
- ☆135May 29, 2025Updated last year