CUDA by practice
☆140Jan 7, 2020Updated 6 years ago
Alternatives and similar repositories for CUDA_by_practice
Users that are interested in CUDA_by_practice are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19May 17, 2016Updated 10 years ago
- Winograd-based convolution implementation in OpenCL☆29Jan 22, 2017Updated 9 years ago
- Numba GPU tutorial notebooks for PyData Amsterdam 2019☆23Jun 25, 2026Updated 2 months ago
- Tensorflow Operation Wrapper of cppjieba (Chinese Word Segamentation)☆10Oct 21, 2019Updated 6 years ago
- ☆19Apr 6, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Sparse matrix computation library for GPU☆59Jul 12, 2020Updated 6 years ago
- Publish sensor data from iOS device to ROS topic☆16Oct 4, 2016Updated 9 years ago
- 小彭老师推出 SyCL 2020 课程(施工中,日后会在直播中放出)☆15Sep 3, 2023Updated 3 years ago
- A tool for simulating UART through NET☆11Jul 16, 2021Updated 5 years ago
- A brief tutorial for eBPF: Verifier, observability, networking, and security.☆14Sep 19, 2024Updated last year
- Resnet-50 + FPN + Keypoint RCNN☆14Jun 18, 2019Updated 7 years ago
- Source code of the paper "OpSparse: a Highly Optimized Framework for Sparse General Matrix Multiplication on GPUs"☆16Aug 23, 2022Updated 4 years ago
- A simple Mali 6xx/7xx register interface model that doesn't do any rendering.☆13Jan 29, 2016Updated 10 years ago
- ☆14Aug 31, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A visual dataflow programming language for NVIDIA's RAPIDS, based on AlvarBer/Persimmon☆14Jun 1, 2019Updated 7 years ago
- ☆15Oct 14, 2025Updated 10 months ago
- study of cutlass☆22Nov 10, 2024Updated last year
- MAchine Micro Management UTilities☆12Nov 5, 2020Updated 5 years ago
- Effective transpose on Hopper GPU☆29Sep 6, 2025Updated 11 months ago
- ☆37Oct 12, 2024Updated last year
- This repo is "NTHU Parallel Programing" course project.☆10Dec 5, 2017Updated 8 years ago
- A high performance implementation of kmeans algorithm with cuda☆18Sep 7, 2014Updated 11 years ago
- Pipelined Processor which implements RV32i Instruction Set. Also contains pipelined L1 4-way set-associative Instruction Cache, direct-ma…☆15Dec 23, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- UCSD CSE231 Advanced Compiler - LLVM project☆12Mar 28, 2017Updated 9 years ago
- It is an annoying thing of preparing the openCL environment, so I wapper the initialization part of OpenCL and setting parameters for ker…☆16May 16, 2018Updated 8 years ago
- ☆11Mar 7, 2018Updated 8 years ago
- Source code that accompanies The CUDA Handbook.☆598Aug 15, 2026Updated 3 weeks ago
- High-performance integer factorization suite implementing GNFS, MPQS, and QS algorithms with optimized lattice reduction, vectorization, …☆10Mar 25, 2026Updated 5 months ago
- Parallel GMRES (Generalized Minimal Residual) linear solver on GPU platforms☆26Oct 5, 2015Updated 10 years ago
- An ATPG tool using PODEM algorithm in C++ that generates a test to detect any given list of Single-Stuck-at Faults☆11Oct 29, 2017Updated 8 years ago
- Convolutional Neural Network of vgg19 model using Cuda to accelerate☆12Jun 11, 2018Updated 8 years ago
- A desktop pet program.☆11Aug 22, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- AutoParBench is a benchmark framework to evaluate compilers and tools designed to automatically insert OpenMP directives.☆12Nov 6, 2020Updated 5 years ago
- Let's build a better future with code!☆13Jan 20, 2024Updated 2 years ago
- Initial framework source code for CSE240A branch predictor project☆10Nov 30, 2017Updated 8 years ago
- C++ heterogeneous and lock-free containers☆13Sep 5, 2018Updated 7 years ago
- Matrix Multiply-Accumulate with CUDA and WMMA( Tensor Core)☆149Aug 18, 2020Updated 6 years ago
- ☆40Feb 28, 2020Updated 6 years ago
- ☆13Nov 15, 2017Updated 8 years ago