CUDA by practice
☆138Jan 7, 2020Updated 6 years ago
Alternatives and similar repositories for CUDA_by_practice
Users that are interested in CUDA_by_practice are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19May 17, 2016Updated 10 years ago
- ☆19Apr 6, 2024Updated 2 years ago
- Sparse matrix computation library for GPU☆59Jul 12, 2020Updated 6 years ago
- Publish sensor data from iOS device to ROS topic☆16Oct 4, 2016Updated 9 years ago
- 小彭老师推出 SyCL 2020 课程(施工中,日后会在直播中放出)☆15Sep 3, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- An MIPS pipelined processor with hazard detection for the course VE370 (FA2020) at UMJI.☆11Dec 28, 2020Updated 5 years ago
- ☆50Jun 27, 2019Updated 7 years ago
- Sim-to-Real via Sim-to-Sim using fast.ai's U-net☆10Nov 25, 2019Updated 6 years ago
- generative-camouflaged-spam-detector☆11Aug 20, 2020Updated 5 years ago
- Resnet-50 + FPN + Keypoint RCNN☆14Jun 18, 2019Updated 7 years ago
- Source code of the paper "OpSparse: a Highly Optimized Framework for Sparse General Matrix Multiplication on GPUs"☆16Aug 23, 2022Updated 3 years ago
- A simple Mali 6xx/7xx register interface model that doesn't do any rendering.☆13Jan 29, 2016Updated 10 years ago
- study of cutlass☆22Nov 10, 2024Updated last year
- Prediction pipeline to generate prognosis predictors for Ebola Virus Disease☆12Feb 22, 2016Updated 10 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Tensorflow to TensorRT Model Converter☆30Feb 11, 2018Updated 8 years ago
- MAchine Micro Management UTilities☆12Nov 5, 2020Updated 5 years ago
- Effective transpose on Hopper GPU☆29Sep 6, 2025Updated 10 months ago
- North Carolina State University: ECE 745 : Project: LC3 Microcontroller Functional Verification using SystemVerilog☆11Jun 5, 2017Updated 9 years ago
- UCSD CSE240A Project: Branch Predictor☆11Jul 24, 2017Updated 9 years ago
- ☆38Oct 12, 2024Updated last year
- Benchmarking Analysis of Vision Kernels on Embedded CPU, GPU and FPGA☆16Apr 21, 2019Updated 7 years ago
- A high performance implementation of kmeans algorithm with cuda☆18Sep 7, 2014Updated 11 years ago
- Source code that accompanies The CUDA Handbook.☆597Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Parallel GMRES (Generalized Minimal Residual) linear solver on GPU platforms☆26Oct 5, 2015Updated 10 years ago
- I optimized some code for the European Space Agency, achieving significant speedups - using OpenMP, SSE, Eigen and CUDA. This is the resu…☆16Aug 4, 2017Updated 8 years ago
- An ATPG tool using PODEM algorithm in C++ that generates a test to detect any given list of Single-Stuck-at Faults☆11Oct 29, 2017Updated 8 years ago
- Convolutional Neural Network of vgg19 model using Cuda to accelerate☆12Jun 11, 2018Updated 8 years ago
- A desktop pet program.☆11Aug 22, 2022Updated 3 years ago
- AutoParBench is a benchmark framework to evaluate compilers and tools designed to automatically insert OpenMP directives.☆12Nov 6, 2020Updated 5 years ago
- Initial framework source code for CSE240A branch predictor project☆10Nov 30, 2017Updated 8 years ago
- C++ heterogeneous and lock-free containers☆13Sep 5, 2018Updated 7 years ago
- Large matrix multiplication in CUDA☆17Oct 20, 2023Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆40Feb 28, 2020Updated 6 years ago
- We have developed Symbol Demonstration Direct Preference Optimization (SymDPO) and validating its effectiveness across multiple benchmark…☆23Nov 22, 2024Updated last year
- ☆13Nov 15, 2017Updated 8 years ago
- Learn CUDA Programming, published by Packt☆1,262Dec 30, 2023Updated 2 years ago
- A PyTorch implementation of our proposed loss function from the paper "SimLoss: Class Similarities in Cross Entropy"☆25Jun 18, 2021Updated 5 years ago
- Dehazing algorithm implemented on CUDA.☆10May 5, 2015Updated 11 years ago
- testbed for different SIMD implementations for set intersection and set union☆41Jan 29, 2020Updated 6 years ago