Code appendix to an OpenCL matrix-multiplication tutorial
☆179Feb 7, 2017Updated 9 years ago
Alternatives and similar repositories for myGEMM
Users that are interested in myGEMM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tuned OpenCL BLAS☆1,171Updated this week
- a software library containing BLAS functions written in OpenCL☆865Aug 2, 2024Updated last year
- A portable high-level API with CUDA or OpenCL back-end☆56Oct 8, 2017Updated 8 years ago
- Sample program to compare calculation performance between CPU and GPU☆16Oct 27, 2016Updated 9 years ago
- Assembler for NVIDIA Maxwell architecture☆1,060Jan 3, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting with the flexibility to host WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Cloudways by DigitalOcean.
- A tool which profiles OpenCL devices to find their peak capacities☆485Updated this week
- The repository targets the OpenCL gemm function performance optimization. It compares several libraries clBLAS, clBLAST, MIOpenGemm, Inte…☆17Mar 28, 2019Updated 7 years ago
- Open single and half precision gemm implementations☆398Apr 2, 2023Updated 2 years ago
- OpenCL tool to detect buffer overflows in GPU kernels☆23Jan 7, 2019Updated 7 years ago
- An Android application which allows to execute native NDK program.☆34Aug 11, 2015Updated 10 years ago
- Sequential and parallel GEMM implementations with C interface + Benchmark.☆12May 24, 2016Updated 9 years ago
- Learn OpenCL step by step.☆140Aug 30, 2022Updated 3 years ago
- ☆2,002Jul 29, 2023Updated 2 years ago
- Caffe deep learning framework - optimized for Xeon Phi☆14May 12, 2015Updated 10 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Efficient SpGEMM on GPU using CUDA and CSR☆59Jul 18, 2023Updated 2 years ago
- FFI for OpenCL☆12Dec 19, 2015Updated 10 years ago
- Sparse Matrix-Vector Multiplication implementations in C☆22Dec 7, 2022Updated 3 years ago
- assembler for NVIDIA FERMI. Imported from Google Code☆74Mar 22, 2015Updated 11 years ago
- Materials for workshop on GPU computation for statistics, data science, machine learning applications.☆14Sep 8, 2016Updated 9 years ago
- Winograd-based convolution implementation in OpenCL☆28Jan 22, 2017Updated 9 years ago
- This is a tuned sparse matrix dense vector multiplication(SpMV) library☆23Mar 21, 2016Updated 10 years ago
- ☆27Oct 26, 2019Updated 6 years ago
- OpenCL API, OpenCL C, Extensions, SPIR-V Environment Specs, Ref page, and C++ for OpenCL doc sources.☆405Mar 10, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling on Cloudways • AdFully Managed hosting built for WordPress-powered businesses that need reliable, auto-scalable hosting. Cloudways SafeUpdates now available.
- load word embeddings to Torch.Tensor☆14May 12, 2016Updated 9 years ago
- An implementation of SGEMV with performance comparable to cuBLAS.☆12May 21, 2021Updated 4 years ago
- ☆32Aug 24, 2022Updated 3 years ago
- Benchmark for Co-running Single Applications on Integrated Architectures☆12Jul 7, 2016Updated 9 years ago
- OpenCL memory benchmark☆15Dec 21, 2016Updated 9 years ago
- Lecture Slide Issue Tracking☆256May 20, 2018Updated 7 years ago
- MAFIA: Multiple Application Framework for GPU architectures☆28Jan 21, 2022Updated 4 years ago
- Caffe: a fast open framework for deep learning.☆14Aug 26, 2015Updated 10 years ago
- An OpenCL device simulator and debugger☆371Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- The OpenDwarfs project provides a benchmark suite consisting of different computation/communication idioms, i.e., dwarfs, for state-of-ar…☆101Sep 18, 2019Updated 6 years ago
- Python implementations of fixed size hardware types (Bit, BitVector, UInt, SInt, ...) based on the SMT-LIB2 semantics☆18Sep 13, 2023Updated 2 years ago
- ☆257Sep 15, 2023Updated 2 years ago
- Winograd minimal convolution algorithm generator for convolutional neural networks.☆627Feb 9, 2026Updated last month
- row-major matmul optimization☆712Feb 24, 2026Updated last month
- OpenCL ICD Loader (free software)☆89Jan 23, 2026Updated 2 months ago
- ☆120Apr 11, 2024Updated last year