SpMV using CUDA
☆20Mar 5, 2018Updated 8 years ago
Alternatives and similar repositories for Sparse-Matrix-Vector-Multiplication
Users that are interested in Sparse-Matrix-Vector-Multiplication are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CUDA Sparse-Matrix Vector Multiplication, using Sliced Coordinate format☆22Jun 8, 2018Updated 8 years ago
- Implementation and analysis of five different GPU based SPMV algorithms in CUDA☆39Feb 5, 2019Updated 7 years ago
- ☆98Feb 10, 2017Updated 9 years ago
- Source code of the IPDPS '21 paper: "TileSpMV: A Tiled Algorithm for Sparse Matrix-Vector Multiplication on GPUs" by Yuyao Niu, Zhengyang…☆13Aug 12, 2022Updated 4 years ago
- Parallel SpMV using CSR representation, built in CUDA☆14Jun 27, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- CSR5-based SpMV on CPUs, GPUs and Xeon Phi☆111Jun 10, 2024Updated 2 years ago
- This is a tuned sparse matrix dense vector multiplication(SpMV) library☆23Mar 21, 2016Updated 10 years ago
- This package includes the implementation for four sparse linear algebra kernels: Sparse-Matrix-Vector-Multiplication (SpMV), Sparse-Trian…☆29Jun 1, 2020Updated 6 years ago
- Parallelized and vectorized SpMV on Intel Xeon Phi (Knights Landing, AVX512, KNL)☆24Feb 12, 2024Updated 2 years ago
- ☆24Jul 30, 2026Updated 2 weeks ago
- Different implementation of sparse matrix multiplication. All matrices are in CSR format. The code contains different CUDA kernels for mu…☆17Nov 15, 2010Updated 15 years ago
- Implementation of paper "GraphACT: Accelerating GCN Training on CPU-FPGA Heterogeneous Platform".☆12Jun 25, 2020Updated 6 years ago
- SpV8 is a SpMV kernel written in AVX-512. Artifact for our SpV8 paper @ DAC '21.☆29Mar 16, 2021Updated 5 years ago
- Public repostory for the DAC 2021 paper "Scaling up HBM Efficiency of Top-K SpMV forApproximate Embedding Similarity on FPGAs"☆16Aug 29, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An artificial matrix generator in C☆13Feb 16, 2023Updated 3 years ago
- CSR-based SpMV on Heterogeneous Processors (Intel Broadwell, AMD Kaveri and nVidia Tegra K1)☆26May 12, 2015Updated 11 years ago
- Source code of the PPoPP '22 paper: "TileSpGEMM: A Tiled Algorithm for Parallel Sparse General Matrix-Matrix Multiplication on GPUs" by Y…☆48May 22, 2024Updated 2 years ago
- A Vector Caching Scheme for Streaming FPGA SpMV Accelerators☆10Sep 7, 2015Updated 10 years ago
- A intelligent matrix format designer for SpMV☆10Oct 10, 2023Updated 2 years ago
- BigDataBench Spark workloads☆11Jul 15, 2016Updated 10 years ago
- 我的《剑指Offer》(第 2 版)刷题记录,包括代码和笔记。☆14May 27, 2019Updated 7 years ago
- HIP acceleration of SpMV solver☆13May 17, 2025Updated last year
- ☆10Jun 9, 2017Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Some simple examples for the Magic VLSI physical chip layout tool.☆31Mar 9, 2021Updated 5 years ago
- ☆35Sep 10, 2024Updated last year
- Incomplete-Cholesky preconditioned conjugate gradient algorithm implemented with cuBLAS/cuSPARSE☆12Jun 24, 2022Updated 4 years ago
- Langevin and Hybrid Quantum Monte Carlo Simulations of Electron-Phonon Models☆14Aug 15, 2022Updated 3 years ago
- Efficient Global Optimization☆10Feb 26, 2016Updated 10 years ago
- ☆10Jan 24, 2019Updated 7 years ago
- Source code of the SC '23 paper: "DASP: Specific Dense Matrix Multiply-Accumulate Units Accelerated General Sparse Matrix-Vector Multipli…☆29Jun 18, 2024Updated 2 years ago
- A Chip Design Automation Solution with Open Source EDA Tools.☆17Updated this week
- Efficient and stable Determinant Quantum Monte Carlo simulations in Python☆11Feb 23, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Pytorch implementation of the paper "Debiasing the Cloze Task in Sequential Recommendation with Bidirectional Transformers".☆12Jan 22, 2023Updated 3 years ago
- ☆12Nov 23, 2020Updated 5 years ago
- OpenDesign Flow Database☆17Oct 31, 2018Updated 7 years ago
- ☆21Nov 22, 2020Updated 5 years ago
- accessing internal network of USTB outside school. USTB/BUPT(或其他基于webvpn的高校) 网络访问代理工具,可在校外无缝连接校内网络(网站、ssh、git、vnc、win远程桌面等)☆58Jan 28, 2026Updated 6 months ago
- IMPORTANT NOTICE: This implementation is long outdated. Whole-Function Vectorization is an algorithm that transforms a scalar function in…☆23May 16, 2012Updated 14 years ago
- Mirror of MAGMA - Next-generation linear algebra libraries for heterogeneous architectures. Please use the official repository, https://b…☆25Aug 29, 2017Updated 8 years ago