Different implementation of sparse matrix multiplication. All matrices are in CSR format. The code contains different CUDA kernels for multiply sparse matrix vs dense vector and sparse matrix vs another sparse matrix. It contains several cuda kernel for sparse matrix dense vector product and sparse matrix sparse matrix product.
☆17Nov 15, 2010Updated 15 years ago
Alternatives and similar repositories for CudaDotProd
Users that are interested in CudaDotProd are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Accelerating CNN's convolution operation on GPUs by using memory-efficient data access patterns.☆14Dec 8, 2017Updated 8 years ago
- Dust Streams in Space☆15Mar 5, 2016Updated 10 years ago
- SpMV using CUDA☆20Mar 5, 2018Updated 8 years ago
- This is a clone of an SVN repository at https://cudafy.svn.codeplex.com/svn. It had been cloned by http://svn2github.com/ , but the servi…☆12Jun 15, 2015Updated 11 years ago
- A simple but efficient C++ thread/worker pool library for asynchronous task management.☆10Jul 11, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- a Go package for encoding and decoding time-series data point in similar way to Facebook Gorilla time-series database☆14Feb 12, 2017Updated 9 years ago
- An intuitive user interface for hp-finite element analysis of three-dimensional piezoelectric beams☆10Feb 27, 2016Updated 10 years ago
- Homemade Pixel Art Tool (WIP)☆17Oct 18, 2024Updated last year
- ☆12Feb 23, 2023Updated 3 years ago
- CNN learns feature mapping between corrupted and clean speech☆12Aug 14, 2017Updated 8 years ago
- Reinforcement learning training project for a SLG game☆13Dec 21, 2017Updated 8 years ago
- Zero-Overhead bare-metal GPGPU library for C++ on Windows.☆15Jan 29, 2017Updated 9 years ago
- golang implementation of wilson score, reddit & hackernews scoring algorithm☆11Jul 20, 2015Updated 11 years ago
- A cache that automatically removes the least-recently-used items☆18Dec 16, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- CS510 Advanced Topics in Concurrency Project☆16Jun 4, 2020Updated 6 years ago
- CUDA implementation of the Jacobi method☆24Mar 3, 2015Updated 11 years ago
- TensorRT-5 based inference engine in Python☆14Sep 23, 2018Updated 7 years ago
- Ring network model test to demonstrate the use of CoreNEURON☆11Jul 5, 2026Updated 2 weeks ago
- Exploring data science through Julia programming language☆13Sep 25, 2021Updated 4 years ago
- Implementation of a Neural Network in .NET using the Diagnostic Wisconsin Breast Cancer Database.☆17Jul 22, 2014Updated 12 years ago
- W3C sitesindeki SQL Editörünün, Türkçe veritabanı ile hazırlanmış halidir.☆11Dec 10, 2015Updated 10 years ago
- C++11 Header-only continuous-storage Double ended vector implementation similar to STL's std::vector for efficient insertions/removals at…☆16Dec 29, 2022Updated 3 years ago
- Python caching libraries benchmark - which is better?☆13Jun 2, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A CUDA-C implementation of FOFE and FSMN☆19Aug 5, 2016Updated 9 years ago
- 第二届云原生编程挑战赛: RocketMQ存储系统设计 第4名 我之渺小 队代码☆11Nov 3, 2021Updated 4 years ago
- ☆10May 1, 2020Updated 6 years ago
- Efficient CUDA Stream Compaction Library☆34Jun 9, 2023Updated 3 years ago
- TUS Server Implementation for coffee lovers☆16Jun 3, 2015Updated 11 years ago
- ☆10Aug 4, 2022Updated 3 years ago
- An interactive Unity gpgpu fluid simulation☆14Apr 1, 2014Updated 12 years ago
- Serving Images dynamically based on the client device is an important part of Web Page Resource Optimization. ImgR.NET aims at automating…☆12Sep 28, 2016Updated 9 years ago
- Catmull-Clark Subdivision Surface algorithm for Skinned Mesh in Unity3d☆16Jan 1, 2017Updated 9 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- An implementation of the Pregel graph processing system on the Spark cluster computing framework. Merged into Spark; please see:☆11Apr 9, 2011Updated 15 years ago
- Sparse Recurrent Neural Networks -- Pruning Connections and Hidden Sizes (TensorFlow)☆74Jul 25, 2020Updated 6 years ago
- A working example for developing Alexa Skills Kit for your Amazon Echo with AWS Lambda.☆18Feb 11, 2016Updated 10 years ago
- Hadoop-based tool for extraction of large scale synchronous grammars for paraphrasing and machine translation☆15Dec 2, 2016Updated 9 years ago
- Context2Bundle: Diversified Personalized Bundle Recommendation☆12Feb 22, 2018Updated 8 years ago
- Webcam Image Processing with CUDA using OpenCV☆16Aug 30, 2014Updated 11 years ago
- hybrid computing engine executed by both GPU and multicore to accelerate PH matrix reduction☆13Dec 2, 2019Updated 6 years ago