🎃 GPU load-balancing library for regular and irregular computations.
☆67Jun 25, 2026Updated 3 weeks ago
Alternatives and similar repositories for loops
Users that are interested in loops are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Apr 24, 2024Updated 2 years ago
- mini is mini☆20Jan 19, 2020Updated 6 years ago
- Runs a single CUDA/OpenCL kernel, taking its source from a file and arguments from the command-line☆26Jun 10, 2026Updated last month
- Open-source library for Graph Streaming. Solves the connected components problem using sub-linear space. Published in SIGMOD'22.☆11Apr 6, 2026Updated 3 months ago
- A vectorizable multi-dimensional iterator for C++ using the Coroutines TS☆12Jun 5, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- CUDA Dynamic Memory Allocator for SOA Data Layout☆39Dec 29, 2021Updated 4 years ago
- Source code for the paper: Accelerating Dynamic Graph Analytics on GPUs☆30Jun 19, 2023Updated 3 years ago
- ☆20Jan 17, 2024Updated 2 years ago
- ☆655Updated this week
- cuASR: CUDA Algebra for Semirings☆49Aug 22, 2022Updated 3 years ago
- Chapel HyperGraph Library (CHGL) - HPC-class Hypergraphs in Chapel☆33Oct 29, 2020Updated 5 years ago
- ☆11Aug 8, 2021Updated 4 years ago
- LonestarGPU: Irregular algorithms parallelized for GPUs☆38Nov 11, 2019Updated 6 years ago
- Efficient and High-quality Graph Coloring on the GPU☆16Apr 3, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆11Apr 10, 2019Updated 7 years ago
- Source code supporting the High Performance Graphics 2022 paper: Supporting Unified Shader Specialization by Co-opting C++ Features☆14Jul 9, 2022Updated 4 years ago
- ☆18Oct 15, 2020Updated 5 years ago
- Record GPU memory accesses of a CUDA program and visualize the access pattern in a browser☆13Nov 17, 2020Updated 5 years ago
- Multi-GPU dynamic scheduler using PGAS style cross-GPU communication☆29Jul 23, 2023Updated 3 years ago
- Programmable CUDA/C++ GPU Graph Analytics☆1,096Feb 28, 2026Updated 4 months ago
- Evaluating different memory managers for dynamic GPU memory☆26Dec 16, 2020Updated 5 years ago
- GPUDirect Async implementation of HPGMG-FV CUDA☆11May 11, 2018Updated 8 years ago
- PilotFish harvests the free GPU cycles of cloud gaming with deep learning training☆14Jul 2, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Department of Energy Standard Utility Library☆34Jun 26, 2026Updated 3 weeks ago
- A pattern-based algorithmic autotuner for graph processing on GPUs.☆33Jun 25, 2025Updated last year
- A Lightweight Graph Processing Framework for Multi-GPUs☆14Apr 15, 2015Updated 11 years ago
- A Collection of Parallel Algorithms for Computational Geometry☆12Mar 10, 2022Updated 4 years ago
- Matrix multiplication on GPUs for matrices stored on a CPU. Similar to cublasXt, but ported to both NVIDIA and AMD GPUs.☆33Apr 2, 2025Updated last year
- Fast SGEMM emulation on Tensor Cores☆17Feb 16, 2025Updated last year
- Statistics on GPUs☆33May 5, 2026Updated 2 months ago
- Artifact for PPoPP20 "Understanding and Bridging the Gaps in Current GNN Performance Optimizations"☆42Nov 16, 2021Updated 4 years ago
- Using C++ magic to capture CUDA kernels and tune them with Kernel Tuner☆22Sep 12, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- GenDP: A Dynamic Programming Framework for Genome Sequencing Analysis☆17Jan 12, 2024Updated 2 years ago
- A reference implementation of std::simd, providing data parallel types in the C++ standard☆14Mar 9, 2020Updated 6 years ago
- Artifact for OSDI'21 GNNAdvisor: An Adaptive and Efficient Runtime System for GNN Acceleration on GPUs.☆71Mar 2, 2023Updated 3 years ago
- SparseTIR: Sparse Tensor Compiler for Deep Learning☆145Mar 31, 2023Updated 3 years ago
- Repository for artifact evaluation of ASPLOS 2023 paper "SparseTIR: Composable Abstractions for Sparse Compilation in Deep Learning"☆25Feb 24, 2023Updated 3 years ago
- Scale-out system monitoring☆25Updated this week
- ☆31Aug 28, 2020Updated 5 years ago