❤️ CUDA/C++ GPU graph analytics simplified.
☆32Sep 19, 2022Updated 3 years ago
Alternatives and similar repositories for essentials
Users that are interested in essentials are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- mini is mini☆20Jan 19, 2020Updated 6 years ago
- 🎃 GPU load-balancing library for regular and irregular computations.☆67Jun 25, 2026Updated last month
- Multi-GPU dynamic scheduler using PGAS style cross-GPU communication☆29Jul 23, 2023Updated 3 years ago
- ☆13May 21, 2020Updated 6 years ago
- Programmable CUDA/C++ GPU Graph Analytics☆1,096Feb 28, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A pattern-based algorithmic autotuner for graph processing on GPUs.☆33Jun 25, 2025Updated last year
- A package for constructing sparse tensors from CSV-like data sources.☆11Dec 24, 2017Updated 8 years ago
- This is a mirror of https://gitlab.com/tiro-is/tiro-speech-core☆15Jun 19, 2023Updated 3 years ago
- Medusa: Building GPU-based Parallel Sparse Graph Applications with Sequential C/C++ Code☆63Oct 17, 2020Updated 5 years ago
- Statistics on GPUs☆33May 5, 2026Updated 3 months ago
- ☆23Feb 16, 2022Updated 4 years ago
- ☆16Feb 26, 2020Updated 6 years ago
- Sparse matrix-matrix multiplication on CPU+GPU systems.☆13Mar 17, 2014Updated 12 years ago
- Mallacc: Accelerating Memory Allocation☆13Jan 2, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A CUDA-based multi-GPU vertex-centric graph processing framework based on Warp Segmentation and Vertex Refinement techniques.☆12Mar 20, 2017Updated 9 years ago
- ☆31Aug 28, 2020Updated 5 years ago
- This code base represents "faimGraph: High Performance Management of Fully-dynamic Graphs under tight Memory Constraints on the GPU"☆16Apr 23, 2021Updated 5 years ago
- High-Performance Linear Algebra-based Graph Primitives on GPUs☆238Jul 2, 2021Updated 5 years ago
- Reference implementation of the draft C++ GraphBLAS specification.☆32Feb 19, 2025Updated last year
- A dynamic GPU memory allocator, suitable for warp synchronized scenarios.☆11Aug 20, 2019Updated 6 years ago
- rust-library to wrap GraphBLAS.h☆31Mar 19, 2023Updated 3 years ago
- GPU MemoryManager based on virtualized queues☆27Jun 25, 2022Updated 4 years ago
- ☆13Nov 4, 2020Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Record GPU memory accesses of a CUDA program and visualize the access pattern in a browser☆13Nov 17, 2020Updated 5 years ago
- CUDA Kernel Benchmarking Library☆915Updated this week
- SIMD-X: Programming and Processing of Graph Algorithms on GPUs [USENIX ATC '19]☆23Jun 14, 2020Updated 6 years ago
- Aspen is a Low-Latency Graph Streaming System built using Compressed Purely-Functional Trees☆90Jun 24, 2019Updated 7 years ago
- ☆12Jul 28, 2022Updated 4 years ago
- A tracing JIT compiler for PyTorch☆14Dec 11, 2021Updated 4 years ago
- Source code supporting the High Performance Graphics 2022 paper: Supporting Unified Shader Specialization by Co-opting C++ Features☆14Jul 9, 2022Updated 4 years ago
- Lightweight speaker anonymization [IEEE SLT2021]☆27Jun 6, 2022Updated 4 years ago
- Galois: C++ library for multi-core and multi-node parallelization☆354May 16, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- LU Decomposition using CUDA☆13Dec 7, 2013Updated 12 years ago
- A Memory-efficient Graph Store for Interactive Queries☆13Sep 1, 2021Updated 4 years ago
- Itoyori: A distributed multi-threading runtime system for global-view fork-join task parallelism☆23Feb 9, 2024Updated 2 years ago
- A recommendation model kernel optimizing system☆12Jun 5, 2025Updated last year
- GBDT-based model with efficient unlearning (SIGMOD 2023)☆10Sep 7, 2025Updated 11 months ago
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆145Updated this week
- Sandia OpenSHMEM is an implementation of the OpenSHMEM specification over multiple Networking APIs, including Portals 4, the Open Fabric …☆78Jul 13, 2026Updated 3 weeks ago