The Foundation for All Legate Libraries
☆241Jul 17, 2026Updated this week
Alternatives and similar repositories for legate
Users that are interested in legate are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- NumPy and SciPy on Multi-Node Multi-GPU systems☆980Updated this week
- An Aspiring Drop-In Replacement for Pandas at Scale☆74Oct 19, 2021Updated 4 years ago
- Legate Hello World Pedagogical Library☆10Apr 5, 2023Updated 3 years ago
- Legate Sparse is a Legate library that aims to provide a distributed and accelerated drop-in replacement for the scipy.sparse library on …☆26Apr 6, 2026Updated 3 months ago
- GBM implementation on Legate☆14Jul 10, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The Legion Parallel Programming System☆763Jul 2, 2026Updated 2 weeks ago
- Microbenchmarks showing relative performance of different Python functions/patterns.☆13Oct 3, 2025Updated 9 months ago
- ☆11Jul 13, 2022Updated 4 years ago
- Contains the xSDK community policies. The master branch is the latest accepted version of the policies and will be applied to future xSDK…☆11Jun 14, 2024Updated 2 years ago
- best CPU/GPU sparse solver for large sparse matrices☆21Oct 5, 2021Updated 4 years ago
- Tools and libraries for writing Kokkos-enabled HPC C++ in E3SM ecosystem☆22Updated this week
- An Attention Superoptimizer☆22Jan 20, 2025Updated last year
- [ARCHIVED] cuDF [alpha] - RAPIDS Merge of GoAi into cuDF☆34Oct 28, 2018Updated 7 years ago
- Unified Collective Communication Library☆310Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- CUDA Kernel Benchmarking Library☆901Updated this week
- cuNumeric.jl wraps the cuPyNumeric C++ API providing a simple array programming interface that executes code on distributed clusters.☆21Jul 9, 2026Updated last week
- Example codes demonstrating the use of various XSDK packages in combination.☆19Jun 1, 2023Updated 3 years ago
- Matrix multiplication on GPUs for matrices stored on a CPU. Similar to cublasXt, but ported to both NVIDIA and AMD GPUs.☆33Apr 2, 2025Updated last year
- [ARCHIVED] The C++ Standard Library for your entire system. See https://github.com/NVIDIA/cccl☆2,304Feb 7, 2024Updated 2 years ago
- Implementation of AMD HIP for CPUs☆22Jun 16, 2020Updated 6 years ago
- The Exascale Computing Project Software Technologies Capability Assessment Report - Public Version☆21Aug 18, 2022Updated 3 years ago
- Standard interface for collecting HPC run metadata☆16Nov 7, 2025Updated 8 months ago
- List all available information about all SYCL devices and platforms☆15Sep 14, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A tracing JIT compiler for PyTorch☆14Dec 11, 2021Updated 4 years ago
- DaCe - Data Centric Parallel Programming☆591Updated this week
- Distributed SDDMM Kernel☆12Jul 8, 2022Updated 4 years ago
- A task benchmark☆46Apr 17, 2026Updated 3 months ago
- Examples demonstrating available options to program multiple GPUs in a single node or a cluster☆908Sep 26, 2025Updated 9 months ago
- A single-header C++ library for simplifying the use of CUDA Runtime Compilation (NVRTC).☆573Sep 15, 2025Updated 10 months ago
- MPI accelerator-integrated communication extensions☆39Apr 4, 2023Updated 3 years ago
- A simple yet powerful tool to turn traditional container/OS images into unprivileged sandboxes.☆978Jun 9, 2026Updated last month
- Codes for supersymmetric lattice gauge theories☆18Jun 26, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Python bindings for UCX☆139Sep 18, 2025Updated 10 months ago
- Distributed Communication-Optimal Matrix-Matrix Multiplication Algorithm☆215Apr 18, 2026Updated 3 months ago
- Collection of scripts to build PyTorch and the domain libraries from source.☆14Jul 9, 2026Updated last week
- Abstraction Library for Parallel Kernel Acceleration☆419Jun 25, 2026Updated 3 weeks ago
- [ARCHIVED] Cooperative primitives for CUDA C++. See https://github.com/NVIDIA/cccl☆1,840Oct 9, 2023Updated 2 years ago
- LaunchMON is a software infrastructure that enables HPC run-time tools to co-locate tool daemons with a parallel job. Its API allows a to…☆13Updated this week
- YAKL is A Kokkos Layer: A simple C++ framework for performance portability and Fortran code porting☆71May 8, 2026Updated 2 months ago