NVIDIA Math Libraries for the Python Ecosystem
☆600Jul 14, 2026Updated 2 months ago
Alternatives and similar repositories for nvmath-python
Users that are interested in nvmath-python are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The CUDA target for Numba☆296Updated this week
- Numbast is a tool to build an automated pipeline that converts CUDA APIs into Numba bindings.☆62Oct 2, 2026Updated last week
- CUDA Python: Performance meets Productivity☆3,395Updated this week
- CUDA Core Compute Libraries☆2,530Updated this week
- A Python framework for GPU-accelerated simulation, robotics, and machine learning.☆7,179Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- repo for Numba-CUDA-MLIR☆69Updated this week
- NumPy & SciPy for GPU☆12,360Updated this week
- CUDA Templates and Python DSLs for High-Performance Linear Algebra☆10,547Sep 23, 2026Updated 2 weeks ago
- cuTile is a programming model for writing parallel kernels for NVIDIA GPUs☆2,154Updated this week
- Functional algorithms - definitions and implementations☆16Oct 17, 2025Updated 11 months ago
- An efficient C++20 GPU numerical computing library with Python-like syntax☆1,447Updated this week
- End of life; no longer maintained. Final public release: v26.06.01.☆982Oct 1, 2026Updated last week
- NVIDIA curated collection of educational resources related to general purpose GPU programming.☆2,040Updated this week
- A Python-embedded DSL that makes it easy to write fast, scalable ML kernels with minimal boilerplate.☆960Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Tilus is a tile-level kernel programming language with explicit control over shared memory and registers.☆495Sep 17, 2026Updated 3 weeks ago
- MSLK (Meta Superintelligence Labs Kernels) is a collection of PyTorch GPU operator libraries that are designed and optimized for GenAI tr…☆155Updated this week
- CUDA Tile IR is an MLIR-based intermediate representation and compiler infrastructure for CUDA kernel optimization, focusing on tile-base…☆1,039Sep 28, 2026Updated last week
- A Fusion Code Generator for NVIDIA GPUs (commonly known as "nvFuser")☆406May 31, 2026Updated 4 months ago
- CUDA Library Samples☆2,525Sep 28, 2026Updated last week
- NVIDIA Performance Libraries: Sample code☆24May 28, 2026Updated 4 months ago
- A stand-alone implementation of several NumPy dtype extensions used in machine learning.☆361Sep 24, 2026Updated 2 weeks ago
- Communication patterns for AI, built on top of NCCL device and host APIs☆68Updated this week
- Parrot is an array fusion GPU library built on NVIDIA's CCCL libaries (Thrust/CUB).☆288Oct 1, 2026Updated last week
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- cuTile Rust provides a safe, tile-based kernel programming DSL for the Rust programming language. It features a safe host-side API for pa…☆1,059Updated this week
- JAX-Toolbox☆432Updated this week
- Test data for DALI project☆45Aug 28, 2026Updated last month
- ☆48Oct 1, 2026Updated last week
- GPU accelerated decision optimization☆1,056Updated this week
- ☆21Mar 3, 2025Updated last year
- PyTorch native quantization for training and inference☆2,995Updated this week
- An Aspiring Drop-In Replacement for Pandas at Scale☆74Oct 19, 2021Updated 4 years ago
- A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on H…☆3,571Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- jax-triton contains integrations between JAX and OpenAI Triton☆472Sep 28, 2026Updated last week
- CUDA Kernel Benchmarking Library☆935Updated this week
- Development repository for the Triton language and compiler☆20,321Updated this week
- Pytorch routines for (Ker)nel (Mac)hines☆11Oct 10, 2025Updated 11 months ago
- PyTorch Single Controller☆1,078Updated this week
- An MPI ABI compatibility layer☆34Jun 17, 2026Updated 3 months ago
- A pure-Python implementation of the Nvidia CuTe layout algebra intended to be approachable and easy to learn.☆253Jun 29, 2026Updated 3 months ago