No-GIL Python environment featuring NVIDIA Deep Learning libraries.
☆72Apr 14, 2025Updated last year
Alternatives and similar repositories for free-threaded-python
Users that are interested in free-threaded-python are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Jan 7, 2025Updated last year
- [WIP] Better (FP8) attention for Hopper☆33Feb 24, 2025Updated last year
- FA4-based Relative Attention Kernel developed by TML and Colfax☆18Jul 17, 2026Updated 3 weeks ago
- A GPU shallow-water equation (SWE) solver☆15May 5, 2022Updated 4 years ago
- ☆15Apr 7, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Octoconda runner☆18Updated this week
- A Claude Code marketplace and skill for working in the cpython repo.☆17Jul 30, 2026Updated last week
- study of cutlass☆22Nov 10, 2024Updated last year
- ☆11Feb 26, 2024Updated 2 years ago
- Samples demonstrating how to use the Compute Sanitizer Tools and Public API☆99Nov 6, 2023Updated 2 years ago
- FLA but cuTile☆27Apr 17, 2026Updated 3 months ago
- #UAI2020 Codes for PAC-Bayesian Contrastive Unsupervised Representation Learning☆14May 23, 2022Updated 4 years ago
- A survey of manufacturer-provided DRAM operating parameters and timings as specified by DRAM chip datasheets from between 1970 and 2021. …☆11May 4, 2022Updated 4 years ago
- Matrix multiplication on GPUs for matrices stored on a CPU. Similar to cublasXt, but ported to both NVIDIA and AMD GPUs.☆33Apr 2, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Triton kernels and PyTorch ops for Block Attention Residuals (AttnRes)☆87May 29, 2026Updated 2 months ago
- Handwritten GEMM using Intel AMX (Advanced Matrix Extension)☆17Jan 11, 2025Updated last year
- Tensor Basis Neural Network for Scalar Mixing☆10Mar 24, 2023Updated 3 years ago
- A shell-friendly hyperparameter search tool inspired by Optuna☆18Dec 17, 2024Updated last year
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆145Updated this week
- Official repository Flash Local Linear Attention☆38May 28, 2026Updated 2 months ago
- Monitor parameter and gradient statistics during neural network training with Chainer☆13Jan 24, 2017Updated 9 years ago
- Scalable GPU Kernel Fission/Fusion Transformation for Memory-Bound Kernels☆14Aug 26, 2015Updated 10 years ago
- TiledLower is a Dataflow Analysis and Codegen Framework written in Rust.☆13Nov 23, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Reimplementation of the paper `Human Attention Maps for Text Classification: Do Humans and Neural Networks Focus on the Same Words? (ACL2…☆17Jul 10, 2020Updated 6 years ago
- ☆32Jul 2, 2025Updated last year
- GVProf: A Value Profiler for GPU-based Clusters☆54Mar 24, 2024Updated 2 years ago
- General purpose, language-agnostic Continuous Benchmarking (CB) framework☆35Apr 15, 2020Updated 6 years ago
- WheelNext Website☆55Dec 19, 2025Updated 7 months ago
- Scale Optuna with Dask☆37Oct 1, 2020Updated 5 years ago
- Multi-Level Triton Runner supporting Python, IR, PTX, AMDGCN, cubin and hasco.☆99May 8, 2026Updated 3 months ago
- An experimental communicating attention kernel based on DeepEP.☆34Jul 29, 2025Updated last year
- CUDA 12.2 HMM demos☆21Jul 26, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- GPU Performance Advisor☆66Jul 25, 2022Updated 4 years ago
- ☆27Mar 14, 2024Updated 2 years ago
- WaveSpeedAI Python Client — Official Python SDK for WaveSpeedAI inference platform. This library provides a clean, unified, and high-perf…☆25Jul 20, 2026Updated 3 weeks ago
- Open Source SSD Controller. NVMe and Lightstor variants☆17May 21, 2014Updated 12 years ago
- ☆40Dec 14, 2025Updated 7 months ago
- A memory profiler for NVIDIA GPUs to explore memory inefficiencies in GPU-accelerated applications.☆38May 30, 2026Updated 2 months ago
- Policies, Configurations, and Documentation of NumFOCUS Managed Infrastructure☆16Aug 1, 2026Updated last week