DLPrimitives/OpenCL out of tree backend for pytorch
☆399Nov 26, 2025Updated 7 months ago
Alternatives and similar repositories for pytorch_dlprim
Users that are interested in pytorch_dlprim are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deep Learning Primitives and Mini-Framework for OpenCL☆211Sep 9, 2024Updated last year
- Example of using pytorch's open device registration API☆31Oct 14, 2022Updated 3 years ago
- OpenCL port of TensorFlow using SYCL, generic instructions for building are here:☆62Mar 31, 2020Updated 6 years ago
- chipStar is a tool for compiling and running HIP/CUDA on SPIR-V via OpenCL or Level Zero APIs.☆364Updated this week
- MLIR 中文文档☆22Dec 1, 2025Updated 7 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A prototype CUDA-to-OpenCL source-to-source translator, built on the Clang compiler framework☆209Jul 12, 2020Updated 6 years ago
- Build NVIDIA® CUDA™ code for OpenCL™ 1.2 devices☆877Apr 23, 2025Updated last year
- Tuned OpenCL BLAS☆1,186Apr 13, 2026Updated 3 months ago
- JAX interpreter for Vulkan☆17Jun 1, 2021Updated 5 years ago
- Tensor Tiling Library☆42Sep 23, 2025Updated 9 months ago
- A portable GPU/CPU Path Tracer library powered by SYCL. (OpenCL/CUDA/OpenMP)☆16Feb 19, 2019Updated 7 years ago
- ☆19Sep 10, 2024Updated last year
- ☆11Updated this week
- Useful assemblies of neural network software.☆14Oct 17, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Easy to run kernels using OpenCL☆188Apr 22, 2025Updated last year
- Recording models☆12Sep 19, 2023Updated 2 years ago
- Community emulator for TI nspire handhelds☆12Jun 25, 2026Updated 3 weeks ago
- CUDA on non-NVIDIA GPUs☆14,630Updated this week
- Artifacts of EVT ASPLOS'24☆29Mar 6, 2024Updated 2 years ago
- Antares: an automatic engine for multi-platform kernel generation and optimization. Supporting CPU, CUDA, ROCm, DirectX12, GraphCore, SYC…☆464Apr 20, 2025Updated last year
- A Python package for extending the official PyTorch that can easily obtain performance on Intel platform☆2,014Mar 30, 2026Updated 3 months ago
- Implementation of OpenCL 3.0 on Vulkan☆442Updated this week
- ☆15Jan 12, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A guide to help developers get up and running quickly with the OpenCL programming framework☆699Aug 7, 2024Updated last year
- A Symbolic Emulator for Shuffle Synthesis on the NVIDIA PTX Code☆16Mar 19, 2023Updated 3 years ago
- OpenCL library to train deep convolutional neural networks☆881Jan 5, 2018Updated 8 years ago
- Simple starter CMake project that uses NVBench.☆15May 6, 2025Updated last year
- A synthetic micro-benchmark that measures peak compute, bandwidth, and matrix throughput of GPUs and CPUs☆505Updated this week
- Configure NVMe by CLI, and test it with fio!☆17Updated this week
- The Torch-MLIR project aims to provide first class support from the PyTorch ecosystem to the MLIR ecosystem.☆1,871Updated this week
- Exocompilation for productive programming of hardware accelerators☆736Jul 3, 2026Updated 2 weeks ago
- Compiler for multiple programming models (SYCL, C++ standard parallelism, HIP/CUDA) for CPUs and GPUs from all vendors: The independent, …☆1,910Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆21Jan 21, 2026Updated 6 months ago
- GPGPU array on Vulkan☆17Jun 3, 2023Updated 3 years ago
- General purpose GPU compute framework built on Vulkan to support 1000s of cross vendor graphics cards (AMD, Qualcomm, NVIDIA & friends). …☆2,541Updated this week
- Chunky Loop Interaction☆25Aug 13, 2019Updated 6 years ago
- ☆17Mar 2, 2020Updated 6 years ago
- ☆13Aug 1, 2024Updated last year
- ExaWorks SDK☆11Feb 1, 2024Updated 2 years ago