☆99Jul 23, 2026Updated this week
Alternatives and similar repositories for torch-xpu-ops
Users that are interested in torch-xpu-ops are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SYCL* Templates for Linear Algebra (SYCL*TLA) - SYCL based CUTLASS implementation for Intel GPUs☆77Updated this week
- OpenAI Triton backend for Intel® GPUs☆261Updated this week
- ☆61Mar 6, 2026Updated 4 months ago
- KFunca: A minimalist, high-performance GPU-based automatic differentiation framework☆31Aug 14, 2025Updated 11 months ago
- ONNX Runtime: cross-platform, high performance scoring engine for ML models☆88Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Multi-stage LLM agent pipeline for optimizing Triton kernels on Intel XPU — from analysis to autotuning.☆16Updated this week
- A Python package for extending the official PyTorch that can easily obtain performance on Intel platform☆2,014Mar 30, 2026Updated 3 months ago
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.☆14Jan 8, 2026Updated 6 months ago
- Intel® Extension for DeepSpeed* is an extension to DeepSpeed that brings feature support with SYCL kernels on Intel GPU(XPU) device. Note…☆65May 27, 2026Updated last month
- This repository contains Dockerfiles, scripts, yaml files, Helm charts, etc. used to scale out AI containers with versions of TensorFlow …☆79May 27, 2026Updated last month
- ☆710Updated this week
- Fast SGEMM emulation on Tensor Cores☆17Feb 16, 2025Updated last year
- Helper Files for IDC☆45Oct 23, 2023Updated 2 years ago
- ☆26Oct 9, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Triton-only attention backend for vLLM☆27Jul 14, 2026Updated last week
- oneAPI - Data Parallel C++ course for students☆43Nov 4, 2024Updated last year
- Intel® Graphics Compute Runtime for oneAPI Level Zero and OpenCL™ Driver☆1,420Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆90Jul 13, 2026Updated last week
- AI PC starter app for doing AI image creation, image stylizing, and chatbot on a PC powered by an Intel® Arc™ GPU.☆939Updated this week
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆57May 28, 2026Updated last month
- Easy and lightning fast training of 🤗 Transformers on Habana Gaudi processor (HPU)☆212Jul 6, 2026Updated 2 weeks ago
- Intel® Tensor Processing Primitives extension for Pytorch*☆19Jul 4, 2026Updated 2 weeks ago
- ☆430Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆17Jul 15, 2026Updated last week
- A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support…☆1,534Updated this week
- SYCL implementation of Fused MLPs for Intel GPUs☆51Jul 17, 2026Updated last week
- Intel® Extension for MLIR. A staging ground for MLIR dialects and tools for Intel devices using the MLIR toolchain.☆153Updated this week
- This repository hosts code that supports the testing infrastructure for the PyTorch organization. For example, this repo hosts the logic …☆110Updated this week
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆17Mar 11, 2026Updated 4 months ago
- ☆24Jun 12, 2023Updated 3 years ago
- SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, …☆2,684Updated this week
- ☆290Updated this week
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆18Jul 13, 2026Updated last week
- SynapseAI Core is a reference implementation of the SynapseAI API running on Habana Gaudi☆46Feb 3, 2025Updated last year
- ☆16Jun 4, 2026Updated last month
- Developing multi platform gesture detector application by applying concepts learnt in Embedded Systems course on peripheral devices.☆21Dec 8, 2023Updated 2 years ago
- ☆21Jan 21, 2026Updated 6 months ago
- Intel staging area for llvm.org contribution. Home for Intel LLVM-based projects.☆1,511Updated this week
- Reference models for Intel(R) Gaudi(R) AI Accelerator☆172Jan 8, 2026Updated 6 months ago