☆102Aug 12, 2026Updated this week
Alternatives and similar repositories for torch-xpu-ops
Users that are interested in torch-xpu-ops are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SYCL* Templates for Linear Algebra (SYCL*TLA) - SYCL based CUTLASS implementation for Intel GPUs☆80Updated this week
- OpenAI Triton backend for Intel® GPUs☆265Updated this week
- ☆61Mar 6, 2026Updated 5 months ago
- The vLLM XPU kernels for Intel GPU☆60Updated this week
- KFunca: A minimalist, high-performance GPU-based automatic differentiation framework☆31Aug 14, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Multi-stage LLM agent pipeline for optimizing Triton kernels on Intel XPU — from analysis to autotuning.☆20Aug 4, 2026Updated last week
- A Python package for extending the official PyTorch that can easily obtain performance on Intel platform☆2,013Mar 30, 2026Updated 4 months ago
- Explore our open source AI portfolio! Develop, train, and deploy your AI solutions with performance- and productivity-optimized tools fro…☆78Mar 27, 2026Updated 4 months ago
- A repository of Dockerfiles, scripts, yaml files, Helm Charts, etc. used to build and scale the sample AI workflows with python, kubernet…☆12Feb 22, 2024Updated 2 years ago
- Intel® Extension for DeepSpeed* is an extension to DeepSpeed that brings feature support with SYCL kernels on Intel GPU(XPU) device. Note…☆65May 27, 2026Updated 2 months ago
- This repository contains Dockerfiles, scripts, yaml files, Helm charts, etc. used to scale out AI containers with versions of TensorFlow …☆79May 27, 2026Updated 2 months ago
- ☆711Updated this week
- Fast SGEMM emulation on Tensor Cores☆17Feb 16, 2025Updated last year
- Helper Files for IDC☆45Oct 23, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆26Oct 9, 2025Updated 10 months ago
- A Triton-only attention backend for vLLM☆28Jul 14, 2026Updated 3 weeks ago
- Large Language Model Text Generation Inference on Habana Gaudi☆34Mar 20, 2025Updated last year
- Intel® Graphics Compute Runtime for oneAPI Level Zero and OpenCL™ Driver☆1,429Updated this week
- Profiling Tools Interfaces for GPU (PTI for GPU) is a set of Getting Started Documentation and Tools Library to start performance analysi…☆272Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆90Jul 27, 2026Updated 2 weeks ago
- AI PC starter app for doing AI image creation, image stylizing, and chatbot on a PC powered by an Intel® Arc™ GPU.☆952Updated this week
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆57May 28, 2026Updated 2 months ago
- Easy and lightning fast training of 🤗 Transformers on Habana Gaudi processor (HPU)☆212Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Intel® Tensor Processing Primitives extension for Pytorch*☆19Updated this week
- ☆18Jul 23, 2026Updated 3 weeks ago
- A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support…☆1,564Updated this week
- SYCL implementation of Fused MLPs for Intel GPUs☆51Jul 17, 2026Updated 3 weeks ago
- Intel® Optimization for Chainer*, a Chainer module providing numpy like API and DNN acceleration using MKL-DNN.☆180Jul 23, 2026Updated 3 weeks ago
- The repository contains a reference end-to-end pipeline for a real-time video analytics application. Realtime data is provided to an infe…☆12Nov 3, 2025Updated 9 months ago
- Sources for the Oak Ridge Leadership Computing Facility User Documentation☆67Updated this week
- Intel® Extension for MLIR. A staging ground for MLIR dialects and tools for Intel devices using the MLIR toolchain.☆156Updated this week
- This repository hosts code that supports the testing infrastructure for the PyTorch organization. For example, this repo hosts the logic …☆112Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Cosmic Tagging Network for Neutrino Physics☆13Jun 26, 2024Updated 2 years ago
- ☆24Jun 12, 2023Updated 3 years ago
- ☆290Updated this week
- SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, …☆2,698Updated this week
- ☆56Jul 16, 2026Updated 3 weeks ago
- ☆19Jul 13, 2026Updated last month
- SynapseAI Core is a reference implementation of the SynapseAI API running on Habana Gaudi☆46Feb 3, 2025Updated last year