☆114Sep 22, 2026Updated this week
Alternatives and similar repositories for torch-xpu-ops
Users that are interested in torch-xpu-ops are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SYCL* Templates for Linear Algebra (SYCL*TLA) - SYCL based CUTLASS implementation for Intel GPUs☆86Updated this week
- OpenAI Triton backend for Intel® GPUs☆271Updated this week
- ☆65Mar 6, 2026Updated 6 months ago
- The vLLM XPU kernels for Intel GPU☆71Updated this week
- KFunca: A minimalist, high-performance GPU-based automatic differentiation framework☆31Aug 14, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ONNX Runtime: cross-platform, high performance scoring engine for ML models☆91Updated this week
- A Python package for extending the official PyTorch that can easily obtain performance on Intel platform☆2,009Mar 30, 2026Updated 5 months ago
- Explore our open source AI portfolio! Develop, train, and deploy your AI solutions with performance- and productivity-optimized tools fro…☆78Mar 27, 2026Updated 5 months ago
- A repository of Dockerfiles, scripts, yaml files, Helm Charts, etc. used to build and scale the sample AI workflows with python, kubernet…☆12Feb 22, 2024Updated 2 years ago
- Intel® Extension for DeepSpeed* is an extension to DeepSpeed that brings feature support with SYCL kernels on Intel GPU(XPU) device. Note…☆65May 27, 2026Updated 3 months ago
- ☆718Updated this week
- Fast SGEMM emulation on Tensor Cores☆17Feb 16, 2025Updated last year
- Helper Files for IDC☆45Oct 23, 2023Updated 2 years ago
- ☆27Oct 9, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Triton-only attention backend for vLLM☆28Jul 14, 2026Updated 2 months ago
- Profiling Tools Interfaces for GPU (PTI for GPU) is a set of Getting Started Documentation and Tools Library to start performance analysi…☆273Sep 11, 2026Updated last week
- Intel® Graphics Compute Runtime for oneAPI Level Zero and OpenCL™ Driver☆1,447Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆91Sep 2, 2026Updated 3 weeks ago
- AI PC starter app for doing AI image creation, image stylizing, and chatbot on a PC powered by an Intel® Arc™ GPU.☆984Updated this week
- [DEPRECATED] Moved to ROCm/rocm-libraries repo☆57May 28, 2026Updated 3 months ago
- Easy and lightning fast training of 🤗 Transformers on Habana Gaudi processor (HPU)☆213Sep 7, 2026Updated 2 weeks ago
- Intel® Tensor Processing Primitives extension for Pytorch*☆20Aug 24, 2026Updated 3 weeks ago
- ☆538Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A simple and effective quantization toolkit for high-accuracy low-bit LLM inference|简洁且高效的量化工具包☆1,622Updated this week
- SYCL implementation of Fused MLPs for Intel GPUs☆50Sep 7, 2026Updated 2 weeks ago
- Sources for the Oak Ridge Leadership Computing Facility User Documentation☆67Updated this week
- This repository hosts code that supports the testing infrastructure for the PyTorch organization. For example, this repo hosts the logic …☆112Updated this week
- Intel® Extension for MLIR. A staging ground for MLIR dialects and tools for Intel devices using the MLIR toolchain.☆156Updated this week
- Cosmic Tagging Network for Neutrino Physics☆13Aug 22, 2026Updated last month
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆17Updated this week
- ☆24Jun 12, 2023Updated 3 years ago
- ☆291Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, …☆2,709Updated this week
- ☆19Jul 13, 2026Updated 2 months ago
- ☆59Sep 14, 2026Updated last week
- SynapseAI Core is a reference implementation of the SynapseAI API running on Habana Gaudi☆46Feb 3, 2025Updated last year
- ☆16Jun 4, 2026Updated 3 months ago
- Developing multi platform gesture detector application by applying concepts learnt in Embedded Systems course on peripheral devices.☆21Dec 8, 2023Updated 2 years ago
- ☆21Jan 21, 2026Updated 8 months ago