Community maintained hardware plugin for vLLM on Intel Gaudi
☆61Oct 7, 2026Updated this week
Alternatives and similar repositories for vllm-gaudi
Users that are interested in vllm-gaudi are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Jun 4, 2026Updated 4 months ago
- ☆19Jul 13, 2026Updated 2 months ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆91Sep 22, 2026Updated 2 weeks ago
- SGLang kernel library for Intel XPU☆36Updated this week
- DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.☆14Jan 8, 2026Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The vLLM XPU kernels for Intel GPU☆75Updated this week
- Easy and lightning fast training of 🤗 Transformers on Habana Gaudi processor (HPU)☆213Updated this week
- fzf-based test selection with pytest☆16Jan 26, 2026Updated 8 months ago
- Community maintained hardware plugin for vLLM on Spyre☆53Oct 1, 2026Updated last week
- Explore training for quantized models☆28Jul 12, 2025Updated last year
- vLLM plugin for RBLN NPU☆59Updated this week
- Intel® Extension for DeepSpeed* is an extension to DeepSpeed that brings feature support with SYCL kernels on Intel GPU(XPU) device. Note…☆65May 27, 2026Updated 4 months ago
- SYCL* Templates for Linear Algebra (SYCL*TLA) - SYCL based CUTLASS implementation for Intel GPUs☆88Sep 23, 2026Updated 2 weeks ago
- Containerization and cloud native suite for OPEA☆74Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆19Updated this week
- PM Workshop China☆10Apr 11, 2019Updated 7 years ago
- ☆41Sep 29, 2026Updated last week
- ☆15May 8, 2025Updated last year
- Very simple and stupid TCP/IP stack written in C☆10Mar 25, 2016Updated 10 years ago
- helm charts for deploying models with llm-d☆32Sep 26, 2026Updated last week
- SPDK fork of nvme-cli. No longer supported - use standard nvme-cli with SPDK nvme CUSE instead. See https://spdk.io/doc/nvme.html#nvme_…☆15Apr 10, 2024Updated 2 years ago
- Dashboard for InferenceX™, Open Source Continuous Inference | InferenceX 仪表板☆46Updated this week
- Community maintained hardware plugin for vLLM on AWS Neuron☆52Aug 17, 2026Updated last month
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆115Updated this week
- A demo to showcase the power of ServiceWorker and the Physical Web☆17Apr 26, 2015Updated 11 years ago
- This is a clone of an SVN repository at http://pagecache-mangagement.googlecode.com/svn/trunk. It had been cloned by http://svn2github.co…☆10May 23, 2013Updated 13 years ago
- ☆18Oct 2, 2026Updated last week
- Workload Services Framework (WSF) is a benchmarking framework on Intel(R) Xeon(R) Platforms☆59Updated this week
- ☆26Updated this week
- Simple Robin Hood hash table implemented using C macros☆15Feb 7, 2025Updated last year
- ☆11Jan 7, 2023Updated 3 years ago
- Software kit for Qualcomm Cloud AI 100☆19Sep 3, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- GPU Functional Descriptor for memory access☆35May 24, 2026Updated 4 months ago
- ☆13Updated this week
- Memory Address Tracer☆14Jun 12, 2020Updated 6 years ago
- Falcon: Fast OLTP Engine for Persistent Cache and Non-Volatile Memory☆11Nov 1, 2023Updated 2 years ago
- SQL Server on OpenShift Workshop☆16Jun 27, 2023Updated 3 years ago
- A good book, uploaded for myself and those who are interested.☆14May 6, 2022Updated 4 years ago
- Tools and pipelines for automated LLM performance evaluation☆15May 20, 2026Updated 4 months ago