☆16Jun 4, 2026Updated 3 months ago
Alternatives and similar repositories for vllm-hpu-extension
Users that are interested in vllm-hpu-extension are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Community maintained hardware plugin for vLLM on Intel Gaudi☆57Updated this week
- ☆19Jul 13, 2026Updated 2 months ago
- A PyTorch native platform for training generative AI models☆17Jun 30, 2026Updated 2 months ago
- Large Language Model Text Generation Inference on Habana Gaudi☆34Mar 20, 2025Updated last year
- SynapseAI Core is a reference implementation of the SynapseAI API running on Habana Gaudi☆46Feb 3, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆19Jul 24, 2025Updated last year
- Efficient kernel for RMS normalization with fused operations, includes both forward and backward passes, compatibility with PyTorch.☆13Jun 5, 2024Updated 2 years ago
- SPDK fork of nvme-cli. No longer supported - use standard nvme-cli with SPDK nvme CUSE instead. See https://spdk.io/doc/nvme.html#nvme_…☆15Apr 10, 2024Updated 2 years ago
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- This is a clone of an SVN repository at http://pagecache-mangagement.googlecode.com/svn/trunk. It had been cloned by http://svn2github.co…☆10May 23, 2013Updated 13 years ago
- ☆11Jan 7, 2023Updated 3 years ago
- kernelboard is the webapp for https://www.gpumode.com☆19Sep 11, 2026Updated last week
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- Full End-to-End examples showing how to use First-gen Gaudi and Gaudi2 in common use cases☆13Dec 2, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Reference models for Intel(R) Gaudi(R) AI Accelerator☆172Jan 8, 2026Updated 8 months ago
- Runtime for creating OS to use A9N Microkernel☆11Sep 7, 2026Updated last week
- SQL Server on OpenShift Workshop☆15Jun 27, 2023Updated 3 years ago
- Tools and pipelines for automated LLM performance evaluation☆15May 20, 2026Updated 3 months ago
- ☆18Apr 20, 2018Updated 8 years ago
- Your finetuned model's back to its original safety standards faster than you can say "SafetyLock"!☆11Oct 16, 2024Updated last year
- ☆14May 25, 2022Updated 4 years ago
- A tool that converts clang generated assembly code into Go ASM.☆18Aug 13, 2025Updated last year
- The official implement of paper S2-VER: Semi-Supervised Visual Emotion Recognition☆11Apr 28, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A conda-smithy repository for ambertools.☆12Jul 15, 2026Updated 2 months ago
- OpenVINO LLM Benchmark☆11Dec 7, 2023Updated 2 years ago
- This repo contains codes and references for Udemy course by Aman on "Agentic AI For Beginner"☆36Aug 24, 2026Updated 3 weeks ago
- **ASCM4ABSA** - Our code and proposed data for NLPCC 2022 paper titled "Aspect-specific Context Modeling for Aspect-based Sentiment Analy…☆12Mar 26, 2023Updated 3 years ago
- NeurIPS 2024 + NeuroAI and SSL Workshops (Oral)☆11Dec 6, 2024Updated last year
- [EMNLP 2022] Official Pytorch implementation for "Tiny-NewsRec: Efficient and Effective PLM-based News Recommendation"☆19Sep 18, 2023Updated 3 years ago
- Experiments with representation engineering☆14Feb 28, 2024Updated 2 years ago
- ☆19Jan 28, 2026Updated 7 months ago
- This repository contains Dockerfiles, scripts, yaml files, Helm charts, etc. used to scale out AI containers with versions of TensorFlow …☆79Sep 8, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆19Updated this week
- [NLPCC 2024] Shared Task 10: Regulating Large Language Models☆14Jun 12, 2024Updated 2 years ago
- Official PyTorch implementation of CD-MOE☆12Mar 18, 2026Updated 6 months ago
- Data Files for "Deep diversification of an AAV capsid protein by machine learning"☆18Mar 9, 2021Updated 5 years ago
- Residual vector quantization for KV cache compression in large language model☆12Oct 22, 2024Updated last year
- Clustered Compositional Embeddings☆13Oct 25, 2023Updated 2 years ago
- ☆34Oct 2, 2024Updated last year