☆16Jun 4, 2026Updated 4 months ago
Alternatives and similar repositories for vllm-hpu-extension
Users that are interested in vllm-hpu-extension are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A high-throughput and memory-efficient inference and serving engine for LLMs☆91Sep 22, 2026Updated 2 weeks ago
- Community maintained hardware plugin for vLLM on Intel Gaudi☆61Updated this week
- ☆19Jul 13, 2026Updated 2 months ago
- A PyTorch native platform for training generative AI models☆17Jun 30, 2026Updated 3 months ago
- SynapseAI Core is a reference implementation of the SynapseAI API running on Habana Gaudi☆46Feb 3, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆19Jul 24, 2025Updated last year
- PM Workshop China☆10Apr 11, 2019Updated 7 years ago
- Benchmark Suite Invocation Scripting☆11Mar 16, 2022Updated 4 years ago
- SPDK fork of nvme-cli. No longer supported - use standard nvme-cli with SPDK nvme CUSE instead. See https://spdk.io/doc/nvme.html#nvme_…☆15Apr 10, 2024Updated 2 years ago
- Tensor parallelism is all you need. Run LLMs on an AI cluster at home using any device. Distribute the workload, divide RAM usage, and in…☆18Nov 11, 2024Updated last year
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- This is a clone of an SVN repository at http://pagecache-mangagement.googlecode.com/svn/trunk. It had been cloned by http://svn2github.co…☆10May 23, 2013Updated 13 years ago
- ☆15Mar 3, 2025Updated last year
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Runtime for creating OS to use A9N Microkernel☆11Sep 7, 2026Updated last month
- Memory Address Tracer☆14Jun 12, 2020Updated 6 years ago
- A good book, uploaded for myself and those who are interested.☆14May 6, 2022Updated 4 years ago
- Tools and pipelines for automated LLM performance evaluation☆15May 20, 2026Updated 4 months ago
- A Prot paper related materials☆11Sep 5, 2022Updated 4 years ago
- Your finetuned model's back to its original safety standards faster than you can say "SafetyLock"!☆11Oct 16, 2024Updated last year
- Operator installing the Telemetry stack in a Kubernetes cluster and installing the metrics and alerts☆18Nov 3, 2023Updated 2 years ago
- A repository of Dockerfiles, scripts, yaml files, Helm Charts, etc. used to build and scale the sample AI workflows with python, kubernet…☆12Feb 22, 2024Updated 2 years ago
- ☆11Aug 5, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆14May 25, 2022Updated 4 years ago
- A tool that converts clang generated assembly code into Go ASM.☆19Aug 13, 2025Updated last year
- ☆14Aug 20, 2026Updated last month
- The official implement of paper S2-VER: Semi-Supervised Visual Emotion Recognition☆11Apr 28, 2024Updated 2 years ago
- ☆17Aug 21, 2023Updated 3 years ago
- RISC-V Software Porting and Optimization Championship☆18Oct 2, 2026Updated last week
- ☆12Sep 23, 2024Updated 2 years ago
- Evaluate gpt-4o on CLIcK (Korean NLP Dataset)☆20May 18, 2024Updated 2 years ago
- Experiments with representation engineering☆14Feb 28, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆19Sep 25, 2026Updated 2 weeks ago
- A repo for LLM jailbreak☆14Sep 5, 2023Updated 3 years ago
- Reference implementation of models from Nyonic Model Factory☆12May 13, 2024Updated 2 years ago
- Residual vector quantization for KV cache compression in large language model☆12Oct 22, 2024Updated last year
- Clustered Compositional Embeddings☆13Oct 25, 2023Updated 2 years ago
- ☆22Jan 25, 2023Updated 3 years ago
- ☆34Oct 2, 2024Updated 2 years ago