☆429Jul 16, 2026Updated this week
Alternatives and similar repositories for llm-scaler
Users that are interested in llm-scaler are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Inference engine for Intel devices. Serve LLMs, VLMs, Whisper, Kokoro-TTS, Embedding and Rerank models over OpenAI endpoints.☆486Updated this week
- ☆180Updated this week
- The vLLM XPU kernels for Intel GPU☆55Updated this week
- ☆54Updated this week
- SGLang kernel library for Intel XPU☆27Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆97Updated this week
- AI PC starter app for doing AI image creation, image stylizing, and chatbot on a PC powered by an Intel® Arc™ GPU.☆937Updated this week
- A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support…☆1,530Updated this week
- Run Generative AI models with simple C++/Python API and using OpenVINO Runtime☆556Updated this week
- 🤗 Optimum Intel: Accelerate inference with Intel optimization tools☆606Updated this week
- llama-benchy - llama-bench style benchmarking tool for all backends☆580Jul 10, 2026Updated last week
- Cache-DiT Node for Comfyui☆295Apr 15, 2026Updated 3 months ago
- The AI PC Application Installer provides a unified way to set up Intel AI PC development environments.☆40May 21, 2026Updated 2 months ago
- This repository contains Dockerfiles, scripts, yaml files, Helm charts, etc. used to scale out AI containers with versions of TensorFlow …☆79May 27, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆150Updated this week
- With OpenVINO Test Drive, users can run large language models (LLMs) and models trained by Intel Geti on their devices, including AI PCs …☆39Mar 12, 2026Updated 4 months ago
- Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V,…☆8,864Jan 28, 2026Updated 5 months ago
- vLLM for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆24Jul 2, 2026Updated 2 weeks ago
- OpenVINO™ is an open source toolkit for optimizing and deploying AI inference☆10,549Updated this week
- Intel® NPU (Neural Processing Unit) Driver☆437Jun 12, 2026Updated last month
- ☆22Jun 10, 2026Updated last month
- Intel® AI Builder (SuperClaw, SuperBuilder)☆196Updated this week
- Edge Insights for Vision (eiv) is a package that helps to auto install Intel® GPU drivers and setup environment for Inference application…☆22Sep 29, 2025Updated 9 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Open-source cross-modal and multimodal prompt injection test suite. 250,000+ attack payloads across text, image, document, and audio moda…☆64Jul 12, 2026Updated last week
- A Python package for extending the official PyTorch that can easily obtain performance on Intel platform☆2,014Mar 30, 2026Updated 3 months ago
- ☆17Updated this week
- Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https…☆4,997Updated this week
- Messy repo filled with messy tests about hardware and LLMs. Built for me, public for you.☆41Jun 1, 2026Updated last month
- Fully automated installation scripts for ComfyUI optimized for Intel Arc GPUs (A-Series) and Intel Core Ultra iGPUs with XPU backend, Tri…☆154Feb 10, 2026Updated 5 months ago
- OpenVINO LLM Benchmark☆11Dec 7, 2023Updated 2 years ago
- Performance optimized libraries, microservices, and tools to support the development of Edge AI applications.☆152Updated this week
- Large Language Model Text Generation Inference on Habana Gaudi☆34Mar 20, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ONNX Runtime: cross-platform, high performance scoring engine for ML models☆88Updated this week
- AI plays Doom — pit Vision Language Models against demons and each other. Solo scenarios, deathmatch arena, 1-4 agents with any OpenAI-co…☆20Mar 12, 2026Updated 4 months ago
- llama.cpp fork with additional SOTA quants and improved performance☆2,943Updated this week
- FORK of VLLM for AMD MI25/50/60. A high-throughput and memory-efficient inference and serving engine for LLMs☆71May 4, 2025Updated last year
- Get aid from local LLMs right in your PowerShell☆16May 2, 2025Updated last year
- Enable true multi gpu capability in Comfy UI using XDiT XFuser and FSDP managed by Ray☆367Updated this week
- Generate a llama-quantize command to copy the quantization parameters of any GGUF☆34Apr 20, 2026Updated 3 months ago