☆536Sep 17, 2026Updated this week
Alternatives and similar repositories for llm-scaler
Users that are interested in llm-scaler are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Inference engine for Intel devices. Serve LLMs, VLMs, Whisper, Kokoro-TTS, Embedding and Rerank models over OpenAI endpoints.☆525Updated this week
- ☆195Updated this week
- Makes Intel Arc Pro B70 GPUs actually fast on Ubuntu Server. 11 llama.cpp cherry-picks that fix the big B70 bugs (MoE slot-init SEGV, Q8_…☆25May 10, 2026Updated 4 months ago
- ☆59Updated this week
- SGLang kernel library for Intel XPU☆35Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- AI PC starter app for doing AI image creation, image stylizing, and chatbot on a PC powered by an Intel® Arc™ GPU.☆982Updated this week
- A SOTA quantization toolkit for high-accuracy low-bit LLM inference|简洁且高效的量化工具包☆1,621Updated this week
- Run Generative AI models with simple C++/Python API and using OpenVINO Runtime☆588Updated this week
- 🤗 Optimum Intel: Accelerate inference with Intel optimization tools☆622Updated this week
- llama-benchy - llama-bench style benchmarking tool for all backends☆719Jul 10, 2026Updated 2 months ago
- The AI PC Application Installer provides a unified way to set up Intel AI PC development environments.☆43Jul 23, 2026Updated last month
- Cache-DiT Node for Comfyui☆308Aug 4, 2026Updated last month
- This repository contains Dockerfiles, scripts, yaml files, Helm charts, etc. used to scale out AI containers with versions of TensorFlow …☆79Sep 8, 2026Updated last week
- Fully automated installation scripts for ComfyUI optimized for Intel Arc GPUs (A-Series) and Intel Core Ultra iGPUs with XPU backend, Tri…☆162Feb 10, 2026Updated 7 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆157Updated this week
- With OpenVINO Test Drive, users can run large language models (LLMs) and models trained by Intel Geti on their devices, including AI PCs …☆41Sep 1, 2026Updated 2 weeks ago
- Developer kits reference setup scripts for various kinds of Intel platforms and GPUs☆54Updated this week
- Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V,…☆8,853Jan 28, 2026Updated 7 months ago
- vLLM for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆39Aug 25, 2026Updated 3 weeks ago
- Community maintained hardware plugin for vLLM on Intel Gaudi☆57Updated this week
- OpenVINO™ is an open source toolkit for optimizing and deploying AI inference☆10,872Updated this week
- Intel® NPU (Neural Processing Unit) Driver☆459Sep 10, 2026Updated last week
- ☆28Jun 10, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Open-source cross-modal and multimodal prompt injection test suite. 250,000+ attack payloads across text, image, document, and audio moda…☆74Jul 22, 2026Updated last month
- Intel® AI Builder (SuperClaw, SuperBuilder)☆249Updated this week
- Edge Insights for Vision (eiv) is a package that helps to auto install Intel® GPU drivers and setup environment for Inference application…☆22Sep 29, 2025Updated 11 months ago
- Ubuntu Server edition: automated setup script for Intel Arc Pro B70 GPU LLM inference server with vLLM tensor parallelism. 140 tok/s on 2…☆34Apr 26, 2026Updated 4 months ago
- Messy repo filled with messy tests about hardware and LLMs. Built for me, public for you.☆49Aug 17, 2026Updated last month
- Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https…☆5,747Updated this week
- OpenVINO LLM Benchmark☆11Dec 7, 2023Updated 2 years ago
- Large Language Model Text Generation Inference on Habana Gaudi☆34Mar 20, 2025Updated last year
- Explore our open source AI portfolio! Develop, train, and deploy your AI solutions with performance- and productivity-optimized tools fro…☆78Mar 27, 2026Updated 5 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Libraries, microservices, tools, and other reference software, supporting development of performance-optimized Edge AI applications.☆163Updated this week
- ONNX Runtime: cross-platform, high performance scoring engine for ML models☆91Updated this week
- AI plays Doom — pit Vision Language Models against demons and each other. Solo scenarios, deathmatch arena, 1-4 agents with any OpenAI-co…☆20Mar 12, 2026Updated 6 months ago
- A rapid prototyping library for connected, autonomous vehicles developed at RISE Dependable Transport Systems☆14Aug 18, 2026Updated last month
- llama.cpp fork with additional SOTA quants and improved performance☆3,244Updated this week
- FORK of VLLM for AMD MI25/50/60. A high-throughput and memory-efficient inference and serving engine for LLMs☆70May 4, 2025Updated last year
- Samples running deep learning models on Intel GPU Arc A770☆15Jul 4, 2024Updated 2 years ago