A fork of vLLM enabling Pascal architecture GPUs
☆37Feb 21, 2025Updated last year
Alternatives and similar repositories for vllm-pascal
Users that are interested in vllm-pascal are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The main repository for building Pascal-compatible versions of ML applications and libraries.☆219Aug 23, 2025Updated last year
- CI scripts designed to build a Pascal-compatible version of vLLM.☆13Aug 10, 2024Updated 2 years ago
- Stable Diffusion and Flux in pure C/C++☆26Updated this week
- Personal voice assistant, with voice interruption and Twilio support☆18Feb 24, 2025Updated last year
- GPU Power and Performance Manager☆73Oct 13, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Model-agnostic code memory MCP server. Budget-aware graph retrieval for AI agents. Sub-millisecond queries, token budgeting, deterministi…☆24May 18, 2026Updated 4 months ago
- Yet Another (LLM) Web UI, made with Gemini☆12Dec 25, 2024Updated last year
- Transform your pdfs into anki flashcard with gpt☆12Jun 10, 2024Updated 2 years ago
- ☆20Jan 11, 2026Updated 8 months ago
- KoboldCpp Smart Launcher with GPU Layer and Tensor Override Tuning☆30May 18, 2025Updated last year
- xllamacpp - a Python wrapper of llama.cpp☆87Sep 26, 2026Updated 2 weeks ago
- 🎥 AI-powered webinar platform with live streaming, smart breakout rooms, and autonomous sales agents.☆13Jun 21, 2025Updated last year
- ☆14Sep 24, 2024Updated 2 years ago
- vLLM for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI60☆435Feb 20, 2026Updated 7 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- LLM proxy specialized in analyzing harness traffic to catch cache invalidation.☆58Sep 27, 2026Updated 2 weeks ago
- The TinyLlama project is an open endeavor to pretrain a 1.1B Llama model on 3 trillion tokens.☆14Mar 30, 2024Updated 2 years ago
- marlin_v100 是一个从 vLLM 主树中提取出来的最小 Marlin 独立开发工作区,聚焦于 Marlin dense 与 Marlin MoE 的源码开发、最小构建和轻量验证。它保留了核心 CUDA/C++ 实现、最小 Python 薄封装、生成器测试与主树回写…☆25Jul 2, 2026Updated 3 months ago
- ☆92Updated this week
- Conversion script adapting vicuna dataset into alpaca format for use with oobabooga's trainer☆12Jun 21, 2023Updated 3 years ago
- ☆18Dec 7, 2023Updated 2 years ago
- Deploy Jupyter Notebook to AWS Lambda☆16Nov 18, 2020Updated 5 years ago
- An arbitrary precision library written in C which supports fractional datum. Arbitraire is fully portable to all linux, mac and netbsd/fr…☆17Jul 14, 2023Updated 3 years ago
- ☆14Aug 25, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Hardware for ESP8266 based walkie talkie (doesn't work particularly well)☆10Jul 16, 2018Updated 8 years ago
- 一个机器人插件 用于管理群聊☆17Nov 24, 2024Updated last year
- NoneBot 黑白名单☆15May 7, 2023Updated 3 years ago
- An OpenAI API compatible images server to generate or manipulate images.☆18Feb 2, 2025Updated last year
- Oobabooga "Hello World" API example for node.js with Express☆13Jul 2, 2023Updated 3 years ago
- PrompFlower 1.0 is a command-line tool that generates AI prompts entirely locally and offline using a local AI engine via Ollama.☆12Updated this week
- A set of bloom filter implementations in pure coffee/javascipt.☆28Jun 19, 2013Updated 13 years ago
- 机器人阿洛娜的额外功能插件☆13Oct 25, 2024Updated last year
- Local LLM Inference Speed Test Tool☆233Sep 2, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A Swiftui Drag and Drop view that is easy to implement☆21May 8, 2023Updated 3 years ago
- Forget about 3D printer spaghetti with Forgetti, an alternative to traditional 3D print failure detection software.☆20May 23, 2026Updated 4 months ago
- minimal fill-in-the-middle autocomplete for vscode/codium, for use with llama.cpp infill or any openai-compatible server☆27May 10, 2026Updated 5 months ago
- ☆16Dec 16, 2024Updated last year
- Run a 105 GB AI model on a Mac that can't hold it. Slotstream streams Qwen3.8-Flash-Next (125B mixture of experts) from your SSD and cach…☆454Updated this week
- At the start of the coronavirus many people were scrambling to gather computer parts. As a result the prices of GPU's (Graphics Processin…☆21Dec 10, 2022Updated 3 years ago
- This is a Next.js based reimplementation of the Bolt OSS project by Stackblitz. It utilizes React, server-side rendering, and other moder…☆21Nov 20, 2025Updated 10 months ago