A fork of vLLM enabling Pascal architecture GPUs
☆34Feb 21, 2025Updated last year
Alternatives and similar repositories for vllm-pascal
Users that are interested in vllm-pascal are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The main repository for building Pascal-compatible versions of ML applications and libraries.☆213Aug 23, 2025Updated 10 months ago
- CI scripts designed to build a Pascal-compatible version of vLLM.☆13Aug 10, 2024Updated last year
- Stable Diffusion and Flux in pure C/C++☆25Jul 14, 2026Updated last week
- A daemon that automatically manages the performance states of NVIDIA GPUs.☆136Feb 24, 2026Updated 4 months ago
- Model-agnostic code memory MCP server. Budget-aware graph retrieval for AI agents. Sub-millisecond queries, token budgeting, deterministi…☆23May 18, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- golang live - rtmp - httpflv - hls☆11Jul 10, 2026Updated last week
- Yet Another (LLM) Web UI, made with Gemini☆12Dec 25, 2024Updated last year
- ☆14Sep 24, 2024Updated last year
- A PyTorch native platform for training generative AI models☆17Jun 30, 2026Updated 3 weeks ago
- LLM proxy specialized in analyzing harness traffic to catch cache invalidation.☆37Updated this week
- The TinyLlama project is an open endeavor to pretrain a 1.1B Llama model on 3 trillion tokens.☆14Mar 30, 2024Updated 2 years ago
- Streamlit OpenAI app to chat with custom text documents of all kinds☆13Apr 11, 2026Updated 3 months ago
- marlin_v100 是一个从 vLLM 主树中提取出来的最小 Marlin 独立开发工作区,聚焦于 Marlin dense 与 Marlin MoE 的源码开发、最小构建和轻量验证。它保留了核心 CUDA/C++ 实现、最小 Python 薄封装、生成器测试与主树回写…☆19Jul 2, 2026Updated 2 weeks ago
- Exploring the minimal architecture required for coherent English language generation.☆14Jun 11, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Stable Diffusion in pure C/C++☆18Jun 21, 2026Updated last month
- An arbitrary precision library written in C which supports fractional datum. Arbitraire is fully portable to all linux, mac and netbsd/fr…☆17Jul 14, 2023Updated 3 years ago
- Firmware for Xilinx Platform Cable 1 USB Jtag adapter☆10Jul 24, 2016Updated 9 years ago
- ☆14Aug 25, 2024Updated last year
- ☆11Jan 28, 2024Updated 2 years ago
- Hardware for ESP8266 based walkie talkie (doesn't work particularly well)☆10Jul 16, 2018Updated 8 years ago
- 一个机器人插件 用于管理群聊☆17Nov 24, 2024Updated last year
- Startup app with some base features enabled☆13Mar 13, 2023Updated 3 years ago
- An OpenAI API compatible images server to generate or manipulate images.☆18Feb 2, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Local LLM Inference Speed Test Tool☆116Jul 2, 2026Updated 2 weeks ago
- Oobabooga "Hello World" API example for node.js with Express☆13Jul 2, 2023Updated 3 years ago
- A collection of block script for Automation Blocks for Premiere Pro☆16Updated this week
- 机器人阿洛娜的额外功能插件☆13Oct 25, 2024Updated last year
- Orchestrator managing Erlang containers and routing Erlang Distribution through itself.☆12Dec 19, 2025Updated 7 months ago
- 公共API服务,Power by ThinkJS☆20May 11, 2025Updated last year
- minimal fill-in-the-middle autocomplete for vscode/codium, for use with llama.cpp infill or any openai-compatible server☆16May 10, 2026Updated 2 months ago
- A high-throughput and memory-efficient inference and serving engine for LLMs - Optimized for AMD gfx906 GPUs, e.g. Radeon VII / MI50 / MI…☆73Jun 23, 2026Updated 3 weeks ago
- ☆13Nov 29, 2025Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆16Dec 16, 2024Updated last year
- 适用于hoshino的冰祈插件集合☆20Jun 15, 2023Updated 3 years ago
- Code for short tutorials☆17Aug 9, 2024Updated last year
- ☆18Sep 4, 2024Updated last year
- Persys desktop. Electron based application to access your Persys server.☆16May 16, 2025Updated last year
- 使用Live2D与GPT-sovits的AI全自动直播一站式解决方案。☆18Jun 10, 2026Updated last month
- ARK (Automated Resource Knowledge-base) revolutionizes personal computing by creating an open-source, decentralized assistant. Harnessing…☆17Updated this week