vLLM Docker Container for the latest Qwen 27b
☆52Sep 27, 2026Updated this week
Alternatives and similar repositories for qwen-27b-docker
Users that are interested in qwen-27b-docker are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆118Apr 28, 2026Updated 5 months ago
- ☆45Sep 20, 2026Updated last week
- Local-first AI workflow orchestration for chaining models, agents, tools, and scripts into repeatable workflows.☆34Aug 10, 2026Updated last month
- Community recipes for serving LLMs on RTX 3090/4090/5090 CUDA gpus. Multi-engine (vLLM, llama.cpp, ik_llama) and model-agnostic. Currentl…☆2,333Updated this week
- Spec-driven iterative development companion CLI for OpenCode.☆19Jun 7, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Qwen3.6-27B on dual RTX 3090 — TP=2 recipe, vLLM nightly, MTP + fp8 KV, validated for concurrent serving☆58Apr 28, 2026Updated 5 months ago
- SNDR Core Engine (Genesis) — vLLM runtime patch-overlay for Qwen3.6 + Gemma4 on consumer NVIDIA (Ampere sm_86, 2× A5000/3090). Qwen3.6-35…☆133Updated this week
- pebkac Chrome Nonautomation - A Local LLM-Driven Web Co-Browser using Smolagents, Zendriver, Trafilatura.☆80Apr 14, 2026Updated 5 months ago
- ☆43May 4, 2026Updated 4 months ago
- VellumForge2 is a Golang CLI for generating high-quality Direct Preference Optimization datasets via a hierarchical prompt pipeline with …☆22Feb 15, 2026Updated 7 months ago
- llama.cpp/ik_llama.cpp launcher: loads big MoE models across mismatched multi-GPU rigs by exact VRAM math.☆280Updated this week
- ☆20Jul 21, 2026Updated 2 months ago
- llama-swap + a minimal ollama compatible api☆63Updated this week
- StatelessChatUI is a lightweight, single-file, OpenAI-compatible chat frontend. It’s built to be portable and privacy-friendly: no accoun…☆19Dec 18, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Headless Matrix WebRTC voice AND video agent — auto-answers calls, bridges audio to any AI agent via PipeWire, optional camera-frame visi…☆20Jun 28, 2026Updated 3 months ago
- A comprehensive collection of atomic Python scripts for learning AI agent development from scratch☆20Jul 17, 2026Updated 2 months ago
- DGX Spark research and tests - containers, benchmarks, and investigation notes for running models on GB10 (SM 12.1)☆28Aug 6, 2026Updated last month
- Information Processing Evaluation for Large Language Models☆70Updated this week
- ☆26Aug 3, 2026Updated 2 months ago
- TradingAgents with a polished local GUI — install once, run multi-agent LLM stock analyses locally.☆34May 25, 2026Updated 4 months ago
- Nemotron Speech ASR Docker deployment☆30Jun 7, 2026Updated 3 months ago
- Modular, agentic framework and MCP platform you self-host. Build and run your own AI agents, connect Claude Code, Claude Desktop, or Curs…☆23Updated this week
- A simple Gradio WebUI for loading/unloading models and loras in tabbyAPI.☆20Nov 21, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Docker images for LLM inference (SGLang + vLLM) on NVIDIA Blackwell GPUs (SM120, CUDA 13.2)☆83Updated this week
- GPU monitor for Linux terminal supporting single or multiple gpu's in realtime☆20Jul 19, 2026Updated 2 months ago
- ☆13Jun 18, 2024Updated 2 years ago
- ☆12May 30, 2025Updated last year
- Crashbench is a LLM benchmark to measure bug-finding and reporting capabilities of LLMs☆14Aug 7, 2026Updated last month
- Specialized fork for (relatively) fast single-GPU inference (in CUDA) using large MoE models that don't fit fully into VRAM☆17May 6, 2026Updated 4 months ago
- Stateless LLM runtime that dynamically routes, loads, executes, and unloads models per request with bounded VRAM caching and intelligent …☆42Apr 12, 2026Updated 5 months ago
- ☆10Jan 23, 2025Updated last year
- Deploy Sourcegraph on Kubernetes using Helm☆19Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆16Jun 16, 2026Updated 3 months ago
- Self-hosted Codex, Your Codex in Your Container☆16Mar 26, 2026Updated 6 months ago
- A Fabric mod that renders emojis and emotes from 7TV, BTTV and FFZ in the chat.☆18Apr 15, 2025Updated last year
- Install Comfyui Custom NOdes Easy☆19Feb 1, 2026Updated 8 months ago
- Mod Manager for World of Tanks Blitz (PC)☆10Mar 6, 2018Updated 8 years ago
- Generate Duolingo-style quiz courses from PDFs with spaced repetition, adaptive difficulty, and tutor chat.☆18Apr 6, 2026Updated 5 months ago
- code for Towards Data Science article on prompt-loss-weight☆11Jun 4, 2025Updated last year