End-to-end documentation to set up your own local & fully private LLM server on Debian. Equipped with chat, web search, RAG, model management, MCP servers, image generation, and TTS.
☆815Jun 30, 2026Updated 3 weeks ago
Alternatives and similar repositories for llm-server-docs
Users that are interested in llm-server-docs are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- User-friendly AI Interface (Supports Ollama, OpenAI API, ...)☆146,342Updated this week
- implementation of paint-by-example on comfyui☆11Apr 30, 2026Updated 2 months ago
- Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.☆176,579Updated this week
- Go manage your Ollama models☆1,821Updated this week
- A proxy server for multiple ollama instances with Key security☆640Apr 23, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Reliable model swapping for any local OpenAI/Anthropic compatible server - llama.cpp, vllm, etc☆5,103Updated this week
- ☆22Oct 19, 2024Updated last year
- Open‑WebUI Tools is a modular toolkit designed to extend and enrich your Open WebUI instance, turning it into a powerful AI workstation. …☆766Jun 30, 2026Updated 3 weeks ago
- my ai-roles like gpt claude gemini glm 豆包 扣子 comfyui also include some prompts in local ai, try to let user use their ai agent/workflow i…☆13Dec 14, 2024Updated last year
- Using image caption models to extract prompts in ComfyUI☆12May 21, 2025Updated last year
- A ComfyUI node prompt generator and CLIP encoder using AI provided by Ollama☆18Nov 29, 2024Updated last year
- Stop configuring your AI stack. Start using it. One command brings a complete pre-wired LLM stack with hundreds of services to explore.☆3,144Updated this week
- LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.☆47,746Updated this week
- flux☆10Aug 29, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Vane is an AI-powered answering engine.☆35,839Apr 11, 2026Updated 3 months ago
- LLM inference in C/C++☆121,220Updated this week
- Deprecated; Please use https://github.com/open-webui/desktop instead☆184May 27, 2024Updated 2 years ago
- Docker container for suno-ai bark model☆12Jun 26, 2023Updated 3 years ago
- API up your Ollama Server.☆202Mar 30, 2026Updated 3 months ago
- Unsloth is a local UI for training and running Gemma 4, Qwen3.6, DeepSeek, Kimi, GLM and other models.☆68,666Updated this week
- Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience☆63,683Updated this week
- Open-source NotebookLM alternative. Research the open web with live data, through one platform, API or MCP server. Join our Discord: http…☆15,295Updated this week
- Graphiti-based knowledge graph memory extensions for Open WebUI☆24May 11, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Run frontier AI locally.☆46,414Jun 23, 2026Updated 3 weeks ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆86,804Updated this week
- ☆19Oct 25, 2025Updated 8 months ago
- a lightweight, open-source blueprint for building powerful and scalable LLM chat applications☆28Jun 7, 2024Updated 2 years ago
- Open WebUI Desktop 🌐☆2,375May 6, 2026Updated 2 months ago
- comfyui的InternVL2插件,InternVL2是当前不错的开源多模态大语言模型,在文档vqa上表现很好☆13Aug 10, 2024Updated last year
- Claraverse is a opesource privacy focused ecosystem to replace ChatGPT, Claude, N8N, ImageGen with your own hosted llm, keys and compute.…☆3,837Updated this week
- Ligand-Receptor docking with AutoDock Vina☆12Mar 21, 2024Updated 2 years ago
- The fastest, litest AI Gateway. Rust core with Python SDK. Call 100+ LLM APIs in OpenAI (or native) format with cost tracking, guardrails…☆54,241Updated this week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- https://docs.openwebui.com☆809Updated this week
- beep boop 🤖 (experimental)☆118Jan 8, 2025Updated last year
- Fully Local Manus AI. No APIs, No $200 monthly bills. Enjoy an autonomous agent that thinks, browses the web, and code for the sole cost …☆26,666Jul 12, 2026Updated last week
- A minimal LLM chat app that runs entirely in your browser☆1,168Oct 12, 2025Updated 9 months ago
- A powerful document AI question-answering tool that connects to your local Ollama models. Create, manage, and interact with RAG systems f…☆1,095Aug 9, 2025Updated 11 months ago
- Pipelines: Versatile, UI-Agnostic OpenAI-Compatible Plugin Framework☆2,422Aug 18, 2025Updated 11 months ago
- Enchanted is iOS and macOS app for chatting with private self hosted language models such as Llama2, Mistral or Vicuna using Ollama.☆5,976Jul 7, 2026Updated 2 weeks ago