llama-swap + a minimal ollama compatible api
☆62May 26, 2026Updated 3 months ago
Alternatives and similar repositories for llama-swappo
Users that are interested in llama-swappo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Llama.cpp runner/swapper and proxy that emulates LMStudio / Ollama backends☆60Aug 21, 2025Updated last year
- ☆58Oct 10, 2025Updated 10 months ago
- A proxy that hosts multiple single-model runners such as LLama.cpp and vLLM☆12May 30, 2025Updated last year
- A simple, "Ollama-like" tool for managing and running GGUF language models from your terminal.☆25Jan 2, 2026Updated 7 months ago
- ☆99Mar 28, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- An fully autonomous agent that accesses the browser and performs tasks.☆18Aug 17, 2026Updated last week
- Yet Another (LLM) Web UI, made with Gemini☆12Dec 25, 2024Updated last year
- ☆18Jul 1, 2025Updated last year
- Reliable model swapping for any local OpenAI/Anthropic compatible server - llama.cpp, vllm, etc☆5,510Updated this week
- Chat WebUI is an easy-to-use user interface for interacting with AI, and it comes with multiple useful built-in tools such as web search …☆53Feb 10, 2026Updated 6 months ago
- Run Orpheus 3B Locally with Gradio UI, Standalone App☆25Apr 1, 2025Updated last year
- ☆93Jul 7, 2025Updated last year
- VellumForge2 is a Golang CLI for generating high-quality Direct Preference Optimization datasets via a hierarchical prompt pipeline with …☆21Feb 15, 2026Updated 6 months ago
- Bookmarklet to pull and run hugging face GGUF models in Ollama☆18Oct 17, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ACE-Step: A Step Towards Music Generation Foundation Model☆50May 20, 2025Updated last year
- Llama Server Launcher (llama.cpp/ik_llama) GUI☆124Jul 22, 2026Updated last month
- A custom LiteLLM provider enabling local execution of Hugging Face models with streaming, quantization, and async support☆30Jun 22, 2025Updated last year
- A command-line client for Bluesky☆20May 29, 2026Updated 3 months ago
- ☆20Sep 28, 2024Updated last year
- CompChomper is a framework for measuring how LLMs perform at code completion.☆21Apr 29, 2025Updated last year
- ☆21Jan 25, 2025Updated last year
- vLLM Docker Container for Qwen3.6 27b☆50Jun 15, 2026Updated 2 months ago
- Electron speech-to-speech app for your voice calls based on 100% locally run AI models☆35Jul 23, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Fast LLM swapping with sleep/wake support, compatible with vllm, llama.cpp, etc. llama-swap fork.☆54Apr 5, 2026Updated 4 months ago
- Scripts and tools for optimizing quantizations in llama.cpp with GGUF imatrices.☆19Jan 10, 2025Updated last year
- Spec-driven iterative development companion CLI for OpenCode.☆18Jun 7, 2026Updated 2 months ago
- A TTS model capable of generating ultra-realistic dialogue in one pass.☆32May 1, 2025Updated last year
- A bare-bones GUI application for the local inference engine, llama.cpp. Built-in TPE optimiser to find the best flags for your system☆18Jul 3, 2026Updated last month
- A local-first web search agent☆30Jun 20, 2026Updated 2 months ago
- ☆16Dec 16, 2024Updated last year
- 🔥 Alternative to Ollama — multi-model serving with sub-ms model switching · CPU-only 20B inference for Edge AI · llama.cpp + stablediffu…☆40Aug 22, 2026Updated last week
- BROKEN REPO. DO NOT USE UNDER ANY CIRCUMSTANCES☆21Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Teaching AI to play the classic text adventure Zork using Large Language Models☆36Apr 5, 2026Updated 4 months ago
- Newman reporter allowing to decorate pull request with postman collection results.☆10Nov 2, 2023Updated 2 years ago
- SPLAA is an AI assistant framework that utilizes voice recognition, text-to-speech, and tool-calling capabilities to provide a conversati…☆29May 6, 2025Updated last year
- Running Microsoft's BitNet inference framework via FastAPI, Uvicorn and Docker.☆41Jul 2, 2025Updated last year
- Visual Tagger is a JavaScript tool that visually highlights HTML elements for AIs, aiding in identifying interactive components on web pa…☆12Oct 28, 2024Updated last year
- Linux & Powershell scripts to easily set up and run the Qwen 3.5 series locally on Windows and Linux with llama.cpp.☆104Aug 22, 2026Updated last week
- A minimal CLI tool for piping anything into an LLM.☆21Jan 1, 2026Updated 7 months ago