Simple node proxy for llama-server that enables MCP use
☆19May 10, 2025Updated last year
Alternatives and similar repositories for llama-server_mcp_proxy
Users that are interested in llama-server_mcp_proxy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16May 8, 2025Updated last year
- ☆16Dec 16, 2024Updated last year
- Llama.cpp runner/swapper and proxy that emulates LMStudio / Ollama backends☆60Aug 21, 2025Updated 11 months ago
- 🔍📃 LLM-powered PDF Table Extractor☆19Jun 26, 2025Updated last year
- Anthropic's Contextual Retrieval implementation with visual chunk comparison. Preview context enrichment before/after embedding.☆30Sep 25, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Vigil- API security☆16Jan 8, 2026Updated 6 months ago
- The most feature-complete local AI workstation. Multi-GPU inference, integrated Stable Diffusion + ADetailer, voice cloning, research-gra…☆64Updated this week
- This project contains the original white paper for Language Construct Modeling (LCM) v1.13, authored by Vincent Shing Hin Chong. It intro…☆15Jul 23, 2025Updated last year
- ☆18Jul 1, 2025Updated last year
- a browser gui for nvidia smi☆21Mar 17, 2025Updated last year
- ☆20Jul 4, 2025Updated last year
- ☆15Mar 18, 2026Updated 4 months ago
- Scan any agent skill — authored or compiled — for prompt injection, secrets, and malicious execution before it touches your agent. Or com…☆32Jul 17, 2026Updated last week
- Local RAG as a simple CLI, for standalone use or as a gptme tool☆48Jul 5, 2026Updated 2 weeks ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A web application that converts speech to speech 100% private using VAD (voice activity detection)☆18Aug 17, 2025Updated 11 months ago
- A small tool to dump the contents of a Binary glTF (.glb) file☆19Feb 26, 2024Updated 2 years ago
- ☆25Aug 26, 2025Updated 10 months ago
- TLS & API keys for your LLM APIs☆20Dec 17, 2025Updated 7 months ago
- Testbench for llama.cpp llama-server☆15Aug 20, 2025Updated 11 months ago
- A TTS model capable of generating ultra-realistic dialogue in one pass.☆32May 1, 2025Updated last year
- An OpenAI API compatible LLM inference server based on ExLlamaV2.☆24Feb 9, 2024Updated 2 years ago
- A user-friendly GUI for llama.cpp — convert, quantize, and run GGUF models without touching the terminal.☆21Jun 9, 2026Updated last month
- A lightweight LLaMA.cpp HTTP server Docker image based on Alpine Linux.☆39Jun 8, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A novel hybrid AI architecture leveraging Titan's-like memory and HRM-like reasoning☆26Jul 17, 2026Updated last week
- Llama Server Launcher (llama.cpp/ik_llama) GUI☆123Updated this week
- ☆99Mar 28, 2026Updated 3 months ago
- Create text chunks which end at natural stopping points without using a tokenizer☆26Nov 26, 2025Updated 7 months ago
- Vector functions and indexing for SQLite☆10Mar 26, 2023Updated 3 years ago
- Get up and running with Llama 2 and other large language models locally☆15Updated this week
- Python language chat with Ollama models locally, anthropic and openai☆24Mar 5, 2026Updated 4 months ago
- ☆13Feb 5, 2026Updated 5 months ago
- A web application that converts speech to speech 100% private☆86Jun 3, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Controllable Language Model Interactions in TypeScript☆10May 17, 2024Updated 2 years ago
- Sherpa-onnx-tts-stt source for homeassisstant addon with Kroko Onnx Streaming STT integration.☆30Dec 18, 2025Updated 7 months ago
- Super simple python connectors for llama.cpp, including vision models (Gemma 3, Qwen2-VL). Compile llama.cpp and run!☆31Dec 11, 2025Updated 7 months ago
- Code for the ICML 2025 Paper "Product of Experts with LLMs: Boosting Performance on ARC is a Matter of Perspective"☆55Nov 9, 2025Updated 8 months ago
- A powerful and user-friendly tool that generates detailed captions for your images☆21Nov 11, 2024Updated last year
- ☆21Sep 28, 2024Updated last year
- Fused BF16 Huffman GEMV Inference kernel☆22Apr 22, 2026Updated 3 months ago