☆99Mar 28, 2026Updated 4 months ago
Alternatives and similar repositories for YALS
Users that are interested in YALS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official API server for Exllama. OAI compatible, lightweight, and fast.☆1,298Updated this week
- An optimized quantization and inference library for running LLMs locally on modern consumer-class GPUs☆1,127Updated this week
- A simple Gradio WebUI for loading/unloading models and loras in tabbyAPI.☆20Nov 21, 2024Updated last year
- Prompt Jinja2 templates for LLMs☆36Jul 9, 2025Updated last year
- Yet Another (LLM) Web UI, made with Gemini☆12Dec 25, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Croco.Cpp is fork of KoboldCPP infering GGML/GGUF models on CPU/Cuda with KoboldAI's UI. It's powered partly by IK_LLama.cpp, and compati…☆178Jul 3, 2026Updated last month
- ONNX speech pipeline library for ASR, diarization, VAD, and denoising☆20Jun 14, 2026Updated last month
- Eternal is an experimental platform for machine learning models and workflows.☆70Mar 9, 2025Updated last year
- llama-swap + a minimal ollama compatible api☆61May 26, 2026Updated 2 months ago
- Transplants vocabulary between language models, enabling the creation of draft models for speculative decoding WITHOUT retraining.☆54Oct 29, 2025Updated 9 months ago
- Train Llama Loras Easily☆30Aug 3, 2023Updated 3 years ago
- Simple node proxy for llama-server that enables MCP use☆19May 10, 2025Updated last year
- A simple no-install web UI for Ollama and OAI-Compatible APIs!☆31Jan 30, 2025Updated last year
- entropix style sampling + GUI☆27Oct 30, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆23May 14, 2026Updated 2 months ago
- Lightweight C inference for Qwen3 GGUF. Multiturn prefix caching & batch processing.☆25Sep 1, 2025Updated 11 months ago
- ☆352Mar 5, 2026Updated 5 months ago
- Loader extension for tabbyAPI in SillyTavern☆27Jun 30, 2025Updated last year
- Efficient 3bit/4bit quantization of LLaMA models☆18May 18, 2023Updated 3 years ago
- AI Based "Happiness Optimizer"☆12Oct 20, 2024Updated last year
- The serverside backend created for use with the LoRA Easy Training Scripts Frontend☆15Updated this week
- ☆24Jan 22, 2025Updated last year
- A Python package designed to simplify the process of creating and managing function calls to OpenAI's API, as well as models using LiteLL…☆17May 25, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Web UI for ExLlamaV2☆513Feb 5, 2025Updated last year
- llama.cpp's official website☆17Updated this week
- Large-scale LLM inference engine☆1,826Updated this week
- Service for testing out the new Qwen2.5 omni model☆62Apr 30, 2025Updated last year
- A fast inference library for running LLMs locally on modern consumer-class GPUs☆4,604Mar 4, 2026Updated 5 months ago
- LLamaHTML is a simple html file to communicate with a running llamacpp llama-server☆25Aug 5, 2025Updated last year
- Some random tools for working with the GGUF file format☆32Nov 24, 2023Updated 2 years ago
- win32 native frontend for llama-cli☆14Nov 2, 2024Updated last year
- An fully autonomous agent that accesses the browser and performs tasks.☆18Apr 25, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Local-first desktop AI workbench for roleplay, multi-character chat, long-form writing, RAG, MCP tools, plugins, and local models.☆106Jul 24, 2026Updated 2 weeks ago
- run ollama & gguf easily with a single command☆53May 15, 2024Updated 2 years ago
- Oobabooga "Hello World" API example for node.js with Express☆13Jul 2, 2023Updated 3 years ago
- A Windows tool to query various LLM AIs. Supports branched conversations, history and summaries among others.☆36May 11, 2026Updated 3 months ago
- Chatbot-to-speech using Orpheus TTS model. Interactive console app.☆21May 1, 2025Updated last year
- llama.cpp fork with additional SOTA quants and improved performance☆3,026Updated this week
- LLM Frontend in a single html file☆750Dec 27, 2025Updated 7 months ago