WebUI for managing llama-server sessions
☆30Jul 22, 2026Updated last month
Alternatives and similar repositories for llama-studio
Users that are interested in llama-studio are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Local-first AI orchestration via Transformers.js & WebGPU. Express/Electron hybrid for low-end hardware. Vision, TTS, STT, and Music Gene…☆22Jul 17, 2026Updated last month
- lets build claude code from scratch☆35May 13, 2026Updated 4 months ago
- Run local AI models privately with GPU acceleration. Modern UI, HuggingFace integration. Built with Tauri + React + Rust.☆62Apr 23, 2026Updated 4 months ago
- A lightweight CUDA-based local inference platform built around Z-Image Turbo by Tongyi☆25Jul 17, 2026Updated last month
- Turn any Mac or GPU into an OpenAI-compatible inference node. One-command setup, automatic HTTPS, model management, and distributed reque…☆25Sep 5, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A bare-bones GUI application for the local inference engine, llama.cpp. Built-in TPE optimiser to find the best flags for your system☆18Sep 6, 2026Updated last week
- Full AI context and content layer for coding agents over one MCP server — tree-sitter code-map, document RAG, shared memory, multi-agent …☆98Updated this week
- AI agent framework, written from scratch (not based on openclaw), focused on stripping it down to the bare necessities, optimizing token …☆482Updated this week
- Agent-first knowledge base — wiki over RAG. Plain markdown with wikilinks, background quality agents, and an MCP server for agent navigat…☆27Apr 15, 2026Updated 4 months ago
- your private, personal assistant☆81Apr 16, 2026Updated 4 months ago
- 249M-param MoE transformer built from scratch in PyTorch. GQA, RoPE, SwiGLU, sparse MoE with 3 aux losses, AMP training loop no Trainer a…☆19May 22, 2026Updated 3 months ago
- Pretraining codebase for Apertus models, based on Megatron-LM☆21Sep 25, 2025Updated 11 months ago
- Specialized fork for (relatively) fast single-GPU inference (in CUDA) using large MoE models that don't fit fully into VRAM☆17May 6, 2026Updated 4 months ago
- finetune method to create think/model/requires tags to allow LLMs to write programs for things they can calculate instead of hallucinatin…☆15Apr 8, 2026Updated 5 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A simple example of how to solve Kaggle's "Titanic: Machine Learning from Disaster" challenge using Python and scikit-learn☆12May 21, 2014Updated 12 years ago
- High-performance batched Top-K selection for CPU inference. Up to 80x faster than PyTorch, optimized for LLM sampling with AVX2 SIMD.☆18Mar 20, 2026Updated 5 months ago
- Open Imi is a open source claude desktop alternative for developers, engineers and tech teams to hack MCP's and agents to their own likin…☆10Nov 16, 2025Updated 9 months ago
- FIO Load Testing framework - for Openshift☆11Jun 8, 2023Updated 3 years ago
- Open-source framework for superagents.☆106Updated this week
- NVIDIA Nemo Parakeet TDT 0.6B V2 Audio to Text Python Script☆21May 8, 2025Updated last year
- Command-line package manager for open-sourced large language models. Download and run 100,000+ models, and share LLMs with a single comma…☆16Mar 11, 2026Updated 6 months ago
- HA Kubernetes cluster ready in less than 15 minutes☆16Dec 13, 2017Updated 8 years ago
- Template of frequently deployed Kubernetes resources on new clusters☆17May 14, 2020Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Total commander WFX plugin: S3☆13Jan 21, 2024Updated 2 years ago
- A command-line client for Bluesky☆22May 29, 2026Updated 3 months ago
- Grease monkey script for jenkins role-strategy UI enhancement☆16Aug 24, 2012Updated 14 years ago
- ☆12Oct 29, 2019Updated 6 years ago
- My Ansible role which sets some "favorite" defaults on the server (CentOS / Ubuntu / Windows)☆14May 6, 2025Updated last year
- Open CC BY 4.0 dataset: which local LLMs (Ollama) fit which hardware — params, quantization, RAM/VRAM. Powered by modelfit.io☆23Jul 6, 2026Updated 2 months ago
- Plan, visualize, and manage complex broadcast cabling systems with real-world production integrations.☆18Updated this week
- Run Claude Code against DigitalOcean Gradient AI. Spins up a local LiteLLM proxy to bridge Claude Code's Anthropic API format to DO's Ope…☆15Feb 18, 2026Updated 6 months ago
- ☆27Sep 3, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Run on TWO-DGX-Spark - vLLm-0.24.0 dual cache optimized DSV4F+DSpark+NVFP4 KV (Concurrency 12 with 1.5M context/3M KV token Pool) >0.58-0…☆21Aug 17, 2026Updated 3 weeks ago
- A modern Flask microservice providing accurate geolocation and comprehensive WHOIS information for IP addresses and domains. Features aut…☆53Mar 21, 2026Updated 5 months ago
- A Jellyfin plugin that automatically tags movies that were nominated for or won an Oscar.☆21May 17, 2026Updated 3 months ago
- The Standalone Agentic Memory Tool is a lightweight drop-in proxy for standard LLMs. It provides two superpowers: On-Demand RAG to select…☆16Mar 10, 2026Updated 6 months ago
- ShellGPT analogue for use with russian GigaChat☆12Jan 14, 2024Updated 2 years ago
- Fulloch - The Fully Local Home Voice Assistant☆139Updated this week
- A PHP portal to aggregate all your server services☆11Mar 28, 2023Updated 3 years ago