WebUI for managing llama-server sessions
☆29Jul 22, 2026Updated last month
Alternatives and similar repositories for llama-studio
Users that are interested in llama-studio are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Run local AI models privately with GPU acceleration. Modern UI, HuggingFace integration. Built with Tauri + React + Rust.☆62Apr 23, 2026Updated 4 months ago
- A lightweight CUDA-based local inference platform built around Z-Image Turbo by Tongyi☆25Jul 17, 2026Updated last month
- 🦙 Turn your idle Mac or GPU into a free public AI API — OpenAI-compatible, one command, zero config (Powered by Llama.cpp)☆24Jul 31, 2026Updated 3 weeks ago
- A bare-bones GUI application for the local inference engine, llama.cpp. Built-in TPE optimiser to find the best flags for your system☆18Jul 3, 2026Updated last month
- Full AI context and content layer for coding agents over one MCP server — tree-sitter code-map, document RAG, shared memory, multi-agent …☆83Aug 8, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AI agent framework, written from scratch (not based on openclaw), focused on stripping it down to the bare necessities, optimizing token …☆469Updated this week
- your private, personal assistant☆78Apr 16, 2026Updated 4 months ago
- Agent-first knowledge base — wiki over RAG. Plain markdown with wikilinks, background quality agents, and an MCP server for agent navigat…☆27Apr 15, 2026Updated 4 months ago
- 249M-param MoE transformer built from scratch in PyTorch. GQA, RoPE, SwiGLU, sparse MoE with 3 aux losses, AMP training loop no Trainer a…☆19May 22, 2026Updated 3 months ago
- Pretraining codebase for Apertus models, based on Megatron-LM☆21Sep 25, 2025Updated 10 months ago
- Specialized fork for (relatively) fast single-GPU inference (in CUDA) using large MoE models that don't fit fully into VRAM☆17May 6, 2026Updated 3 months ago
- Tensor parallelism is all you need. Run LLMs on an AI cluster at home using any device. Distribute the workload, divide RAM usage, and in…☆18Nov 11, 2024Updated last year
- finetune method to create think/model/requires tags to allow LLMs to write programs for things they can calculate instead of hallucinatin…☆15Apr 8, 2026Updated 4 months ago
- A simple example of how to solve Kaggle's "Titanic: Machine Learning from Disaster" challenge using Python and scikit-learn☆12May 21, 2014Updated 12 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- MIMIC: A privacy-first, local-first AI desktop assistant with VRM avatar integration, multi-model logic, and persistent persona memory. R…☆15Mar 3, 2026Updated 5 months ago
- High-performance batched Top-K selection for CPU inference. Up to 80x faster than PyTorch, optimized for LLM sampling with AVX2 SIMD.☆18Mar 20, 2026Updated 5 months ago
- Open Imi is a open source claude desktop alternative for developers, engineers and tech teams to hack MCP's and agents to their own likin…☆10Nov 16, 2025Updated 9 months ago
- FIO Load Testing framework - for Openshift☆11Jun 8, 2023Updated 3 years ago
- Handle Android "draw over other apps" permissions and queries in a version-agnostic way☆22Jun 17, 2019Updated 7 years ago
- Director Agent + vision critic + image, video, music & voice models - all on a single AMD Instinct MI300X.☆44May 22, 2026Updated 3 months ago
- NVIDIA Nemo Parakeet TDT 0.6B V2 Audio to Text Python Script☆21May 8, 2025Updated last year
- HA Kubernetes cluster ready in less than 15 minutes☆16Dec 13, 2017Updated 8 years ago
- Template of frequently deployed Kubernetes resources on new clusters☆17May 14, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Total commander WFX plugin: S3☆13Jan 21, 2024Updated 2 years ago
- ☆26Updated this week
- A plugins for FreeIPA 4+ to manage DHCP configurations☆17May 6, 2018Updated 8 years ago
- WASM BusyTeX API for TeX Live 2026 (01.03.26) supporting LuaTeX, PdfTeX, & XeTeX in the browser☆28Updated this week
- Free AI-900 Microsoft Azure AI Fundamentals Certification preparation Content☆13Feb 7, 2021Updated 5 years ago
- A project combining roguelike with LLMs, RAG, Text2Speech, and Speech2Text☆19Mar 4, 2025Updated last year
- Grafana dashboard for viewing ZFS pool metrics with data collected by zpool_influxdb☆19Sep 16, 2023Updated 2 years ago
- ☆16Nov 27, 2025Updated 8 months ago
- Python based script to get twitter followers and output to a csv file☆10Jan 31, 2013Updated 13 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Termux friendly Android Automation skill for you Openclaw☆22Apr 8, 2026Updated 4 months ago
- A modern Flask microservice providing accurate geolocation and comprehensive WHOIS information for IP addresses and domains. Features aut…☆53Mar 21, 2026Updated 5 months ago
- Run on TWO-DGX-Spark - vLLm-0.24.0 dual cache optimized DSV4F+DSpark+NVFP4 KV (Concurrency 12 with 1.5M context/3M KV token Pool) >0.58-0…☆21Aug 17, 2026Updated last week
- A Jellyfin plugin that automatically tags movies that were nominated for or won an Oscar.☆18May 17, 2026Updated 3 months ago
- Local setup of Kubernetes using VMware Fusion, ansible and Kubespray☆16Feb 12, 2019Updated 7 years ago
- The Standalone Agentic Memory Tool is a lightweight drop-in proxy for standard LLMs. It provides two superpowers: On-Demand RAG to select…☆16Mar 10, 2026Updated 5 months ago
- ShellGPT analogue for use with russian GigaChat☆12Jan 14, 2024Updated 2 years ago