Everything I know about running LLMs locally
☆1,846Jul 10, 2026Updated 2 months ago
Alternatives and similar repositories for local-llm
Users that are interested in local-llm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Let Claude (or any LLM) actually watch a video — scene-aware, deduplicated frames + transcript, from a URL or local file. Runs locally, M…☆2,187Updated this week
- Local-first AI agents with governed, approval-gated memory. Any model provider; MCP tools and web search built in. Nothing remembered wit…☆360Updated this week
- Wireshark for MCP. A transparent proxy that shows every real tool call between your AI client and your MCP servers, live in your terminal…☆355Sep 11, 2026Updated 2 weeks ago
- Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦☆37,896Updated this week
- Shadow any website for offline viewing, with the JavaScript stripped out☆3,439Aug 10, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Python framework for self-hosted LLM tool-calling and multi-step agentic workflows☆2,250Sep 1, 2026Updated 3 weeks ago
- Stop wasting tokens and re-explaining your project every session. Recall gives Claude Code , Opencode durable memory — entirely offline.☆751Sep 7, 2026Updated 3 weeks ago
- A visualization tool that replays coding-agent sessions on a 3D map of your codebase.☆1,361Aug 10, 2026Updated last month
- Automation foundation model for tiny devices: 2-bit, 8-29 MB, tool calls, structured extraction and embeddings on phones, wearables, smar…☆12,743Updated this week
- Distributed AI/LLM for the people. Share compute privately or publicly to power your agents and chat.☆3,462Updated this week
- 🌱 Private, quiet space for thinking. Simple app for .md files.☆4,157Sep 2, 2026Updated 3 weeks ago
- DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm☆22,736Sep 20, 2026Updated last week
- Hundreds of models & providers. One command to find what runs on your hardware.☆37,208Updated this week
- Generate hands-on, multi-part technical tutorials on demand, with LLM skills tuned to make content approachable. Then you work through th…☆1,677Aug 3, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- autonomous red teaming platform; multi-agent offensive-security meta-harness☆6,260Sep 8, 2026Updated 2 weeks ago
- Simple CLI tool for deterministic routing of queries between local and hosted LLM models☆416Sep 6, 2026Updated 3 weeks ago
- Model router for agentic systems. Routes every prompt to the right model in <50ms. Cut costs 40-70% with just an endpoint change.☆5,324Updated this week
- Use Claude Code's autonomous agent loop with DeepSeek V4 Pro, OpenRouter, or any Anthropic-compatible backend. Same UX, 17x cheaper.☆2,258Jul 23, 2026Updated 2 months ago
- Speedrunning LoRA fine-tuning: frozen task, frozen hardware, public wall-clock leaderboard. modded-nanogpt for fine-tuning.☆149Sep 12, 2026Updated 2 weeks ago
- Running a big model on a small laptop☆4,151Mar 19, 2026Updated 6 months ago
- Secure, fast, efficient, battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, pre…☆41,923Updated this week
- Skills for threat modeling, scanning, triage, patching, plus an autonomous scanning harness you can /customize☆7,530Aug 6, 2026Updated last month
- Run models too big for your Mac's memory☆668Aug 27, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Beautiful, AI-native markdown IDE and LLM wiki☆4,337Updated this week
- Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.☆26,402Jul 29, 2026Updated last month
- Local-first, zero-trust agentic IDE for verifiable autonomous software development.☆868Sep 11, 2026Updated 2 weeks ago
- An embeddable, portable, branchable virtual machine to safely run Agents locally.☆6,392Updated this week
- Fast and Accurate Code Search for Agents. Uses 99% fewer tokens than grep+read☆6,146Updated this week
- an open source, extensible AI agent that goes beyond code suggestions - install, execute, edit, and test with any LLM☆54,706Updated this week
- Coding Agent singularly focused efficiency and context curation. Reduces API costs by 50-80% vs other agent AND improves the code quality…☆1,514Updated this week
- A guide for how to use your smartphone to code anywhere at anytime.☆1,733Jan 15, 2026Updated 8 months ago
- Sudoless Apple Silicon system monitor (native SwiftUI GUI) with ANE / Media Engine / memory-bandwidth tracking☆966Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- RL-training an AI agent to RL-train AI agents.☆240Jul 14, 2026Updated 2 months ago
- A ~9M parameter LLM that talks like a small fish.☆3,813Apr 15, 2026Updated 5 months ago
- OpenWiki is a CLI that writes and maintains agent documentation for your codebase.☆16,809Updated this week
- AI coworker with memory and collaboration☆17,979Updated this week
- Agent Lattice: a knowledge graph for your codebase, written in markdown.☆2,000Updated this week
- ☆577May 30, 2026Updated 3 months ago
- Aegis- a local zero-trust AI gate for OS and Apps packages☆31May 22, 2026Updated 4 months ago