Kon is a minimal coding agent (and also a highly opinionated one)
☆353Aug 1, 2026Updated last month
Alternatives and similar repositories for kon
Users that are interested in kon are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Zero-config AI agent written in pure Go. Enforced ephemeral subagents keep your model smart and your context tiny. From tiny local models…☆414Updated this week
- ☆12May 30, 2025Updated last year
- QLoRA: Efficient Finetuning of Quantized LLMs☆11Jul 22, 2023Updated 3 years ago
- llama.cpp/ik_llama.cpp launcher: loads big MoE models across mismatched multi-GPU rigs by exact VRAM math.☆267Updated this week
- A simple no-install web UI for Ollama and OAI-Compatible APIs!☆31Jan 30, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Coding agent with cool features.☆19Aug 23, 2026Updated last week
- llama-swap + a minimal ollama compatible api☆62May 26, 2026Updated 3 months ago
- A local-first LLM development studio. Build, test, and customize inference workflows with your own models — no cloud, totally local.☆17May 21, 2025Updated last year
- ☆24Jan 22, 2025Updated last year
- Code for boomerang distillation enables zero-shot model size interpolation.☆22Jul 10, 2026Updated last month
- Prompt Jinja2 templates for LLMs☆36Jul 9, 2025Updated last year
- ☆16Oct 28, 2025Updated 10 months ago
- ☆99Mar 28, 2026Updated 5 months ago
- entropix style sampling + GUI☆27Oct 30, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Customizable implementation of the self-instruct paper.☆1,051Mar 7, 2024Updated 2 years ago
- LLM fine-tuning and eval☆344Mar 21, 2024Updated 2 years ago
- Autonomous, agentic, creative story writing system that incorporates stored embeddings and Knowledge Graphs.☆110Feb 16, 2026Updated 6 months ago
- ⛔ DEPRECATED -- use flash-head instead (pip install flash-head)☆29Apr 10, 2026Updated 4 months ago
- The official API server for Exllama. OAI compatible, lightweight, and fast.☆1,338Updated this week
- ☆10Jun 30, 2022Updated 4 years ago
- Fine-tuning LLMs using QLoRA☆267Jun 21, 2026Updated 2 months ago
- Run multiple resource-heavy Large Models (LM) on the same machine with limited amount of VRAM/other resources by exposing them on differe…☆88Aug 24, 2026Updated last week
- Produce your own Dynamic 3.0 Quants and achieve optimum accuracy & SOTA quantization performance! Input a target size and the toolchain w…☆155Aug 21, 2026Updated last week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Spec-driven iterative development companion CLI for OpenCode.☆18Jun 7, 2026Updated 2 months ago
- Lightweight Llama 3 8B Inference Engine in CUDA C☆52Mar 21, 2025Updated last year
- Run autoresearch on any NVIDIA GPUs (Works on 2-4GB+ Cards)☆25Mar 22, 2026Updated 5 months ago
- ik_llama.cpp's Thireus fork with release builds for macOS/Windows/Ubuntu CPU, Vulkan and CUDA☆174Updated this week
- MiniLM (BERT) embeddings from scratch☆21Aug 14, 2025Updated last year
- ☆44Apr 26, 2026Updated 4 months ago
- Code for Papeg.ai☆228Jan 5, 2025Updated last year
- Mixed-vendor GPU inference cluster manager with speculative decoding☆33Jul 2, 2026Updated last month
- Collection of extensions for pi coding agent (sessions, ask_user, handoff)☆61Aug 22, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- AI coding agent optimized for small LLMs. 87% benchmark with 4B-active model.☆2,021Aug 12, 2026Updated 2 weeks ago
- RetroChat is a powerful command-line interface for interacting with various AI language models. It provides a seamless experience for eng…☆87Jul 13, 2025Updated last year
- AI agent framework, written from scratch (not based on openclaw), focused on stripping it down to the bare necessities, optimizing token …☆478Aug 18, 2026Updated 2 weeks ago
- ☆355Mar 5, 2026Updated 5 months ago
- Experimental sampler to make LLMs more creative☆31Aug 2, 2023Updated 3 years ago
- ☆212Jan 5, 2026Updated 7 months ago
- A harness optimized to smaller LLMs☆2,519Updated this week