Based on the implementation of Google's TurboQuant (ICLR 2026) — Quansloth brings elite KV cache compression to local LLM inference. Quansloth is a fully private, air-gapped AI server that runs massive context models natively on consumer hardware with ease
☆151May 13, 2026Updated 2 months ago
Alternatives and similar repositories for Quansloth
Users that are interested in Quansloth are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A CLI toolkit for optimizing energy efficiency in software projects, leveraging local AI to promote sustainable development practices☆22Jun 4, 2026Updated last month
- Describe an app. Get a desktop app. AI-powered generator that builds real Tauri applications from natural language.☆212Apr 2, 2026Updated 3 months ago
- ☆42Apr 5, 2026Updated 3 months ago
- Visual repo documentation skill for AI agents☆31May 12, 2026Updated 2 months ago
- AgentFM is a peer-to-peer network that turns everyday computers into a decentralized AI supercomputer. AgentFM lets you run massive AI wo…☆127Jul 15, 2026Updated last week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- coze api to openai☆15Sep 1, 2024Updated last year
- ☆78Jun 11, 2026Updated last month
- Hierarchical Agent Memory with Git-Like Version Control☆594Jul 10, 2026Updated 2 weeks ago
- Looks through your social media 'saved' and transcribes the videos/posts and builds a document around them. Optional AI features for tran…☆19May 5, 2026Updated 2 months ago
- Your Android companion debug buddy☆215Updated this week
- Continuous learning for Kiro agents☆39May 6, 2026Updated 2 months ago
- Open-source, AI-driven penetration testing framework.☆51Jun 30, 2026Updated 3 weeks ago
- this is a google botguard token gen fully works by node.js, no webdriver☆51May 19, 2026Updated 2 months ago
- Local, git-versioned memory for AI coding agents. No RAG, no Docker, no external service. Capture, compile, recall over a local LLM wiki …☆128Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Pure Rust Inference Engine☆609Updated this week
- Spellbound - your multilingual AI-powered writing assistant☆13May 12, 2025Updated last year
- ☆46Jul 14, 2026Updated last week
- This is the codebase for the HYTOPIA world editor, available at https://build.hytopia.com - Anyone can fork this repository and submit pu…☆17Feb 28, 2026Updated 4 months ago
- ☆50Apr 26, 2026Updated 2 months ago
- 🧠 Bio-Agent OS: 🇻🇳 Bio-Inspired Memory Framework for AI Agents (OpenClaw/ERP). Researched & Developed by Dev Tuan Anh Ha (Locaith Solu…☆20Apr 21, 2026Updated 3 months ago
- ☆16Feb 18, 2024Updated 2 years ago
- Tessera — a workspace for organizing AI coding sessions across projects, collections, tabs, panes, and Git worktrees☆278Updated this week
- 97% token reduction for AI coding sessions — zero deps, 33 languages, MCP server☆601Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Microsoft Azure Data Lake Store Library for Delphi☆17Sep 8, 2020Updated 5 years ago
- ⚡️ The fastest way to run local LLMs on Apple Silicon — sub-second model loads, beats Ollama on throughput, tail latency, and full-respon…☆17Updated this week
- Force any Model, VAE, or CLIP to any GPU or CPU—and keep it there (or don't).☆31Feb 6, 2026Updated 5 months ago
- Give AI agents safe access to your database.☆75Jun 23, 2026Updated last month
- LLM inference in C/C++☆2,169Updated this week
- Build a classic 2D Pong game with Rust and Bevy, step by step, complete with movement, physics, and scoring!☆16Nov 17, 2024Updated last year
- ACE Platform for Agentic Context Engineering☆156Mar 31, 2026Updated 3 months ago
- Fast LLM speculative inference server for consumer hardware.☆2,673Updated this week
- Save 40%+ on agent token costs with code graphs: call graphs, dependency graphs, dead code detection, and blast radius analysis.☆93Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆14Feb 8, 2024Updated 2 years ago
- Website using FinBERT + live financial news scraping to assess short-term investment potential.☆14Mar 7, 2025Updated last year
- ☆29Mar 31, 2026Updated 3 months ago
- Agentic AI server in Rust. Multi-provider LLM routing, tool calling, RAG, MCP, multi-tenant workflows.☆16Jun 18, 2026Updated last month
- Flux.2 image generation on Mac with MLX - Native Swift implementation☆34Updated this week
- 📚 Build knowledge bases for RAG☆32Apr 20, 2026Updated 3 months ago
- Temporal knowledge graph for AI coding agents. Prevents hallucinations, repeated mistakes, regressions. 30 MCP tools, 12 runtime adapters…☆19Updated this week