☆149Sep 16, 2026Updated 2 weeks ago
Alternatives and similar repositories for qllm2
Users that are interested in qllm2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [H] HyperspaceDB is a high-performance, vector database. It features 1-bit quantization, async replication, and native support for hierar…☆157Sep 8, 2026Updated 3 weeks ago
- Ultra-Sparse Adaptation of 1-Bit LLMs via XOR Patches☆91Aug 6, 2026Updated last month
- Open-source implementation of Google's TurboQuant (ICLR 2026) — KV cache compression to 2.5–4 bits with near-zero quality loss. 3.8–5.7x …☆52Mar 29, 2026Updated 6 months ago
- Jacobian-Brainwash : A manual alignment tool for large language models built on Anthropic's Jacobian Lens. Results are exportable.☆230Sep 6, 2026Updated 3 weeks ago
- The first pure SNN language model trained from scratch with a fully original architecture. 618M parameters • 93% sparsity • Runs on phone…☆82Aug 28, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- autoresearch for everything — autonomous iterative improvement for any system☆39Apr 2, 2026Updated 6 months ago
- A fully local document intelligence system that allows users to build a persistent private knowledge base from documents and query it usi…☆20Apr 15, 2026Updated 5 months ago
- Embed, cluster, and visualize any collection of texts in 3D semantic space — then learn a continuous semantic flow field over that space,…☆19May 9, 2026Updated 4 months ago
- Custom nodes for ComfyUI☆15Oct 10, 2025Updated 11 months ago
- nanoGPT using Equinox☆15Mar 3, 2023Updated 3 years ago
- run ollama & gguf easily with a single command☆54May 15, 2024Updated 2 years ago
- ☆29Dec 31, 2025Updated 9 months ago
- ☆19Oct 25, 2025Updated 11 months ago
- A local-first web search agent☆30Jun 20, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Multi-strategy RAG system achieving 74% Recall@10 on MultiHop-RAG. Combines RAPTOR hierarchical retrieval, knowledge graphs, HyDE, BM25, …☆43Feb 3, 2026Updated 7 months ago
- ☆22May 17, 2026Updated 4 months ago
- This bridge integrates Ollama into any chat interface and lets you build your own multi-agent pipeline, including a built-in memory data…☆89Jun 29, 2026Updated 3 months ago
- Simple node proxy for llama-server that enables MCP use☆19May 10, 2025Updated last year
- 🤖 AI GitHub App that automatically reviews PRs, triages issues, and monitors repository health using LLMs.☆24Sep 26, 2026Updated last week
- Mooch - Your Interview Pal☆19Updated this week
- An agentic harness for small language models.☆52Aug 11, 2026Updated last month
- Agentic memory using knowledge graphs☆36May 23, 2026Updated 4 months ago
- Give your AI agent instant API lookups instead of expensive source file reads. MCP server for C#, Go, Java, Python, and TypeScript.☆21Jul 28, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Would you like to recreate GPT-2 124m in a cave with a box of scraps and a 4090 in less than two hours? LETS SPEEDRUN!☆39Nov 26, 2025Updated 10 months ago
- Local-first RAG application for technical documentation and research papers☆28Jun 26, 2026Updated 3 months ago
- An ML engineering plugin for your coding agents.☆195Mar 17, 2026Updated 6 months ago
- LLamaHTML is a simple html file to communicate with a running llamacpp llama-server☆25Aug 5, 2025Updated last year
- A lightweight chat interface for interacting with local models, featuring persistent memory using a seamless SQLite database to store you…☆34Sep 15, 2025Updated last year
- A novel hybrid AI architecture leveraging Titan's-like memory and HRM-like reasoning☆27Aug 28, 2026Updated last month
- SwiftLet is a lightweight Python framework for running open-source Large Language Models (LLMs) locally using safetensors☆29Aug 6, 2025Updated last year
- ☆16Feb 1, 2025Updated last year
- ☆24Apr 3, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- llm chat in text file☆29Sep 12, 2026Updated 3 weeks ago
- A simpler self-hosted alternative to Open WebUI. Bring your own API keys or local models. Native Android and iOS clients.☆97Updated this week
- Book OCR Pipeline → Markdown (PaddleOCR-VL-1.5 + llama-server)☆34Sep 10, 2026Updated 3 weeks ago
- ☆25Jul 10, 2026Updated 2 months ago
- Developing K - a language model to generate OPENSCAD code from prompt☆19Dec 3, 2025Updated 10 months ago
- Turn any data source into an MCP server in 5 minutes. Build AI-agents-ready knowledge bases.☆22Jan 8, 2026Updated 8 months ago
- Validated, private, shareable knowledge-graph memory for AI — per-tenant, write-gated, PostgreSQL-authoritative, served over MCP.☆20Updated this week