Local LLM Search Index for RAG
☆20May 5, 2026Updated 4 months ago
Alternatives and similar repositories for llmsearchindex
Users that are interested in llmsearchindex are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Apr 29, 2026Updated 5 months ago
- An ongoing, collaborative meta-analysis about Human-AI-Interactions. We aggregate data and knowledge to build a non-abrasive, user-friend…☆113Aug 25, 2026Updated last month
- Speculative Decoding Implementations: MTP, EAGLE-3, Medusa-1, PARD, Draft Models, N-gram and Suffix Decoding from scratch☆15May 2, 2026Updated 5 months ago
- Homework for STAT 205A - Berkeley☆13Dec 9, 2014Updated 11 years ago
- Discord chatbot interface to train an LLM on user message history☆27Jun 9, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Lightweight coding agent that runs in your terminal☆20Jun 23, 2026Updated 3 months ago
- Abandoned, unsupported research archive for a local-first, model-agnostic LLM agent harness exploring orchestration, code retrieval, off-…☆23Sep 6, 2026Updated 3 weeks ago
- Self hosted secure gateway for AI agents. One token. Full control. Complete audit trail.☆15Jun 10, 2026Updated 3 months ago
- Build SD card image for NVIDIA ShieldTV☆14Aug 27, 2016Updated 10 years ago
- Dynamic emotion system framework for Hermes Agent - real-time emotion detection,Long-term memory, context management, and recall systems,…☆31Aug 26, 2026Updated last month
- A SKILL.md agent skill that gives Codex, Claude, and other coding agents a 1.0-10.0 Linus Level dial for autonomy, strictness, security, …☆17May 8, 2026Updated 4 months ago
- LLM terms explained from an engineering perspective with the production implications, not just the definition.☆47May 23, 2026Updated 4 months ago
- Photographs to a textured mesh, on one NVIDIA card with 8 GB, offline. Built on TRELLIS.2, with a four view path.☆41Sep 12, 2026Updated 3 weeks ago
- A skill for AI agents. A virtual beer and wine sommelier that learns your tastes and suggests the best choice.☆15Jul 7, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Standalone web dashboard for the Hermes AI agent runtime — chat, sessions, memory, skills, secrets, config, and graph visualization.☆21Aug 22, 2026Updated last month
- handwritten harness for AI, not vibecoded, everything is a module/plugin, tailor made for use with local models, tiny system prompt (arou…☆495Updated this week
- Local-first AI operator☆17Updated this week
- Local screen memory for Claude Code and Codex CLI.☆29Apr 22, 2026Updated 5 months ago
- Fourier-based text morphing demo☆22Apr 12, 2026Updated 5 months ago
- Open platform for running Managed Agents at scale☆27Updated this week
- Multi-strategy RAG system achieving 74% Recall@10 on MultiHop-RAG. Combines RAPTOR hierarchical retrieval, knowledge graphs, HyDE, BM25, …☆43Feb 3, 2026Updated 7 months ago
- Native Excel (.xlsx) generator tool for Open WebUI — multi-sheet, Tables, live formulas, charts. MIT.☆33Jul 19, 2026Updated 2 months ago
- 3.34× faster inference on Apple Silicon — native MLX port of DFlash speculative decoding☆19Apr 11, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- TypeScript SDK for building AI agents with automatic, scoped, persistent memory. `agent-memory` wraps model calls with memory recall and…☆25Jun 15, 2026Updated 3 months ago
- Agent orchestration framework with Git-backed storage. Every turn is a Git commit - inspectable, rewindable, forkable.☆57Aug 5, 2026Updated last month
- Make AI Free Again - Shard is a GUI for Ollama LLM's. Powered by Next.js.☆63Feb 28, 2025Updated last year
- ☆15Feb 8, 2026Updated 7 months ago
- Smart proxy for LLM APIs that enables model-specific parameter control, automatic mode switching (like Qwen3's /think and /no_think), and…☆50May 19, 2025Updated last year
- PolarEngine: vLLM plugin for PolarQuant quantized LLM inference — 75% FP16 speed at 2.3x less VRAM☆36Apr 13, 2026Updated 5 months ago
- Robin LLM is a Java-based service that crawls OpenRouter's website for free LLM models, continuously tests their performance, and provide…☆22May 17, 2026Updated 4 months ago
- Production-ready ternary quantized (1.58-bit) Rust code generation model with mHC-lite, MaxRL training, and comprehensive benchmarking☆22Aug 2, 2026Updated 2 months ago
- Your appetite for code + Claude's capabilities = Limitless creation. No experience required - just pure hunger! 🧠⚡💻☆58Jun 20, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)☆12Aug 1, 2025Updated last year
- Port of Facebook's LLaMA model in C/C++☆15Updated this week
- A collection of Basalt (Bash) packages.☆15Feb 6, 2026Updated 7 months ago
- Fixes Mi Silent Mouse side buttons.☆20May 28, 2025Updated last year
- An open-source, local-first AI teammate and agent operating environment for turning intent into completed work—with tools, memory, permis…☆136Updated this week
- Natural language control for Python CLI tools using locally-trained SLMs (CPU inference)☆33Sep 22, 2026Updated last week
- ☆31Jul 24, 2026Updated 2 months ago