Hierarchical RAG architecture scaling to 693K chunks on consumer hardware (4GB VRAM). Features 3-address routing, hybrid vector+graph fusion, and SetFit classification.
☆39Feb 11, 2026Updated 7 months ago
Alternatives and similar repositories for WiredBrain-Hierarchical-Rag
Users that are interested in WiredBrain-Hierarchical-Rag are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Dec 1, 2025Updated 9 months ago
- UI-based Fine-Tuning for Large Language Models (LLMs)☆20Dec 4, 2025Updated 9 months ago
- ☆22Jan 22, 2026Updated 7 months ago
- I developed a fine-tuned retrieval head for RAG that learns to more reliably retrieve relevant passages by transforming the query embeddi…☆17May 21, 2026Updated 3 months ago
- The rag pipeline for optimizing dynamic data editing.☆23Oct 30, 2025Updated 10 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆18Jun 6, 2026Updated 3 months ago
- Streaming Retrieval-Augmented Generation (RAG) agent in Go. It consumes real-time data from Kafka topics, processes it in configurable wi…☆27Jun 7, 2025Updated last year
- MilimoChat: Privacy-first, self-hosted AI chat with customizable personas, context-aware memory, and local analytics. Built on Python/Str…☆14Mar 12, 2025Updated last year
- This is more like a QA system☆16Jul 20, 2026Updated last month
- ☆16Dec 16, 2024Updated last year
- Multi-agent orchestration framework for AI applications - build, deploy, and manage AI agents across the full lifecycle with Forge, Conve…☆33Mar 28, 2026Updated 5 months ago
- Local runner for Microsoft VibeVoice Realtime TTS Fully compatible with Open-Webui Plug and Play. OpenAI api endpoint .Run the Colab note…☆44Jul 20, 2026Updated last month
- A conversational AI system using Ollama with persistent memory capabilities. Features hybrid context management (sliding window + vector …☆22Mar 20, 2026Updated 5 months ago
- A local dual-layer memory pattern for AI agents: a compact, human-readable markdown index paired with semantic retrieval from a local vec…☆59Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Semantic Web Memory for Intelligent Agents☆19Jul 12, 2026Updated 2 months ago
- ☆20Jan 3, 2026Updated 8 months ago
- One library to split them all: Sentence, Code, Docs. Chunk smarter, not harde; built for LLMs, RAG pipelines, and beyond.☆84Updated this week
- Distributed AI Agents Orchestration☆23Nov 10, 2025Updated 10 months ago
- Inspect LLM's logprobs and perplexity over a piece of text, or compare two LLMs (like a git diff)☆18Aug 3, 2026Updated last month
- ⚡ Debug your RAG pipeline without leaving the terminal. Real-time chunking visualization, batch testing, quality metrics, and one-click e…☆21Sep 3, 2026Updated 2 weeks ago
- ☆30Apr 23, 2025Updated last year
- Qwen2-VL for OCR & VQA☆19Sep 3, 2024Updated 2 years ago
- [H] HyperspaceDB is a high-performance, vector database. It features 1-bit quantization, async replication, and native support for hierar…☆156Sep 8, 2026Updated last week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Demo of fine-tuning QA models for answering FAQ of cloud providers documentation☆11Jun 20, 2026Updated 2 months ago
- A local-first LLM development studio. Build, test, and customize inference workflows with your own models — no cloud, totally local.☆17May 21, 2025Updated last year
- Declarative Document Indexing (DDI) framework for Python. Define schemas, extract structured indices, search smarter.☆47Jun 22, 2026Updated 2 months ago
- A text-grid web renderer for AI agents — see the web without screenshots☆104Mar 10, 2026Updated 6 months ago
- Precision Knowledge Editing (PKE): A novel method to reduce toxicity in LLMs while preserving performance, with robust evaluations and ha…☆12Nov 26, 2024Updated last year
- Find your files with natural language and ask questions.☆64Aug 21, 2026Updated 3 weeks ago
- 💬 ChatGPT application built with OpenAI API, Next.js, TypeScript, and TailwindCSS.☆13Apr 5, 2023Updated 3 years ago
- This repository demonstrated the possibility of using Vocode to create your very own AI Voice Agent☆17May 16, 2024Updated 2 years ago
- Custom Olares Market source optimized for Olares One (RTX 5090M + Core Ultra 9 275HX)☆19Jul 29, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Chrome extension for webmcp-hub to fetch site configs in real-time☆21Mar 6, 2026Updated 6 months ago
- We believe that every SOTA result is only valid on its own dataset. RAGView provides a unified evaluation platform to benchmark different…☆82Dec 5, 2025Updated 9 months ago
- Local ears and mouth for your LLM — offline, private, safe, free & open source. Voice-enable any local LLM stack: Claude Code, OpenCode, …☆44Sep 12, 2026Updated last week
- Production-ready RAG framework for Python — multi-tenant chatbots with streaming, tool calling, agent mode (LangGraph), vector search (FA…☆30May 7, 2026Updated 4 months ago
- [SIGIR 2025] Benchmarking Recommendation, Classification, and Tracing Based on Hugging Face Knowledge Graph☆17Jun 6, 2025Updated last year
- ☆25Feb 10, 2026Updated 7 months ago
- Groq-powered MAD: The first work to explore Multi-Agent Debate with Large Language Models :D☆12Jul 5, 2024Updated 2 years ago