Streaming Retrieval-Augmented Generation (RAG) agent in Go. It consumes real-time data from Kafka topics, processes it in configurable windows, converts the window content into embeddings using Ollama, and stores these embeddings (along with the original text) in Elasticsearch
☆27Jun 7, 2025Updated last year
Alternatives and similar repositories for stream-rag-agent
Users that are interested in stream-rag-agent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- One library to split them all: Sentence, Code, Docs. Chunk smarter, not harde; built for LLMs, RAG pipelines, and beyond.☆85Updated this week
- ☆17Nov 23, 2025Updated 10 months ago
- Hierarchical RAG architecture scaling to 693K chunks on consumer hardware (4GB VRAM). Features 3-address routing, hybrid vector+graph fus…☆39Feb 11, 2026Updated 7 months ago
- CDRAG is a new retrieval framework that uses hierarchical document clustering, and LLM-guided document selection from those clusters to c…☆38Apr 13, 2026Updated 5 months ago
- A conversational AI system using Ollama with persistent memory capabilities. Features hybrid context management (sliding window + vector …☆22Mar 20, 2026Updated 6 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- I developed a fine-tuned retrieval head for RAG that learns to more reliably retrieve relevant passages by transforming the query embeddi…☆17May 21, 2026Updated 4 months ago
- ☆18Dec 1, 2025Updated 9 months ago
- Retrieval-augmented generation (RAG) for remote & local LLM use☆45May 24, 2025Updated last year
- A repository dedicated to exploring and analyzing the latest developments in Generative AI☆27Apr 1, 2025Updated last year
- [COLM '24] Source-Aware Training Enables Knowledge Attribution in Language Models☆22Apr 1, 2025Updated last year
- AI-powered text compression library for RAG systems and API calls. Reduce token usage by up to 50-60% while preserving semantic meaning w…☆90Aug 16, 2025Updated last year
- A VSCode extension for running LLM prompts. It turns VSCode into a powerful prompt IDE.☆33Oct 11, 2024Updated last year
- Multi-strategy RAG system achieving 74% Recall@10 on MultiHop-RAG. Combines RAPTOR hierarchical retrieval, knowledge graphs, HyDE, BM25, …☆42Feb 3, 2026Updated 7 months ago
- Demo of fine-tuning QA models for answering FAQ of cloud providers documentation☆11Jun 20, 2026Updated 3 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆22Jan 22, 2026Updated 8 months ago
- buderus2mqtt is a gateway between a KM200 Buderus internet gateway and MQTT with the https://github.com/mqtt-smarthome topic and payload …☆18Jan 3, 2023Updated 3 years ago
- Local AI Desktop Assistant with a focus on long-term memory, computer vision, and customizable behavior.☆18Feb 26, 2026Updated 6 months ago
- Precision Knowledge Editing (PKE): A novel method to reduce toxicity in LLMs while preserving performance, with robust evaluations and ha…☆12Nov 26, 2024Updated last year
- This project is demo for using MvvmCross navigate between native page and Xamarin Forms page☆11Apr 23, 2017Updated 9 years ago
- Is a high-performance Augmented Recovery-Generation (RAG) solution based on Redis, Qdrant or PostgreSQL. It offers a high-level interface…☆30Jan 6, 2026Updated 8 months ago
- C# Wrapper for SmartPlant APIs☆13Mar 15, 2017Updated 9 years ago
- Groq-powered MAD: The first work to explore Multi-Agent Debate with Large Language Models :D☆12Jul 5, 2024Updated 2 years ago
- ☆11May 2, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- RAG boilerplate with semantic/propositional chunking, hybrid search (BM25 + dense), LLM reranking, query enhancement agents, CrewAI orche…☆78Nov 18, 2025Updated 10 months ago
- A real-time shared memory layer for multi-agent LLM systems.☆68Jan 12, 2026Updated 8 months ago
- Inspect status of Azure DevOps Service inside Visual Studio☆15Dec 27, 2018Updated 7 years ago
- Home of the Stratis Identity proof of concept☆13Sep 18, 2017Updated 9 years ago
- Django implementation of WOPI server for Microsoft Office Online Server.☆13Feb 21, 2018Updated 8 years ago
- This project implements a Reinforcement Learning (RL) enhanced Retrieval-Augmented Generation (RAG) system that optimizes document retrie…☆25Apr 27, 2025Updated last year
- Model Context Protocol (MCP) Gateway & Registry - Central hub for managing tools, resources, and prompts for MCP-compatible LLMs. Transla…☆41Mar 1, 2026Updated 6 months ago
- A data analysis AI built with crewAI☆12Feb 19, 2024Updated 2 years ago
- Code for my youtube video on building a local AI assistant with whisper turbo 3 and llama 3.2☆15Oct 21, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆16Jul 23, 2026Updated 2 months ago
- code for training and using chess embeddings models☆14Jun 9, 2024Updated 2 years ago
- A text analysis library for relevance and subtheme detection☆16Sep 17, 2026Updated last week
- ☆67Jun 24, 2025Updated last year
- We're working on bringing IIS into Let's Encrypt☆14Jun 17, 2015Updated 11 years ago
- Files and script to run a Red Hat OpenShift Dev Spaces demo☆20Feb 5, 2024Updated 2 years ago
- TFS Auto Shelve Extension for Visual Studio☆13Dec 13, 2023Updated 2 years ago