A Python CLI to test, benchmark, and find the best RAG chunking strategy for your Markdown documents.
☆120Jan 18, 2026Updated 8 months ago
Alternatives and similar repositories for rag-chunk
Users that are interested in rag-chunk are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A self-hosted AI toolkit running locally via Docker Compose, bundling an LLM gateway, workflow automation, and a chat UI — all backed by …☆16May 17, 2026Updated 4 months ago
- Self-hosted RAG search engine — 34 formats, BM25+hybrid search, multi-LLM (Gemini/OpenAI/Claude/Ollama), FastAPI + Docker, production-rea…☆109May 19, 2026Updated 4 months ago
- Stop using static chunk sizes. A lightweight, production-ready RAG ingestion toolkit. Uses Docling for layout-aware parsing and applies s…☆69Mar 15, 2026Updated 6 months ago
- One library to split them all: Sentence, Code, Docs. Chunk smarter, not harder; built for LLMs, RAG pipelines, and beyond.☆87Updated this week
- ☆16Aug 7, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Model Context Protocol (MCP) server that enables communication and coordination between multiple AI agents☆33Sep 8, 2025Updated last year
- Cut LLM costs by up to 80% and unlock sub-millisecond responses with intelligent semantic caching.A drop-in, provider-agnostic LLM proxy …☆428Aug 18, 2026Updated last month
- Draw cool web UIs with local LLMs.☆29Jun 4, 2026Updated 4 months ago
- smallevals — CPU-fast, GPU-blazing fast offline retrieval evaluation for RAG systems with tiny QA models.☆23Dec 4, 2025Updated 10 months ago
- Exploring Ultrafast machine learning on FPGAs via Kolmogorov-Arnold Networks☆17Jun 15, 2026Updated 3 months ago
- An MCP server that executes Python code in isolated rootless containers with optional MCP server proxying. Implementation of Anthropic's …☆340Dec 5, 2025Updated 10 months ago
- Hands-on AI agentic patterns workshop material for building real AI and web apps.☆37Sep 3, 2026Updated last month
- Text Match Cut Video Generator Web App☆38Feb 19, 2026Updated 7 months ago
- Oscillink — Self‑Optimizing Coherent Memory for Embedding Workflows☆15Nov 24, 2025Updated 10 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- RLVR Testing and Training☆21Aug 28, 2025Updated last year
- SLMs for personal expenses summaries☆23Sep 7, 2026Updated last month
- ☆21Jan 25, 2025Updated last year
- The rag pipeline for optimizing dynamic data editing.☆23Oct 30, 2025Updated 11 months ago
- Talk to your data. Instantly analyze, visualize, and transform☆22Oct 30, 2025Updated 11 months ago
- Streamable multi-format serialization with schema☆23Dec 10, 2024Updated last year
- Local runner for Microsoft VibeVoice Realtime TTS Fully compatible with Open-Webui Plug and Play. OpenAI api endpoint .Run the Colab note…☆44Jul 20, 2026Updated 2 months ago
- Orchestrator Kit for Agentic Reasoning - OrKa is a modular AI orchestration system that transforms Large Language Models (LLMs) into comp…☆95Jul 7, 2026Updated 3 months ago
- A framework for creating message-driven training systems with PyTorch☆21Aug 2, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- MarrowScript compiler. Welcome to deterministic typed LLM orchestration as a compile-time concern☆32May 21, 2026Updated 4 months ago
- ☆15Sep 23, 2024Updated 2 years ago
- CustomGPT.ai’s RAG API’s Starter Kit, including multi-instance embedded widgets, floating buttons, and standalone application.☆53Dec 23, 2025Updated 9 months ago
- Splat: Affinity diagramming tool in a single HTML file. Cluster notes on a board. Organize data visually. Fully local for privacy. Semant…☆21Dec 31, 2025Updated 9 months ago
- One stop shop - Local-first RAG stack with intelligent polyglot-code/docs, remote code execution, local llama enrichment, progressive dis…☆36Feb 17, 2026Updated 7 months ago
- Geometric AI research: a proven cube-math core, reusable vector-collapse dynamics, and reproducible experiments in embeddings, NLI, gener…☆15Jul 31, 2026Updated 2 months ago
- Embedding and readout for simple multi-categorical and gaussian continuous☆20Jul 5, 2026Updated 3 months ago
- Fetch only relevant tools for the current conversation and save cost while increasing the precision of your LLM Response☆17Sep 9, 2025Updated last year
- Handle-ifying MCP proxy: slash token bills on chatty MCP servers via content-aware backends and local munch verbs☆33Sep 1, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A better API for Python asynchronous UDP☆19May 19, 2026Updated 4 months ago
- Generate a llama-quantize command to copy the quantization parameters of any GGUF☆36Apr 20, 2026Updated 5 months ago
- minimal scripts for 24GB VRAM GPUs. training, inference, whatever☆52Jun 3, 2026Updated 4 months ago
- MCP server that allows Claude to have a voice.☆14May 5, 2025Updated last year
- Z.E.T.A. Zero: Cognitive Construct & Persistent Memory for Local LLMs☆46Feb 13, 2026Updated 7 months ago
- Structured PDF extraction at 318 pages/second on CPU☆108Sep 8, 2026Updated last month
- A simple streamlit app, dockerized, to do OCR on documents. I'm lazy, idk.☆26Aug 18, 2025Updated last year