Provenance-first extractive RAG: return verbatim source spans with citations using local ModernBERT or optional LLM-assisted extraction.
☆204Sep 7, 2026Updated 3 weeks ago
Alternatives and similar repositories for verbatim-rag
Users that are interested in verbatim-rag are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Learn rule-based models from examples using LLM-powered synthesis. Replace expensive LLM calls with fast, deterministic, inspectable rege…☆34Sep 21, 2026Updated last week
- Span-level grounding verification for RAG, code, and tool-grounded AI outputs.☆612Sep 7, 2026Updated 3 weeks ago
- ☆32Jun 22, 2026Updated 3 months ago
- Nearly Inference Free Embeddings: make your RAG queries 500x faster☆86Apr 27, 2026Updated 5 months ago
- Pre-train Static Word Embeddings☆111Jun 9, 2026Updated 3 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Unified Schema-Based Information Extraction☆2,215Updated this week
- The rag pipeline for optimizing dynamic data editing.☆23Oct 30, 2025Updated 10 months ago
- Fast search index for SPLADE sparse retrieval models implemented in Python using Numpy and Numba☆39Oct 16, 2025Updated 11 months ago
- ☆58Dec 27, 2025Updated 9 months ago
- A package for handy processing of semantic graphs and meaning representations, (e.g. AMR) with a special focus on standardized evaluation☆27May 1, 2025Updated last year
- Fast Multimodal Semantic Deduplication & Filtering☆971Updated this week
- Knowledgeable Embedding: Injecting dynamically updatable entity knowledge into embeddings to enhance RAG☆15Aug 31, 2025Updated last year
- Unified multi-layer caching library for AI/agent pipelines — LangChain, LangGraph, AutoGen, CrewAI, Agno, A2A☆30Jun 4, 2026Updated 3 months ago
- Break a PDF into chapters with metadata. Creates a folder with the chapters properly labeled. Do the Truffle Shuffle!☆19Jul 29, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Trully flash implementation of DeBERTa disentangled attention mechanism.☆93Sep 20, 2026Updated last week
- Generalist and Lightweight Model for Text Classification☆540Updated this week
- A collection of experimental Retrieval Augmented Generation (RAG) Techniques to elevate your pipelines, all with code and intuitive expla…☆36Jul 21, 2025Updated last year
- Generalist and Lightweight Model for Named Entity Recognition (Extract any entity types from texts)☆3,956Updated this week
- Architecture pattern for combining a fast LLM voice loop with a slower SLM that tracks hard facts.☆15Apr 27, 2026Updated 5 months ago
- A demonstration of metadata generation for RAG using a Health Canada document☆22Jan 19, 2025Updated last year
- Fully local governed RAG for code, documents, and tables. No LLM API key. No GPU. No infrastructure.☆41Updated this week
- Fast Diversification for Search & Retrieval☆500May 24, 2026Updated 4 months ago
- Production inference for encoder models - ColBERT, GLiNER, ColPali, embeddings etc. - as vLLM plugins for online and in-process deploymen…☆80Jul 6, 2026Updated 2 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A framework for benchmarking embedding models in hybrid search scenarios (BM25 + vector search) using Weaviate.☆40Aug 19, 2026Updated last month
- An Open Source Python package to strengthen Gemini Citations- Validated, Controlled, Reliable and Structured Citations☆15Sep 9, 2025Updated last year
- ☆18Jul 1, 2025Updated last year
- Late Interaction Models Training & Retrieval☆895Jul 23, 2026Updated 2 months ago
- Squeeze verbose LLM agent tool output down to only the relevant lines☆23Apr 27, 2026Updated 5 months ago
- Datamodels for hugging face tokenizers☆112Updated this week
- I got tired of manually creating training datasets, so I built this. Transform your PDFs/docs into fine-tuning data automatically.☆32Sep 2, 2025Updated last year
- A research toolkit for decomposing and explaining text similarity across neural, structured, and symbolic levels.☆32Aug 13, 2026Updated last month
- Plug-and-play document AI with zero-shot models.☆126May 11, 2026Updated 4 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A Python package for zero-shot text anonymization using Transformer-based NER models.☆84Updated this week
- Efficient and modular GraphRAG system☆50Jul 7, 2026Updated 2 months ago
- SpanMarker for Named Entity Recognition☆477Apr 10, 2026Updated 5 months ago
- DSPydantic: Auto-Optimize Your Prompts and Pydantic Models with DSPy☆333Mar 20, 2026Updated 6 months ago
- A multi-lingual approach to AllenNLP CoReference Resolution along with a wrapper for spaCy.☆111Apr 16, 2024Updated 2 years ago
- CustomGPT.ai’s RAG API’s Starter Kit, including multi-instance embedded widgets, floating buttons, and standalone application.☆53Dec 23, 2025Updated 9 months ago
- Attend - to what matters.☆17Feb 22, 2025Updated last year