Examples of RAG using Llamaindex with local LLMs - Gemma, Mixtral 8x7B, Llama 2, Mistral 7B, Orca 2, Phi-2, Neural 7B
☆132Feb 25, 2024Updated 2 years ago
Alternatives and similar repositories for LlamaIndex-RAG-WSL-CUDA
Users that are interested in LlamaIndex-RAG-WSL-CUDA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Testing speed and accuracy of RAG with, and without Cross Encoder Reranker.☆50Jan 12, 2024Updated 2 years ago
- Experimenting text-embeddings-inference server on both CPU and GPU☆18Oct 25, 2023Updated 2 years ago
- Code for react youtube tutorial☆31Feb 14, 2024Updated 2 years ago
- ☆45Dec 21, 2023Updated 2 years ago
- OpenAI document chatbot using llama-index, pinecone and chainlit. With incremental features, giving you the tools to go from a basic RAG …☆80Apr 26, 2024Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- a local RAG LLM with persistent database to query your PDFs☆17Feb 8, 2024Updated 2 years ago
- ☆20Jan 7, 2024Updated 2 years ago
- A local rag demo☆33Mar 18, 2024Updated 2 years ago
- Notebooks and code with some RAG techniques using llamaindex☆31Apr 19, 2024Updated 2 years ago
- LLM Chatbot w/ Retrieval Augmented Generation using Llamaindex. It demonstrates how to impl. chunking, indexing, and source citation.☆45Oct 18, 2023Updated 2 years ago
- Full stack advanced chatbot over LlamaIndex.TS documentation with preview feature using Multi-documents-agents, bootstrapped with create-…☆154Mar 10, 2024Updated 2 years ago
- Fully local RAG setup: GPT4ALL, HuggingFace Embeddings model, FAISS, LangChain☆10May 10, 2023Updated 3 years ago
- API to load and query documents using RAG☆14Sep 25, 2023Updated 3 years ago
- Create-tsi is a generative AI RAG toolkit which generates AI Applications with low code.☆234Nov 4, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A Hands-on Practical Guide to LlamaIndex☆33Oct 14, 2024Updated last year
- Tutorials from AutoGen Basics to Use Cases☆32Nov 15, 2023Updated 2 years ago
- This project enhances the construction of RAG applications by addressing challenges, improving accessibility, scalability, and managing d…☆144Apr 10, 2024Updated 2 years ago
- LM Studio: RAG (Retrieval-Augmented Generation) Local LLM vs GPT-4☆22Jan 16, 2024Updated 2 years ago
- ☆90Dec 31, 2023Updated 2 years ago
- Chat Complex PDF with Tables Using IBM WatsonX, Langchain and LlamaParser.☆14Sep 18, 2025Updated last year
- ☆14Dec 1, 2023Updated 2 years ago
- Summarize SEC documents using LLMs☆14Aug 23, 2023Updated 3 years ago
- Developing a Korean LLM model : Hate Speech Filtering, Improving conversational skills, Finetuning with the RLHF method☆20May 27, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆14May 30, 2024Updated 2 years ago
- LLM Agent that performs sentiment analysis of drawings and natural language using a combination of Google Gemini Vision model and GPT-4 T…☆13Dec 22, 2023Updated 2 years ago
- ☆42Apr 17, 2025Updated last year
- A Docker image with Llama Index, Lang Chain, and a few other popular AI packages installed by default☆11Nov 19, 2025Updated 10 months ago
- ☆10Dec 17, 2023Updated 2 years ago
- Simple demo for chatting with a PDF - and optionally point the RAG implementation to a local LLM☆28Nov 29, 2023Updated 2 years ago
- ☆15Feb 7, 2024Updated 2 years ago
- Distributed compute that processes documents.☆20Jul 11, 2026Updated 2 months ago
- A collection of COVID-19 question-answer pairs and transformer baselines for evaluating QA models (Official Repository)☆26Nov 18, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆172Jan 25, 2024Updated 2 years ago
- ☆20Jan 27, 2024Updated 2 years ago
- Dynamic Metadata based RAG Framework☆77Dec 7, 2025Updated 9 months ago
- Advanced Coding AI Assistant that uses a Gradio interface to stream coding related responses. ChatRAG supports local and API inference an…☆26May 6, 2025Updated last year
- LlamaIndex wrapper for doing LLM RAG queries on local/private documents☆22Feb 10, 2024Updated 2 years ago
- Abusing Maybank's unauthenticated APIs for real time market data☆14Aug 12, 2024Updated 2 years ago
- Demo code for Gemini Live Integration☆13Jul 29, 2025Updated last year