Experimenting text-embeddings-inference server on both CPU and GPU
☆18Oct 25, 2023Updated 2 years ago
Alternatives and similar repositories for text-embedding-inference-server-edd
Users that are interested in text-embedding-inference-server-edd are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Dec 1, 2023Updated 2 years ago
- Applying Evaluation Driven Development (EDD) to aid in the design decision of RAG pipelines☆31Oct 20, 2023Updated 2 years ago
- Testing speed and accuracy of RAG with, and without Cross Encoder Reranker.☆50Jan 12, 2024Updated 2 years ago
- Chat Complex PDF with Tables Using IBM WatsonX, Langchain and LlamaParser.☆14Sep 18, 2025Updated last year
- Summarize SEC documents using LLMs☆14Aug 23, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆187Oct 9, 2023Updated 2 years ago
- ☆14May 30, 2024Updated 2 years ago
- LLM Agent that performs sentiment analysis of drawings and natural language using a combination of Google Gemini Vision model and GPT-4 T…☆13Dec 22, 2023Updated 2 years ago
- A Docker image with Llama Index, Lang Chain, and a few other popular AI packages installed by default☆11Nov 19, 2025Updated 10 months ago
- ☆19Aug 25, 2025Updated last year
- ☆20Jan 27, 2024Updated 2 years ago
- ☆20Jan 7, 2024Updated 2 years ago
- ☆22Aug 24, 2023Updated 3 years ago
- Demo code for Gemini Live Integration☆13Jul 29, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- a graph definition and execution library for python☆16Mar 22, 2023Updated 3 years ago
- ☆90Dec 31, 2023Updated 2 years ago
- Simple CogVLM client script☆13Dec 20, 2023Updated 2 years ago
- memAry @ University of Texas at Austin☆21Apr 16, 2024Updated 2 years ago
- ☆16Mar 23, 2025Updated last year
- ☆12Jan 25, 2025Updated last year
- Demo for Antigravity agent x LlamaParse/LiteParse for unstructured document parsing☆22May 18, 2026Updated 4 months ago
- Examples of RAG using Llamaindex with local LLMs - Gemma, Mixtral 8x7B, Llama 2, Mistral 7B, Orca 2, Phi-2, Neural 7B☆132Feb 25, 2024Updated 2 years ago
- Official Deployment Scripts for Learning Locker☆11Apr 30, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Efficient BM25 indexing using rust☆19Sep 17, 2024Updated 2 years ago
- ☆76Dec 12, 2023Updated 2 years ago
- Github repo for storing LlamaDatasets☆40May 5, 2026Updated 4 months ago
- LlamaWorksDB is a Retrieval Augmented Generation (RAG) product designed to interact with the documentation of various products such as Ll…☆17May 3, 2024Updated 2 years ago
- ☆31Mar 18, 2024Updated 2 years ago
- Sample GitHub Actions reusable workflows and Terraform reusable modules☆57Mar 25, 2024Updated 2 years ago
- ☆74Sep 27, 2024Updated last year
- ☆25Apr 18, 2023Updated 3 years ago
- ☆14Dec 7, 2023Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Our own point-and-click MCP server.☆17Sep 16, 2026Updated last week
- ☆18Oct 21, 2024Updated last year
- A collection of COVID-19 question-answer pairs and transformer baselines for evaluating QA models (Official Repository)☆26Nov 18, 2022Updated 3 years ago
- Test agent capability with complex generated mazes☆16Aug 7, 2025Updated last year
- A simple example of using LlamaIndex☆16May 24, 2023Updated 3 years ago
- A Hands-on Practical Guide to LlamaIndex☆33Oct 14, 2024Updated last year
- ☆56Nov 6, 2023Updated 2 years ago