Open Source Text Embedding Models with OpenAI Compatible API
☆172Jul 13, 2024Updated 2 years ago
Alternatives and similar repositories for open-text-embeddings
Users that are interested in open-text-embeddings are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- setup the env for vllm users☆16Oct 31, 2023Updated 2 years ago
- Deployment a light and full OpenAI API for production with vLLM to support /v1/embeddings with all embeddings models.☆45Jul 16, 2024Updated 2 years ago
- Infinity is a high-throughput, low-latency serving engine for text-embeddings, reranking models, clip, clap and colpali☆2,931Mar 24, 2026Updated 5 months ago
- Deploy your GGML models to HuggingFace Spaces with Docker and gradio☆37Jun 6, 2023Updated 3 years ago
- ☆10Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Loader extension for tabbyAPI in SillyTavern☆27Jun 30, 2025Updated last year
- 将零一万物 YI-34B 模型 API 转换为各种使用 OpenAI API 的开源软件支持的格式,无需修改开源软件配置或代码。☆12Jan 13, 2024Updated 2 years ago
- Sentence Embedding as a Service☆15Jun 30, 2025Updated last year
- A blazing fast inference solution for text embeddings models☆5,043Updated this week
- ☆22Jan 16, 2026Updated 7 months ago
- Normalize text string☆12Nov 6, 2018Updated 7 years ago
- Topic supervised non-negative matrix factorization with sparse matrices☆12Mar 24, 2020Updated 6 years ago
- llm-inference is a platform for publishing and managing llm inference, providing a wide range of out-of-the-box features for model deploy…☆97May 17, 2024Updated 2 years ago
- ☆11Apr 13, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A framework for evaluating the effectiveness of chain-of-thought reasoning in language models.☆19Feb 6, 2025Updated last year
- Convert LaBSE model from TF Hub to PyTorch.☆15Jan 15, 2026Updated 7 months ago
- GeventMP - Gevent Multiprocessing Extension☆18Aug 29, 2026Updated last week
- An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and FastChat-T5.☆11May 26, 2023Updated 3 years ago
- A fluent, scalable, and easy-to-use LLM data processing framework.☆28Jan 31, 2026Updated 7 months ago
- ☆146Aug 20, 2025Updated last year
- A pipeline to isolate and transcribe one language in mixed-language speech☆20Oct 25, 2022Updated 3 years ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆131Jun 25, 2024Updated 2 years ago
- Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-p…☆9,550Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Generate interesting clips using Youtube chat archives☆12Jun 6, 2024Updated 2 years ago
- ♟️ Browser Extension - Shows Win/Lose/Draw and average accuracy stats for games on chess.com☆12Updated this week
- Lightweight continuous batching OpenAI compatibility using HuggingFace Transformers include T5 and Whisper.☆29Mar 15, 2025Updated last year
- Open Source WizardCoder Dataset☆166Jul 12, 2023Updated 3 years ago
- Python bindings for llama.cpp☆68Feb 29, 2024Updated 2 years ago
- An implementation of the Computer Craft peripheral API allowing the use of wireless modems☆15Jan 25, 2024Updated 2 years ago
- Realtime tts reading of large textfiles by your favourite voice. +Translation via LLM (Python script)☆52Oct 18, 2024Updated last year
- prompt提示词工程快速上手☆28Aug 30, 2024Updated 2 years ago
- unoficial python api for responsive voice☆16Nov 8, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 4 months ago
- Enable tool-use ability for any LLM model (DeepSeek V3/R1, etc.)☆58May 27, 2025Updated last year
- Imitate OpenAI with Local Models☆91Aug 27, 2024Updated 2 years ago
- Langport is a language model inference service☆94Sep 9, 2024Updated last year
- ☆36Sep 6, 2024Updated 2 years ago
- Experiment deploying Rstudio to Google AppEngine☆11Sep 3, 2017Updated 9 years ago
- Code implement reposity of Paper HiQA☆106Mar 2, 2025Updated last year