A stable, fast and easy-to-use inference library with a focus on a sync-to-async API
☆46Sep 26, 2024Updated last year
Alternatives and similar repositories for embed
Users that are interested in embed are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Infinity is a high-throughput, low-latency serving engine for text-embeddings, reranking models, clip, clap and colpali☆2,942Mar 24, 2026Updated 5 months ago
- "a towel is about the most massively useful thing an interstellar AI hitchhiker can have"☆48Oct 9, 2024Updated last year
- TaCo: Enhancing Cross-Lingual Transfer for Low-Resource Languages in LLMs through Translation-Assisted Chain-of-Thought Processes☆14Jul 1, 2025Updated last year
- ☆137Jun 30, 2026Updated 2 months ago
- ☆15Apr 26, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- code for training and using chess embeddings models☆14Jun 9, 2024Updated 2 years ago
- Run multiple resource-heavy Large Models (LM) on the same machine with limited amount of VRAM/other resources by exposing them on differe…☆88Sep 3, 2026Updated 2 weeks ago
- Multilingual Entity Linking model by BELA model☆12Jul 20, 2023Updated 3 years ago
- Simple examples using Argilla tools to build AI