⚡️ A fast and flexible PyTorch inference server that runs locally, on any cloud or AI HW.
☆147Jun 8, 2024Updated 2 years ago
Alternatives and similar repositories for nos
Users that are interested in nos are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12Apr 1, 2024Updated 2 years ago
- ☆14Aug 25, 2024Updated 2 years ago
- ojjson is a library designed to facilitate JSON interactions with Ollama, a large language api (LLM). It leverages the power of Zod for s…☆12Nov 7, 2024Updated last year
- The (open-source part of) code to reproduce "BPPSA: Scaling Back-propagation by Parallel Scan Algorithm".☆13Jun 7, 2021Updated 5 years ago
- Simple LLM inference server☆20Jun 13, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Simple orchestration for EC2 spot containers☆19Sep 27, 2024Updated last year
- ☆37Updated this week
- ☆62Jan 21, 2024Updated 2 years ago
- Repo for training MLMs, CLMs, or T5-type models on the OLM pretraining data, but it should work with any hugging face text dataset.☆98Feb 9, 2023Updated 3 years ago
- [NeurIPS 2024] The official implementation of "Image Copy Detection for Diffusion Models"☆18Oct 1, 2024Updated last year
- AI_Powered_Dev_Search_Engine☆12Mar 10, 2024Updated 2 years ago
- A stable, fast and easy-to-use inference library with a focus on a sync-to-async API☆46Sep 26, 2024Updated last year
- ☆13Feb 22, 2024Updated 2 years ago
- Universal connector to LLMs for Node.js & Bun☆30Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Automated LLM novelist☆46Apr 11, 2024Updated 2 years ago
- ☆15Dec 21, 2025Updated 8 months ago
- A really tiny autograd engine☆98May 26, 2025Updated last year
- ☆24Dec 27, 2024Updated last year
- SGLang is fast serving framework for large language models and vision language models.☆37Aug 21, 2026Updated last week
- ☆17Dec 18, 2023Updated 2 years ago
- Synthetic data for fine tuning LLM☆28Dec 26, 2024Updated last year
- Linum v2 (text-to-video) models☆50Jan 22, 2026Updated 7 months ago
- ☆36May 9, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆20Sep 28, 2024Updated last year
- WIP: Ofen is a toolkit aimed at making transformer models production-ready. API included☆17Oct 2, 2024Updated last year
- 🔥 LitLytics - an affordable, simple analytics platform that leverages LLMs to automate data analysis☆103Nov 25, 2024Updated last year
- Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs☆3,826May 28, 2026Updated 3 months ago
- Ultrafast serverless GPU inference, sandboxes, and background jobs☆1,760Updated this week
- Easily create LLM automation/agent workflows☆60Feb 13, 2024Updated 2 years ago
- ☆66Jun 27, 2024Updated 2 years ago
- Ultra-low-latency, high-throughput multiprocess transport over SHM and mmap. LMAX-Disruptor-style cross-process ring substrate.☆18Aug 6, 2026Updated 3 weeks ago
- Distribute and run AI workloads on Kubernetes magically in Python, like PyTorch for ML infra.☆1,224May 29, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SERAX is a structured data format built for AI-generated content. Unlike JSON/XML that break when AI outputs contain quotes or brackets, …☆19Jun 10, 2025Updated last year
- Implementation of a Tensorflow XLA rematerialization pass☆15Dec 20, 2019Updated 6 years ago
- Github repo of the CHARLIE AI interaction project☆14Aug 2, 2023Updated 3 years ago
- A hybrid router that uses Spot GPU instances to reduce costs and Serverless GPUs for making Cold Starts faster.☆27Mar 13, 2026Updated 5 months ago
- ☆10Aug 11, 2025Updated last year
- OpenAlpaca: A Fully Open-Source Instruction-Following Model Based On OpenLLaMA☆302Jun 13, 2023Updated 3 years ago
- The simplest way to serve AI/ML models in production☆1,196Updated this week