Super-fast Structured Outputs
☆876Sep 25, 2026Updated this week
Alternatives and similar repositories for llguidance
Users that are interested in llguidance are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Fast, Flexible and Portable Structured Generation☆1,937Updated this week
- TensorRT-LLM server with Structured Outputs (JSON) built with Rust☆73Apr 25, 2025Updated last year
- ☆101Feb 12, 2026Updated 7 months ago
- Faster structured generation☆316Aug 5, 2026Updated last month
- Derivative-based regular expression engine for Rust☆28Sep 18, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A guidance language for controlling large language models.☆21,779May 21, 2026Updated 4 months ago
- Structured Outputs☆15,886Sep 21, 2026Updated last week
- A high-performance constrained decoding engine based on context free grammar in Rust☆59May 22, 2025Updated last year
- AICI: Prompts as (Wasm) Programs☆2,076Jan 22, 2025Updated last year
- Enforce the output format (JSON Schema, Regex etc) of a language model☆2,043Apr 4, 2026Updated 5 months ago
- Graph model execution API for Candle☆18Jul 27, 2025Updated last year
- Renderer for the harmony response format to be used with gpt-oss☆4,508Apr 8, 2026Updated 5 months ago
- A curated list of papers related to constrained decoding of LLM, along with their relevant code and resources.☆369Jul 27, 2026Updated 2 months ago
- SGLang is a high-performance serving framework for large language models and multimodal models.☆36,481Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Fast, flexible LLM inference☆7,715Updated this week
- State-of-the-art LLM compression, built for production inference with vLLM☆3,824Updated this week
- DSPy: The framework for programming—not prompting—language models☆38,374Updated this week
- A language for constraint-guided and efficient LLM programming.☆4,218May 22, 2025Updated last year
- FlashInfer: Kernel Library for LLM Serving☆6,510Updated this week
- A minimal, secure Python interpreter written in Rust for use by AI☆8,362Updated this week
- Fast Open-Source Search & Clustering engine × for Vectors & Arbitrary Objects × in C++, C, Python, JavaScript, Rust, Java, Objective-C, S…☆4,319Aug 31, 2026Updated 3 weeks ago
- Official Rust Implementation of Model2Vec☆216Sep 21, 2026Updated last week
- Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.☆728Sep 17, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Datacenter Scale Distributed Inference Serving Framework☆8,166Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆92,771Updated this week
- Tensor library for machine learning☆15,416Updated this week
- Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data ve…☆7,118Updated this week
- Use context-free grammars with an LLM☆176Mar 21, 2024Updated 2 years ago
- The programming language for agents☆9,344Updated this week
- Fast, Accurate, Lightweight Python library to make State of the Art Embedding☆3,220Updated this week
- Fast tokenizer for language models, compatible with SentencePiece, Tokenizers, Tiktoken and more. Supports BPE, Unigram and WordPiece tok…☆62Aug 10, 2026Updated last month
- Any model. Any hardware. Zero compromise. Built with @ziglang / @openxla / MLIR / @bazelbuild☆4,087Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Apr 26, 2025Updated last year
- structured outputs for llms☆13,947Updated this week
- Benchmark structured generation libraries☆31Oct 25, 2024Updated last year
- Minimalist ML framework for Rust☆21,114Updated this week
- A framework for few-shot evaluation of language models.☆14,085Sep 14, 2026Updated 2 weeks ago
- LMCache: Supercharge Your LLM with the Fastest KV Cache Layer☆11,922Updated this week
- A guide to structured generation using constrained decoding☆18Jun 9, 2024Updated 2 years ago