Super-fast Structured Outputs
☆857Sep 4, 2026Updated this week
Alternatives and similar repositories for llguidance
Users that are interested in llguidance are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Fast, Flexible and Portable Structured Generation☆1,863Updated this week
- TensorRT-LLM server with Structured Outputs (JSON) built with Rust☆72Apr 25, 2025Updated last year
- ☆100Feb 12, 2026Updated 6 months ago
- Faster structured generation☆311Aug 5, 2026Updated last month
- Derivative-based regular expression engine for Rust☆27Sep 1, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A guidance language for controlling large language models.☆21,740May 21, 2026Updated 3 months ago
- Structured Outputs☆15,759Updated this week
- A high-performance constrained decoding engine based on context free grammar in Rust☆59May 22, 2025Updated last year
- AICI: Prompts as (Wasm) Programs☆2,077Jan 22, 2025Updated last year
- Enforce the output format (JSON Schema, Regex etc) of a language model☆2,034Apr 4, 2026Updated 5 months ago
- Graph model execution API for Candle☆18Jul 27, 2025Updated last year
- Renderer for the harmony response format to be used with gpt-oss☆4,498Apr 8, 2026Updated 4 months ago
- A curated list of papers related to constrained decoding of LLM, along with their relevant code and resources.☆367Jul 27, 2026Updated last month
- SGLang is a high-performance serving framework for large language models and multimodal models.☆35,594Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Fast, flexible LLM inference☆7,660Updated this week
- Transformers-compatible library for applying various compression algorithms to LLMs for optimized deployment with vLLM☆3,764Updated this week
- DSPy: The framework for programming—not prompting—language models☆37,828Updated this week
- A language for constraint-guided and efficient LLM programming.☆4,215May 22, 2025Updated last year
- FlashInfer: Kernel Library for LLM Serving☆6,343Updated this week
- A minimal, secure Python interpreter written in Rust for use by AI☆8,179Updated this week
- Fast Open-Source Search & Clustering engine × for Vectors & Arbitrary Objects × in C++, C, Python, JavaScript, Rust, Java, Objective-C, S…☆4,293Aug 31, 2026Updated last week
- Official Rust Implementation of Model2Vec☆212May 24, 2026Updated 3 months ago
- Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.☆721Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A Datacenter Scale Distributed Inference Serving Framework☆7,986Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆91,170Updated this week
- Tensor library for machine learning☆15,304Updated this week
- Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data ve…☆7,047Updated this week
- Use context-free grammars with an LLM☆176Mar 21, 2024Updated 2 years ago
- The programming language for agents☆9,147Updated this week
- Fast tokenizer for language models, compatible with SentencePiece, Tokenizers, Tiktoken and more. Supports BPE, Unigram and WordPiece tok…☆59Aug 10, 2026Updated 3 weeks ago
- Any model. Any hardware. Zero compromise. Built with @ziglang / @openxla / MLIR / @bazelbuild☆4,018Updated this week
- Fast, Accurate, Lightweight Python library to make State of the Art Embedding☆3,190Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆15Apr 26, 2025Updated last year
- Benchmark structured generation libraries☆31Oct 25, 2024Updated last year
- structured outputs for llms☆13,836Updated this week
- LMCache: Supercharge Your LLM with the Fastest KV Cache Layer☆11,690Updated this week
- Minimalist ML framework for Rust☆21,006Updated this week
- A framework for few-shot evaluation of language models.☆13,920Sep 1, 2026Updated last week
- A guide to structured generation using constrained decoding☆18Jun 9, 2024Updated 2 years ago