Structured Chain-of-Thought
☆220May 16, 2026Updated 4 months ago
Alternatives and similar repositories for structured-cot
Users that are interested in structured-cot are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PolarEngine: vLLM plugin for PolarQuant quantized LLM inference — 75% FP16 speed at 2.3x less VRAM☆36Apr 13, 2026Updated 5 months ago
- LLM speculative inference server for heterogeneous hardware & consumer GPUs☆2,868Updated this week
- llama.cpp fork with TQ3_1S/4S CUDA kernels — 3.5-bit WHT quantization achieving Q4s quality at 10% smaller size. Based on RaBitQ-inspired…☆229Updated this week
- ☆67Jun 4, 2026Updated 3 months ago
- Multi-tab GUI terminal wrapper for AI coding assistants — image paste, Telegram bridge, session persistence.☆18Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Tree-based speculative decoding for Apple Silicon (MLX). ~10-15% faster than DFlash on code, ~1.5x over autoregressive. First MLX port wi…☆145Apr 15, 2026Updated 5 months ago
- Test LLMs on real tasks. Compare models side-by-side.☆421Aug 10, 2026Updated last month
- ☆397Apr 16, 2026Updated 5 months ago
- ☆37Apr 25, 2026Updated 4 months ago
- DFlash: Block Diffusion for Flash Speculative Decoding☆6,102Aug 18, 2026Updated last month
- ☆15Apr 26, 2025Updated last year
- DFlash block-diffusion speculative decoding running on Apple Silicon via MLX, with an ANE execution path that explores heterogeneous acce…