A minimalistic C++ Jinja templating engine for LLM chat templates
☆228Sep 22, 2025Updated last year
Alternatives and similar repositories for minja
Users that are interested in minja are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A C++17 single-file header-only wrapper for llama.cpp☆30Updated this week
- Editor with LLM generation tree exploration☆87Feb 12, 2025Updated last year
- Rust crate for some audio utilities☆32Jun 17, 2026Updated 3 months ago
- Thin wrapper around GGML to make life easier☆48Jul 26, 2026Updated last month
- DINOv2 inference engine written in C/C++ using ggml and OpenCV.☆103May 6, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- FMO (Friendli Model Optimizer)☆14Jun 29, 2026Updated 2 months ago
- For audio visualization and playback in Jupyter notebooks.☆19Aug 27, 2026Updated 3 weeks ago
- egg is a flexible, high-performance e-graph library☆13Apr 5, 2024Updated 2 years ago
- Example and helpers for building rust projects under cmake☆16Oct 5, 2018Updated 7 years ago
- Semantic emoji finder. Python/dash UI. Uses sentence transformer embeddings and duckdb☆20Sep 15, 2025Updated last year
- HPYLMのC++実装☆11May 2, 2017Updated 9 years ago
- Pure C++ implementation of several models for real-time chatting on your computer (CPU & GPU)☆930Sep 7, 2026Updated 2 weeks ago
- GGUF implementation in C as a library and a tools CLI program☆366May 16, 2026Updated 4 months ago
- An N-gram punctuator for Chinese and English.☆20Oct 14, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Zeta implementation of a reusable and plug in and play feedforward from the paper "Exponentially Faster Language Modeling"☆16Nov 11, 2024Updated last year
- A minimalistic Swift implementation of the Jinja templating engine, specifically designed for parsing and rendering ML chat templates.☆133Sep 14, 2026Updated last week
- Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++☆7,242Updated this week
- Implementation of the LDP module block in PyTorch and Zeta from the paper: "MobileVLM: A Fast, Strong and Open Vision Language Assistant …☆15Mar 11, 2024Updated 2 years ago
- A lovely structopt library for C++! Parse command line arguments by defining a struct! ❤️☆12Apr 24, 2023Updated 3 years ago
- Spezi Module to Handle and Display User Interfaces for Chat-based Interactions☆11Apr 27, 2026Updated 4 months ago
- faster inference☆27Jan 20, 2025Updated last year
- WebAssembly binding for llama.cpp - Enabling on-browser LLM inference☆1,311Updated this week
- Tensor library for machine learning☆15,407Updated this week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆15May 11, 2025Updated last year
- A collection of experiments related to LLM inference with llama.cpp/mlx☆40Aug 26, 2026Updated 3 weeks ago
- iterate quickly with llama.cpp hot reloading. use the llama.cpp bindings with bun.sh☆51Oct 30, 2023Updated 2 years ago
- Disarm — Fast AArch64 Decode/Encoder☆54Sep 7, 2026Updated 2 weeks ago
- Extracts structured data from unstructured input. Programming language agnostic. Uses llama.cpp☆45May 16, 2024Updated 2 years ago
- A utility for async batch jobs in marimo☆13Mar 12, 2025Updated last year
- Kodi's Hy utility functions and macros☆13Jan 11, 2026Updated 8 months ago
- Official implementation of ECCV24 paper: POA☆24Aug 8, 2024Updated 2 years ago
- No-GIL Python environment featuring NVIDIA Deep Learning libraries.☆71Apr 14, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Sep 5, 2026Updated 2 weeks ago
- Rust implementation of RaBitQ + IVF and MSTG (multi-scale tree graph)☆19Feb 26, 2026Updated 6 months ago
- entropix style sampling + GUI☆27Oct 30, 2024Updated last year
- Anthropic’s Model Context Protocol implementation for Oat++☆49Dec 13, 2024Updated last year
- A single-header scheduler aware C++ oneshot channel☆20Sep 8, 2026Updated 2 weeks ago
- lightweight, standalone C++ inference engine for Google's Gemma models.☆7,043Updated this week
- Evaluation framework for GGUF☆15Apr 2, 2026Updated 5 months ago