GGUF implementation in C as a library and a tools CLI program
☆359May 16, 2026Updated 3 months ago
Alternatives and similar repositories for gguf-tools
Users that are interested in gguf-tools are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A small utility library for parsing GGUF file info☆31Jan 27, 2025Updated last year
- A fork of llama3.c used to do some R&D on inferencing☆23Dec 20, 2024Updated last year
- GGUF parser in Python☆28May 1, 2026Updated 4 months ago
- ggml implementation of BERT☆503Feb 23, 2024Updated 2 years ago
- GGUF parser for Go☆14Mar 8, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- CLIP inference in plain C/C++ with no extra dependencies☆568Aug 24, 2026Updated last week
- Inference of Mamba, Mamba2 and Mamba3 models in pure C☆203Mar 18, 2026Updated 5 months ago
- Fast neural codec compression and generation for audio waveforms☆232Dec 4, 2024Updated last year
- Implementation of ModernBERT in MLX☆21Jan 7, 2026Updated 7 months ago
- A new city of code on a cosmopolitan foundation.☆20Mar 19, 2021Updated 5 years ago
- Suno AI's Bark model in C/C++ for fast text-to-speech generation☆866Nov 16, 2024Updated last year
- Tensor library for machine learning☆15,289Updated this week
- Cross-platform binary launcher with Cosmopolitan libc☆34Apr 12, 2025Updated last year
- INT4/INT5/INT8 and FP16 inference on CPU for RWKV language model☆1,580Mar 23, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Train text generation model with JavaScript.☆15Jul 14, 2024Updated 2 years ago
- First token cutoff sampling inference example☆30Jan 15, 2024Updated 2 years ago
- Pure C inference for the GTE Small embedding model☆117Jan 21, 2026Updated 7 months ago
- A collection of some lockfree datastructures☆81Apr 20, 2023Updated 3 years ago
- Code Llama GGUF Demo☆10Aug 28, 2023Updated 3 years ago
- Port of Microsoft's BioGPT in C/C++ using ggml☆88Feb 21, 2024Updated 2 years ago
- A collection of experiments related to LLM inference with llama.cpp/mlx☆41Aug 26, 2026Updated last week
- Local ML voice chat using high-end models.☆189Jun 4, 2026Updated 2 months ago
- run ollama & gguf easily with a single command☆53May 15, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Flux 2 image generation model pure C inference☆1,988Feb 13, 2026Updated 6 months ago
- ☆137Nov 9, 2024Updated last year
- Visual bag of words for fast image matching☆25Apr 27, 2023Updated 3 years ago
- Inference Vision Transformer (ViT) in plain C/C++ with ggml☆319Apr 11, 2024Updated 2 years ago
- A minimalistic C++ Jinja templating engine for LLM chat templates☆228Sep 22, 2025Updated 11 months ago
- Temporary mail - Keep your real mailbox clean and secure. Temp Mail provides temporary, secure, anonymous, free, disposable email address…☆12Mar 17, 2023Updated 3 years ago
- Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++☆6,901Updated this week
- iterate quickly with llama.cpp hot reloading. use the llama.cpp bindings with bun.sh☆51Oct 30, 2023Updated 2 years ago
- lightweight, standalone C++ inference engine for Google's Gemma models.☆7,034Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆68Aug 19, 2024Updated 2 years ago
- Thin wrapper around GGML to make life easier☆48Jul 26, 2026Updated last month
- Yet Another (LLM) Web UI, made with Gemini☆12Dec 25, 2024Updated last year
- Load and run Llama from safetensors files in C☆16Oct 24, 2024Updated last year
- A simple MLX implementation for pretraining LLMs on Apple Silicon.☆84Aug 20, 2025Updated last year
- Inference Llama 2 in one file of pure C☆20,045Aug 6, 2024Updated 2 years ago
- Example of using SDL2 with Cosmopolitan Libc☆39Mar 20, 2024Updated 2 years ago