To better understand the ggml library
☆31Jun 13, 2025Updated last year
Alternatives and similar repositories for GGML-Tutorial
Users that are interested in GGML-Tutorial are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The repository targets the OpenCL gemm function performance optimization. It compares several libraries clBLAS, clBLAST, MIOpenGemm, Inte…☆17Mar 28, 2019Updated 7 years ago
- TPLink IPC Control☆20Jul 24, 2024Updated 2 years ago
- ☆20Apr 13, 2025Updated last year
- Some common CUDA kernel implementations (Not the fastest).☆30Jun 24, 2026Updated 2 months ago
- ☆25Mar 31, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A collection of papers on LLM applications in the IoT field.☆25Jan 21, 2026Updated 7 months ago
- A declarative workflow orchestrator for massive GPU clusters of any kind, including slurm, K8S and docker. Topology-aware node and GPU pl…☆41Sep 11, 2026Updated last week
- qwen2 and llama3 cpp implementation☆50Jun 7, 2024Updated 2 years ago
- ☆26Apr 2, 2026Updated 5 months ago
- ☆50Jun 7, 2024Updated 2 years ago
- 北大编译课程实践,独立完成的C语言子集SysY编译器,实现了从C语言编译到Koopa IR,再从Koopa IR编译到RISC-V汇编的实现☆34Jul 16, 2024Updated 2 years ago
- LLM-guided hyperparameter tuning☆10Oct 7, 2023Updated 2 years ago
- ImageAnnotation(Qt + Opencv) Tool makes tagging object easier, faster and more effective. Hope it can help you a lot. There are also som…☆11Nov 23, 2019Updated 6 years ago
- The 3rd Iteration of the Berkeley RISC-V DMA Accelerator☆28Nov 21, 2019Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- auto grad in rust with video explanation.☆26Jun 19, 2025Updated last year
- Elana: A Simple Energy & Latency Analyzer for LLMs☆16Apr 3, 2026Updated 5 months ago
- Keyformer proposes KV Cache reduction through key tokens identification and without the need for fine-tuning☆59Mar 26, 2024Updated 2 years ago
- GPU-accelerated LLM Training Simulator☆23Jun 26, 2025Updated last year
- ☆18Sep 21, 2025Updated 11 months ago
- Repo for MobiSys 2021 paper: "ClusterFL: A Similarity-Aware Federated Learning System for Human Activity Recognition".☆39Apr 4, 2023Updated 3 years ago
- ☆139Updated this week
- esp32 lvgl v8.1, support touch screen, rely on ESP-IDF environment☆14Nov 30, 2021Updated 4 years ago
- A set of tools to work with cgroup tree and process classification/QoS according to it☆10Oct 1, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An implementation of SGEMV with performance comparable to cuBLAS.☆12May 21, 2021Updated 5 years ago
- Erasure code library for Erlang☆12Sep 5, 2024Updated 2 years ago
- llvm slides and books and other☆63Feb 2, 2025Updated last year
- Enhance CHISEL for Smooth and Comfortable Chip Design☆20Mar 29, 2026Updated 5 months ago
- Official code for "Evaluations of Machine Learning Privacy Defenses are Misleading" (https://arxiv.org/abs/2404.17399)☆13Apr 29, 2024Updated 2 years ago
- ☆14Oct 23, 2023Updated 2 years ago
- flash attention tutorial written in python, triton, cuda, cutlass☆535Jan 20, 2026Updated 8 months ago
- [MobiCom 24] Adaptive DNN inference under memory constraints☆59Jan 22, 2025Updated last year
- [SIGCOMM 2023] PacketGame: Multi-Stream Packet Gating for Concurrent Video Inference at Scale☆15Jul 1, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A repo for llm on ncnn☆246Jul 2, 2026Updated 2 months ago
- LLVM passes with usage instructions☆18Apr 23, 2017Updated 9 years ago
- Recursive Bayesian Networks☆11May 11, 2025Updated last year
- ☆17Oct 19, 2023Updated 2 years ago
- A Valgrind extension for CUDA, unofficial mirror for https://www.hlrs.de/organization/av/spmt/research/cudagrind/☆10Aug 5, 2015Updated 11 years ago
- 音视频分析工具☆12May 10, 2017Updated 9 years ago
- High-speed GEMV kernels, at most 2.7x speedup compared to pytorch baseline.☆130Jul 13, 2024Updated 2 years ago