LLaVA server (llama.cpp).
☆184Oct 20, 2023Updated 2 years ago
Alternatives and similar repositories for llava-cpp-server
Users that are interested in llava-cpp-server are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 生成训练文本检测数据集☆12Jul 1, 2020Updated 6 years ago
- iterate quickly with llama.cpp hot reloading. use the llama.cpp bindings with bun.sh☆51Oct 30, 2023Updated 2 years ago
- ☆22Oct 20, 2023Updated 2 years ago
- ☆1,276Oct 24, 2023Updated 2 years ago
- A simple "Be My Eyes" web app with a llama.cpp/llava backend☆495Nov 28, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Fine-tuning, DPO, RLHF, RLAIF on LLMs - Qwen3, Zephyr 7B GPTQ with 4-Bit Quantization, Mistral-7B-GPTQ☆15Jul 5, 2025Updated last year
- Semantic emoji finder. Python/dash UI. Uses sentence transformer embeddings and duckdb☆20Sep 15, 2025Updated 10 months ago
- Suno AI's Bark model in C/C++ for fast text-to-speech generation☆866Nov 16, 2024Updated last year
- Friendly Terminal Assistant for Developers☆17Mar 23, 2024Updated 2 years ago
- Port of Suno AI's Bark in C/C++ for fast inference☆56Apr 15, 2024Updated 2 years ago
- This repository is a voice search demo using OpenAI Whisper, DuckDB, and the Metaphone algorithm. The associate blog post is here: https:…☆13May 15, 2024Updated 2 years ago
- GPT-2 small trained on phi-like data☆68Feb 18, 2024Updated 2 years ago
- Implementation of "LM-Infinite: Simple On-the-Fly Length Generalization for Large Language Models"☆40Nov 11, 2024Updated last year
- Inference Vision Transformer (ViT) in plain C/C++ with ggml☆318Apr 11, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Fine-tune mistral-7B on 3090s, a100s, h100s☆735Oct 11, 2023Updated 2 years ago
- Port of MiniGPT4 in C++ (4bit, 5bit, 6bit, 8bit, 16bit CPU inference with GGML)☆573Aug 8, 2023Updated 2 years ago
- The llama-cpp-agent framework is a tool designed for easy interaction with Large Language Models (LLMs). Allowing users to chat with LLM …☆651Mar 9, 2026Updated 4 months ago
- Python bindings for llama.cpp☆10,512Updated this week
- A C++17 single-file header-only wrapper for llama.cpp☆30Updated this week
- transformer tokenizers (e.g. BERT tokenizer) in C++ (WIP)☆18Apr 7, 2022Updated 4 years ago
- Port of Microsoft's BioGPT in C/C++ using ggml☆87Feb 21, 2024Updated 2 years ago
- Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++☆6,590Updated this week
- LLM-based code completion engine☆194Jan 23, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Apache Lucene/Solr Guide☆13Oct 14, 2021Updated 4 years ago
- Demo python script app to interact with llama.cpp server using whisper API, microphone and webcam devices.☆47Nov 6, 2023Updated 2 years ago
- ☆133Nov 24, 2023Updated 2 years ago
- This repo is for handling Question Answering, especially for Multi-hop Question Answering☆68Dec 20, 2023Updated 2 years ago
- 🚀 Scale your RAG pipeline using Ragswift: A scalable centralized embeddings management platform☆38Jan 29, 2024Updated 2 years ago
- [ICML 2025] EffiCoder: Enhancing Code Generation in Large Language Models through Efficiency-Aware Fine-tuning☆16May 24, 2025Updated last year
- An implementation of Compositional Attention: Disentangling Search and Retrieval by MILA☆14Jun 1, 2022Updated 4 years ago
- Automatically generate a lip-synced avatar based off of a transcript and audio☆15Feb 17, 2023Updated 3 years ago
- ☆15Sep 8, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Visual Studio Code extension for WizardCoder☆148Aug 1, 2023Updated 2 years ago
- (WACV 2025 - Oral) Vision-language conversation in 10 languages including English, Chinese, French, Spanish, Russian, Japanese, Arabic, H…☆85Aug 5, 2025Updated 11 months ago
- Make any person bald!! Component of the paper: Learning to regulate 3D head shape by removing occluding hair from in-the-wild images.☆12Jun 6, 2022Updated 4 years ago
- ☆63Sep 23, 2024Updated last year
- Go language bindings for the ggwave C++ library☆14Apr 9, 2025Updated last year
- Standalone Flash Attention v2 kernel without libtorch dependency☆113Sep 10, 2024Updated last year
- example of using CoreML from c++☆24Jun 14, 2023Updated 3 years ago