TRELLIS.2 image-to-3D in C++/GGML (CUDA + Vulkan), with a resident HTTP server
☆144Jul 20, 2026Updated this week
Alternatives and similar repositories for trellis.cpp
Users that are interested in trellis.cpp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A bare-bones GUI application for the local inference engine, llama.cpp. Built-in TPE optimiser to find the best flags for your system☆17Jul 3, 2026Updated 2 weeks ago
- Local AI text-to-speech with voice cloning and voice design, powered by GGML. C++17 port of OmniVoice (k2-fsa/OmniVoice). 646 languages, …☆137Updated this week
- An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, a…☆798Updated this week
- A conversational voice-to-voice open-weights LLM-powered assistant designed to run on high-end consumer or workstation class hardware☆71Jul 6, 2026Updated 2 weeks ago
- CPU-Native Language Models☆32May 18, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆13Jun 18, 2024Updated 2 years ago
- RDNA-native LLM inference engine in Rust.☆484Updated this week
- Pure C wrapper library to use llama.cpp with Linux and Windows as simple as possible.☆15Jul 11, 2026Updated last week
- An interface that features barely zero external dependencies beyond the Ollama API itself, making it lightweight and portable to easily i…☆12Mar 25, 2025Updated last year
- Run Orpheus 3B Locally with Gradio UI, Standalone App☆24Apr 1, 2025Updated last year
- Open-source framework for superagents.☆91Jun 16, 2026Updated last month
- A Qt GUI for large language models☆45Nov 17, 2023Updated 2 years ago
- A markdown web renderer for AI agents — see the web without screenshots☆64May 21, 2026Updated last month
- High-performance batched Top-K selection for CPU inference. Up to 80x faster than PyTorch, optimized for LLM sampling with AVX2 SIMD.☆17Mar 20, 2026Updated 4 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Persys desktop. Electron based application to access your Persys server.☆16May 16, 2025Updated last year
- interactive semantic search demo using Qwen3-0.6B-Embedding in your browser☆60Feb 25, 2026Updated 4 months ago
- A framework for efficient model inference with omni-modality models☆29Jul 12, 2026Updated last week
- Image to 3D on your Mac☆145Jul 12, 2026Updated last week
- A framework to allow unlimited text base RPG possibities.☆15Feb 18, 2026Updated 5 months ago
- Generative interactive-textbooks with learning journey stored as a tree structure. Generate recursive branches based off doubts or backtr…☆44May 26, 2026Updated last month
- ☆16Oct 28, 2025Updated 8 months ago
- A minimal, local-first coding agent for the terminal.☆202Jul 12, 2026Updated last week
- A harness optimized to smaller LLMs☆1,778Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Lightweight Llama 3 8B Inference Engine in CUDA C☆53Mar 21, 2025Updated last year
- ☆35Jul 13, 2026Updated last week
- Lightweight C inference for Qwen3 GGUF. Multiturn prefix caching & batch processing.☆25Sep 1, 2025Updated 10 months ago
- ☆57Oct 10, 2025Updated 9 months ago
- Using DSPy to optimize Chat-to-SQL☆15Nov 17, 2025Updated 8 months ago
- Generate Duolingo-style quiz courses from PDFs with spaced repetition, adaptive difficulty, and tutor chat.☆16Apr 6, 2026Updated 3 months ago
- ☆43Oct 23, 2025Updated 8 months ago
- Jacobian-Brainwash : A manual alignment tool for large language models built on Anthropic's Jacobian Lens. Results are exportable.☆182Updated this week
- Production-ready ternary quantized (1.58-bit) Rust code generation model with mHC-lite, MaxRL training, and comprehensive benchmarking☆21Mar 7, 2026Updated 4 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Interactive Jacobian-Lens visualizer and live steerer for GGUF models on llama.cpp☆74Jul 12, 2026Updated last week
- A native .NET LLM inference engine for GGUF models. TensorSharp provides a console application, a web-based chatbot interface, and Ollama…☆225Updated this week
- Using LLMs for iteratively exploring the solution search space at scale.☆732Jul 13, 2026Updated last week
- KVarN, KV cache precision tail, low-bit quants in llama.cpp for longer context of better precision in the same VRAM☆789Updated this week
- OpenMOSS pure C++ pipeline based on GGML☆63Jul 11, 2026Updated last week
- V.I.S.O.R., my in-development AI-powered voice assistant with integrated memory!☆38Nov 20, 2025Updated 8 months ago
- V.I.S.O.R., my in-development AI-powered voice assistant with integrated memory!☆36Nov 20, 2025Updated 8 months ago