KoboldCpp Smart Launcher with GPU Layer and Tensor Override Tuning
☆30May 18, 2025Updated last year
Alternatives and similar repositories for TensorTune
Users that are interested in TensorTune are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Testbench for llama.cpp llama-server☆15Aug 20, 2025Updated last year
- Specialized AI agents for your bare metal. 100% On-Prem & Air-Gap ready.☆26Feb 21, 2026Updated 6 months ago
- Eidos – A Self-Growing AI Agent with Long-Term Memory and Environmental Awareness☆23Jul 4, 2025Updated last year
- CI scripts designed to build a Pascal-compatible version of vLLM.☆13Aug 10, 2024Updated 2 years ago
- Second Brain is a desktop application that acts as a personal knowledge base, using retrieval-augmented generation (RAG), multimodal AI m…☆27Jan 30, 2026Updated 7 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 🤖 AI-powered CLI for file reorganization. Runs fully locally — no data leaves your machine.☆20Jul 2, 2025Updated last year
- Decentralizing distribution of open-source AI models.☆22Updated this week
- ☆20Aug 19, 2025Updated last year
- An OpenAI API compatible images server to generate or manipulate images.☆18Feb 2, 2025Updated last year
- A user-friendly GUI for llama.cpp — convert, quantize, and run GGUF models without touching the terminal.☆23Jun 9, 2026Updated 2 months ago
- A PowerShell script to fully automate the setup of `llama.cpp` on Windows. It installs all prerequisites, including the correct CUDA Tool…☆17May 12, 2026Updated 3 months ago
- A fast batching API to serve LLM models☆189Apr 26, 2024Updated 2 years ago
- ☆16Dec 16, 2024Updated last year
- A highly optimized engine for maya-1 tts model to generate minutes of audio in seconds.☆67Nov 17, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18Jul 1, 2025Updated last year
- Create text chunks which end at natural stopping points without using a tokenizer☆26Nov 26, 2025Updated 9 months ago
- Stable Diffusion and Flux in pure C/C++☆26Aug 13, 2026Updated 2 weeks ago
- ☆47Apr 29, 2026Updated 4 months ago
- Generate a wiki for your research topic, sourcing from the web and your docs.☆58Mar 8, 2025Updated last year
- A project combining roguelike with LLMs, RAG, Text2Speech, and Speech2Text☆19Mar 4, 2025Updated last year
- ☆51Feb 19, 2025Updated last year
- Personal voice assistant, with voice interruption and Twilio support☆18Feb 24, 2025Updated last year
- CompChomper is a framework for measuring how LLMs perform at code completion.☆21Apr 29, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A simple, "Ollama-like" tool for managing and running GGUF language models from your terminal.☆25Jan 2, 2026Updated 7 months ago
- Evolution process to find the best quant tensor weights to build the most optimal GGUF options for an AI model.☆43Aug 22, 2026Updated last week
- FlashQLA TileLang GDN kernels ported to NVIDIA Blackwell consumer (GB10 / DGX Spark)☆17Jun 5, 2026Updated 2 months ago
- Talk to your data. Instantly analyze, visualize, and transform☆22Oct 30, 2025Updated 10 months ago
- 🎮 Material You TUI for monitoring NVIDIA GPUs☆58Jan 16, 2026Updated 7 months ago
- Cohere Toolkit is a collection of prebuilt components enabling users to quickly build and deploy RAG applications.☆30Jan 19, 2025Updated last year
- ☆18Aug 15, 2023Updated 3 years ago
- Implementation of Qwen3-ASR-0.6B in GGML☆110Jul 28, 2026Updated last month
- ☆43Oct 23, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Transform your pdfs into anki flashcard with gpt☆12Jun 10, 2024Updated 2 years ago
- 🚀 Scale your RAG pipeline using Ragswift: A scalable centralized embeddings management platform☆38Jan 29, 2024Updated 2 years ago
- Useful views and functions for postgreSQL DBA's.☆15Apr 18, 2016Updated 10 years ago
- ☆15Jul 3, 2025Updated last year
- ☆51Oct 1, 2025Updated 10 months ago
- A library and CLI utilities for managing performance states of NVIDIA GPUs.☆37Oct 6, 2024Updated last year
- Benchmarks testing the performance of various releases of Pydantic v2 🦀☆13Dec 2, 2024Updated last year