Evolution process to find the best quant tensor weights to build the most optimal GGUF options for an AI model.
☆45Sep 8, 2026Updated this week
Alternatives and similar repositories for MagicQuant
Users that are interested in MagicQuant are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12May 30, 2025Updated last year
- ☆20Sep 28, 2024Updated last year
- Produce your own Dynamic 3.0 Quants and achieve optimum accuracy & SOTA quantization performance! Input a target size and the toolchain w…☆164Updated this week
- ☆22Updated this week
- Desktop application for instant AI-powered text transformation. Translate, correct, summarize, and change the tone of any text, anywhere,…☆35Dec 29, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A systematic empirical study of self-verification strategies in agentic coding harnesses☆29Mar 4, 2026Updated 6 months ago
- ☆16Feb 1, 2025Updated last year
- KoboldCpp Smart Launcher with GPU Layer and Tensor Override Tuning☆30May 18, 2025Updated last year
- A highly optimized engine for maya-1 tts model to generate minutes of audio in seconds.☆66Nov 17, 2025Updated 9 months ago
- CPU-Native Language Models☆34May 18, 2026Updated 3 months ago
- Llama.cpp launcher with integrated huggingface☆71Updated this week
- Llama Server Launcher (llama.cpp/ik_llama) GUI☆125Jul 22, 2026Updated last month
- Containerized Opinionated Agent Orchestration Platform☆15Mar 8, 2026Updated 6 months ago
- GUI tool to QLoRA/LoRA-fine-tune LLMs and deploy to Ollama. Broad GPU support (NVIDIA/AMD/Intel/Apple) + CPU fallback.☆15Feb 18, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Mar 18, 2026Updated 5 months ago
- LLM inference in C/C++☆17Jul 29, 2026Updated last month
- SD.Next Quantization Engine☆133Aug 30, 2026Updated last week
- Browser extension that lets you summarize and chat with any webpage using a local LLM of your choice.☆23Oct 24, 2024Updated last year
- ☆13Mar 5, 2025Updated last year
- Group-relative Trajectory-based Policy Optimization: Increasing Quality and Training Stability☆42Feb 23, 2026Updated 6 months ago
- Auto-Video maker handling many AI's☆11Mar 18, 2024Updated 2 years ago
- ☆23May 14, 2026Updated 3 months ago
- Cohere Toolkit is a collection of prebuilt components enabling users to quickly build and deploy RAG applications.☆30Jan 19, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Bookmarklet to pull and run hugging face GGUF models in Ollama☆18Oct 17, 2024Updated last year
- Add an image reference feature to the Cosmos model or models based on it, such as the Anima model☆55Jun 19, 2026Updated 2 months ago
- Analyze binaries and generate structured reports for AI agents and security research.☆21May 13, 2026Updated 3 months ago
- ☆20Updated this week
- Monorepo for sharing my most commonly used Nix expressions between projects.☆30Updated this week
- Next-gen AI memory layer with importance scoring, temporal decay, hierarchical memory, and YMYL prioritization☆48Updated this week
- Docker network setup with authelia, caddy, crowdsec, and wg-easy☆19Aug 7, 2026Updated last month
- ☆33Jan 28, 2025Updated last year
- A small set of unique adapters meant to bridge the dual_stream_shunt trained for guiding prompt embeddings and diffusion.☆14Nov 26, 2025Updated 9 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Nacrith — Lossless text compression via ensemble neural arithmetic coding. Combines SmolLM2-135M language model with context mixing, adap…☆22Mar 21, 2026Updated 5 months ago
- A Streamlit app for generating high-quality Q&A training datasets from text and PDFs, leveraging Gemini, Claude, and OpenAI for LLM fine-…☆41Jul 5, 2025Updated last year
- Adaptive Precision for EXpert Models: MoE-aware mixed-precision quantization☆467Aug 17, 2026Updated 3 weeks ago
- ☆16Jul 23, 2026Updated last month
- ☆51Feb 19, 2025Updated last year
- empirically chooses -ngl param for llama.cpp☆20Mar 19, 2025Updated last year
- ☆27Aug 16, 2025Updated last year