Evolution process to find the best quant tensor weights to build the most optimal GGUF options for an AI model.
☆40May 13, 2026Updated 2 months ago
Alternatives and similar repositories for MagicQuant-Wiki
Users that are interested in MagicQuant-Wiki are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12May 30, 2025Updated last year
- ☆34Jul 19, 2026Updated 3 weeks ago
- ☆20Sep 28, 2024Updated last year
- Produce your own Dynamic 3.0 Quants and achieve optimum accuracy & SOTA quantization performance! Input a target size and the toolchain w…☆148Aug 1, 2026Updated last week
- Private-first, self-hostable knowledge base. Your data, your server, your control. No cloud, no telemetry, no trust required.☆22May 26, 2026Updated 2 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- A systematic empirical study of self-verification strategies in agentic coding harnesses☆29Mar 4, 2026Updated 5 months ago
- Desktop application for instant AI-powered text transformation. Translate, correct, summarize, and change the tone of any text, anywhere,…☆35Dec 29, 2025Updated 7 months ago
- ☆16Feb 1, 2025Updated last year
- Experimental implementation of DeepSeek v4 flaash in llama.cpp☆24Apr 30, 2026Updated 3 months ago
- KoboldCpp Smart Launcher with GPU Layer and Tensor Override Tuning☆30May 18, 2025Updated last year
- A highly optimized engine for maya-1 tts model to generate minutes of audio in seconds.☆67Nov 17, 2025Updated 8 months ago
- CPU-Native Language Models☆32May 18, 2026Updated 2 months ago
- 🎭 Battle-tested plugins for Hermes Agent — zero core patches. Native vision bypass, multi-agent context injection, and more.☆46Jun 17, 2026Updated last month
- Llama.cpp launcher with integrated huggingface☆68Jul 17, 2026Updated 3 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Llama Server Launcher (llama.cpp/ik_llama) GUI☆124Jul 22, 2026Updated 2 weeks ago
- High-performance batched Top-K selection for CPU inference. Up to 80x faster than PyTorch, optimized for LLM sampling with AVX2 SIMD.☆18Mar 20, 2026Updated 4 months ago
- Containerized Opinionated Agent Orchestration Platform☆15Mar 8, 2026Updated 5 months ago
- GUI tool to QLoRA/LoRA-fine-tune LLMs and deploy to Ollama. Broad GPU support (NVIDIA/AMD/Intel/Apple) + CPU fallback.☆14Feb 18, 2026Updated 5 months ago
- Ultra-Sparse Adaptation of 1-Bit LLMs via XOR Patches☆86Updated this week
- ☆15Mar 18, 2026Updated 4 months ago
- Fast CLI to extract folders or files from GitHub repos using sparse-checkout☆20Jul 4, 2025Updated last year
- SD.Next Quantization Engine☆124Updated this week
- 🧪 A command line pipe inspection utility.☆18Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Browser extension that lets you summarize and chat with any webpage using a local LLM of your choice.☆23Oct 24, 2024Updated last year
- Group-relative Trajectory-based Policy Optimization: Increasing Quality and Training Stability☆42Feb 23, 2026Updated 5 months ago
- ComfyUI-VRAM-Manager is an independent memory management custom node for ComfyUI. Provides Distorch memory management functionality for e…☆46Updated this week
- Auto-Video maker handling many AI's☆11Mar 18, 2024Updated 2 years ago
- Bookmarklet to pull and run hugging face GGUF models in Ollama☆18Oct 17, 2024Updated last year
- A tool that can be used to measure the sequential performance of any OpenAI-compatible LLM API☆25Aug 1, 2024Updated 2 years ago
- Analyze binaries and generate structured reports for AI agents and security research.☆20May 13, 2026Updated 2 months ago
- A Conversational Speech Generation Model☆14Mar 16, 2025Updated last year
- Next-gen AI memory layer with importance scoring, temporal decay, hierarchical memory, and YMYL prioritization☆48Aug 1, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Docker network setup with authelia, caddy, crowdsec, and wg-easy☆19Updated this week
- colab list for video☆12Jul 18, 2026Updated 3 weeks ago
- Official implementation of "Figure It Out: Improve the Frontier of Reasoning with Active Visual Thinking"☆17Jan 13, 2026Updated 6 months ago
- Zoof is a high-efficiency Small Language Model (SLM) engineered from scratch. It demonstrates how modern architectural choices and high-q…☆47Jan 13, 2026Updated 6 months ago
- ☆33Jan 28, 2025Updated last year
- Vector database driven Sillytavern Memory System☆52Updated this week
- A small set of unique adapters meant to bridge the dual_stream_shunt trained for guiding prompt embeddings and diffusion.☆14Nov 26, 2025Updated 8 months ago