☆63Jul 10, 2025Updated last year
Alternatives and similar repositories for rekaquant
Users that are interested in rekaquant are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Yet another frontend for LLM, written using .NET and WinUI 3☆11Sep 14, 2025Updated 11 months ago
- A sleek, customizable interface for managing LLMs with responsive design and easy agent personalization.☆19Aug 30, 2024Updated 2 years ago
- Thin wrapper around GGML to make life easier☆48Jul 26, 2026Updated last month
- The official repo for “Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem” [EMNLP25]☆33Sep 1, 2025Updated last year
- The official implementation of NOSA☆20Jun 11, 2026Updated 2 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Efficient non-uniform quantization with GPTQ for GGUF☆64Sep 17, 2025Updated 11 months ago
- ☆52May 20, 2025Updated last year
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated last year
- Load and run Llama from safetensors files in C☆16Oct 24, 2024Updated last year
- An extention to the GaLore paper, to perform Natural Gradient Descent in low rank subspace☆19Oct 21, 2024Updated last year
- a single interface around speech-to-speech foundation models☆28Jun 27, 2025Updated last year
- RWKV is a RNN with transformer-level LLM performance. It can be directly trained like a GPT (parallelizable). So it's combining the best …☆10Nov 3, 2023Updated 2 years ago
- T-Projection is a method to perform high-quality Annotation Projection of Sequence Labeling datasets.☆13Nov 21, 2023Updated 2 years ago
- Official implementation for DenseMixer: Improving MoE Post-Training with Precise Router Gradient☆68Aug 3, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 4 months ago
- Proteus is an experimental platform that combines the power of Large Language Models with the Genesis physics engine☆25Dec 20, 2024Updated last year
- A prototype agent with the purpose of evaluating the performance of a Large Language Model within a python terminal.☆13Aug 28, 2023Updated 3 years ago
- Sparse Inferencing for transformer based LLMs☆222Mar 25, 2026Updated 5 months ago
- Physics Master is a model fine-tuned from llama3-8B-Instruct. It can answer your physics question!☆16Aug 24, 2024Updated 2 years ago
- Tcurtsni: Reverse Instruction Chat, ever wonder what your LLM wants to ask you?☆23Jun 25, 2024Updated 2 years ago
- ☆21Jun 8, 2025Updated last year
- Jina VDR is a multilingual, multi-domain benchmark for visual document retrieval☆38Aug 4, 2025Updated last year
- ☆112Aug 21, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Use Codestral Mamba with Visual Studio Code and the Continue extension. A local LLM alternative to GitHub Copilot.☆31Jul 18, 2024Updated 2 years ago
- ☆22Jun 5, 2025Updated last year
- ☆51May 31, 2024Updated 2 years ago
- AI Based "Happiness Optimizer"☆12Oct 20, 2024Updated last year
- ☆17Updated this week
- A real-time shared memory layer for multi-agent LLM systems.☆68Jan 12, 2026Updated 7 months ago
- Nexusflow function call, tool use, and agent benchmarks.☆28Dec 13, 2024Updated last year
- ☆13Apr 15, 2024Updated 2 years ago
- MACKO: Sparse matrix vector multiplication for low sparsity☆41Apr 6, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Profile-guided GPU kernel optimizer for AMD/RDNA3. Auto-tunes llama.cpp MMVQ kernels per model shape. 2x decode speedup on 7900 XTX.☆67Updated this week
- REAP: Router-weighted Expert Activation Pruning for SMoE compression☆495Apr 17, 2026Updated 4 months ago
- Pytorch implementation of "Oscillation-Reduced MXFP4 Training for Vision Transformers" on DeiT Model Pre-training☆41May 4, 2026Updated 4 months ago
- Produce your own Dynamic 3.0 Quants and achieve optimum accuracy & SOTA quantization performance! Input a target size and the toolchain w…☆159Aug 21, 2026Updated 2 weeks ago
- Model souping for LLMs☆75Nov 18, 2025Updated 9 months ago
- An optimized quantization and inference library for running LLMs locally on modern consumer-class GPUs☆1,312Updated this week
- Quartet II Official Code☆82May 1, 2026Updated 4 months ago