An Open Source Toolkit For LLM Distillation
☆1,079May 12, 2026Updated 4 months ago
Alternatives and similar repositories for DistillKit
Users that are interested in DistillKit are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A pipeline for LLM knowledge distillation☆119May 7, 2026Updated 5 months ago
- EvolKit is an innovative framework designed to automatically enhance the complexity of instructions used for fine-tuning Large Language M…☆262Oct 30, 2024Updated last year
- Automated Identification of Redundant Layer Blocks for Pruning in Large Language Models☆270Apr 23, 2024Updated 2 years ago
- Tools for merging pretrained large language models.☆7,392Sep 12, 2026Updated 3 weeks ago
- Easy to use, High Performant Knowledge Distillation for LLMs☆98May 5, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆146Sep 29, 2026Updated last week
- a toolkit on knowledge distillation for large language models☆480Sep 25, 2026Updated 2 weeks ago
- Distilabel is a framework for synthetic data and AI feedback for engineers who need fast, reliable and scalable pipelines based on verifi…☆3,412Updated this week
- Best practices for distilling large language models.☆642Feb 1, 2024Updated 2 years ago
- ☆57Nov 6, 2024Updated last year
- Official PyTorch implementation of DistiLLM: Towards Streamlined Distillation for Large Language Models (ICML 2024)☆269Mar 13, 2025Updated last year
- [ICLR 2025] Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing. Your efficient and high-quality synthetic data …☆886Mar 17, 2025Updated last year
- Go ahead and axolotl questions☆12,545Updated this week
- This repository collects papers for "A Survey on Knowledge Distillation of Large Language Models". We break down KD into Knowledge Elicit…☆1,316Mar 9, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Minimalistic large language model 3D-parallelism training☆2,834Updated this week
- This is our own implementation of 'Layer Selective Rank Reduction'☆240Sep 29, 2026Updated last week
- AllenAI's post-training codebase☆3,882Updated this week
- Agentic RL Training at Scale☆2,137Updated this week
- Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends☆2,553Updated this week
- Efficient Triton Kernels for LLM Training☆6,652Updated this week
- Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.☆3,377Updated this week
- A family of compressed models obtained via pruning and knowledge distillation☆386Nov 6, 2025Updated 11 months ago
- Curated list of datasets and tools for post-training.☆4,801Apr 29, 2026Updated 5 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs☆3,836May 28, 2026Updated 4 months ago
- Our library for RL environments + evals☆4,686Updated this week
- Fast Multimodal Semantic Deduplication & Filtering☆973Updated this week
- Democratizing Reinforcement Learning for LLMs☆5,860Updated this week
- A framework for few-shot evaluation of language models.☆14,166Sep 14, 2026Updated 3 weeks ago
- Official implementation of Half-Quadratic Quantization (HQQ)☆959Feb 26, 2026Updated 7 months ago
- Robust recipes to align language models with human and AI preferences☆5,691Sep 23, 2026Updated 2 weeks ago
- State-of-the-art LLM compression, built for production inference with vLLM☆3,859Updated this week
- OpenDiLoCo: An Open-Source Framework for Globally Distributed Low-Communication Training☆596Jul 17, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Fast State-of-the-Art Static Embeddings☆2,232Updated this week
- [ICLR 2025] MiniPLM: Knowledge Distillation for Pre-Training Language Models☆81Nov 23, 2024Updated last year
- The official implementation of Self-Play Fine-Tuning (SPIN)☆1,253May 8, 2024Updated 2 years ago
- slime is an LLM post-training framework for RL Scaling.☆8,615Updated this week
- SGLang is a high-performance serving framework for large language models and multimodal models.☆36,925Updated this week
- Synthetic data curation for post-training and structured data extraction☆1,743Updated this week
- Late Interaction Models Training & Retrieval☆897Jul 23, 2026Updated 2 months ago