Everything about the SmolLM and SmolVLM family of models
☆3,865May 26, 2026Updated 2 months ago
Alternatives and similar repositories for smollm
Users that are interested in smollm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Minimalistic large language model 3D-parallelism training☆2,779May 26, 2026Updated 2 months ago
- A course on aligning smol models.☆6,725May 26, 2026Updated 2 months ago
- The simplest, fastest repository for training/finetuning small-sized VLMs.☆4,979Oct 27, 2025Updated 9 months ago
- Lighteval is your all-in-one toolkit for evaluating LLMs across multiple backends☆2,510Jun 29, 2026Updated last month
- Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek-V4, GLM and other models.☆69,717Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Fully open reproduction of DeepSeek-R1☆26,428Apr 2, 2026Updated 4 months ago
- Train transformer language models with reinforcement learning.☆19,027Updated this week
- 🤗 smolagents: a barebones library for agents that think in code.☆28,728Jul 21, 2026Updated 2 weeks ago
- Minimalistic 4D-parallelism distributed training framework for education purpose☆2,274Aug 26, 2025Updated 11 months ago
- Distilabel is a framework for synthetic data and AI feedback for engineers who need fast, reliable and scalable pipelines based on verifi…☆3,359Updated this week
- Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.☆3,253Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆88,524Updated this week
- AllenAI's post-training codebase☆3,820Updated this week
- NanoGPT (124M) in 90 seconds☆5,648Aug 2, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Sharing both practical insights and theoretical knowledge about LLM evaluation that we gathered while managing the Open LLM Leaderboard a…☆2,136Dec 3, 2025Updated 8 months ago
- Go ahead and axolotl questions☆12,328Updated this week
- 20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.☆13,611Jul 20, 2026Updated 2 weeks ago
- Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We als…☆18,554May 19, 2026Updated 2 months ago
- Minimal reproduction of DeepSeek R1-Zero☆13,217Feb 27, 2026Updated 5 months ago
- MobileLLM Optimizing Sub-billion Parameter Language Models for On-Device Use Cases. In ICML 2024.☆1,456Apr 30, 2026Updated 3 months ago
- Recipes for shrinking, optimizing, customizing cutting edge vision models. 💜☆1,968May 26, 2026Updated 2 months ago
- Structured Outputs☆15,539Updated this week
- Robust recipes to align language models with human and AI preferences☆5,657May 26, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- PyTorch native post-training library☆5,794Updated this week
- SGLang is a high-performance serving framework for large language models and multimodal models.☆31,535Updated this week
- Tools for merging pretrained large language models.☆7,289Jun 17, 2026Updated last month
- Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.☆19,747Jan 30, 2026Updated 6 months ago
- DSPy: The framework for programming—not prompting—language models☆36,702Updated this week
- Large Language Model Text Generation Inference☆10,888Mar 21, 2026Updated 4 months ago
- Meta Lingua: a lean, efficient, and easy-to-hack codebase to research LLMs.☆4,763Jul 18, 2025Updated last year
- A lightweight, local-first, and free experiment tracking library from Hugging Face 🤗☆1,627Updated this week
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,869Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A framework for few-shot evaluation of language models.☆13,576Jul 13, 2026Updated 3 weeks ago
- Modeling, training, eval, and inference code for OLMo☆6,619Nov 24, 2025Updated 8 months ago
- The TinyLlama project is an open endeavor to pretrain a 1.1B Llama model on 3 trillion tokens.☆9,020May 3, 2024Updated 2 years ago
- Fast and memory-efficient exact attention☆24,654Updated this week
- Democratizing Reinforcement Learning for LLMs☆5,770Updated this week
- Efficient Triton Kernels for LLM Training☆6,556Updated this week
- Recipes to scale inference-time compute of open models☆1,131May 26, 2026Updated 2 months ago