This project aims to replicate mainstream open-source model architectures with limited computational resources, implementing mini models with 100-200M parameters.
☆284Jun 14, 2026Updated last month
Alternatives and similar repositories for Mini-LLM
Users that are interested in Mini-LLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.☆3,385Jul 3, 2026Updated 3 weeks ago
- 从零开始学大模型Transformer、GPT2、BERT pre-training and fine-tuning from scratch☆41Jul 1, 2024Updated 2 years ago
- ☆15May 10, 2026Updated 2 months ago
- 👀「大模型」2小时从0训练65M参数的视觉多模态VLM!Train a 65M-parameter VLM from scratch in just 2h!☆8,396Jun 28, 2026Updated last month
- 🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!☆54,029Jul 23, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An LLM training framework built from the ground up, featuring a custom BumbleBee architecture and end-to-end support for multiple open-so…☆100Apr 26, 2026Updated 3 months ago
- 手撕transformer并完成一个简单的机器翻译。☆23Feb 19, 2025Updated last year
- 从零构建大模型:从预训练到RLHF的完整实践☆2,679May 20, 2026Updated 2 months ago
- Nano vLLM☆14,679Apr 26, 2026Updated 3 months ago
- [CVPR 2026 Highlight] Official implementation of Log-linear Sparse Attention (LLSA).☆92May 1, 2026Updated 2 months ago
- GEMV implementation with CUTLASS☆21Aug 21, 2025Updated 11 months ago
- 🎓从0开始训练一个大模型Minimind项目的超详细解析,包括但不限于用到的架构,算法,以及大模型面试经验☆1,271May 25, 2026Updated 2 months ago
- 复现大模型相关算法及一些学习记录☆3,471Jul 2, 2026Updated 3 weeks ago
- slime is an LLM post-training framework for RL Scaling.☆7,696Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- My implementation of Stanford CS336 assignments.☆246Mar 15, 2026Updated 4 months ago
- ☆48Nov 1, 2025Updated 8 months ago
- Based on Nano-vLLM, a simple replication of vLLM with self-contained paged attention and flash attention implementation☆944Jul 22, 2026Updated last week
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,581Updated this week
- 这是一个从头训练大语言模型的项目,包括预训练、微调和直接偏好优化,模型拥有1B参数,支持中英文。☆862Feb 18, 2025Updated last year
- This is the public repository of AAAI 2024 paper "Is a Large Language Model a Good Annotator for Event Extraction"☆10Feb 16, 2024Updated 2 years ago
- ☆143Jan 18, 2026Updated 6 months ago
- A minimal PyTorch re-implementation of Qwen 3.5☆432Jun 15, 2026Updated last month
- 一份面向实践者的 verl 框架使用教程。verl 是字节跳动开源的大语言模型强化学习训练框架,支持 PPO、GRPO 等多种算法,以及分布式训练、AgentRL 等场景。☆112Jul 2, 2026Updated 3 weeks ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆55Nov 22, 2025Updated 8 months ago
- A from-scratch Prefill/Decode disaggregation inference engine for LLMs☆160May 10, 2026Updated 2 months ago
- Agentic RL最详细入门☆200Updated this week
- This repository presents the original implementation of Pretraining Data Detection for Large Language Models: A Divergence-based Calibrat…☆23May 21, 2025Updated last year
- A course of learning LLM inference serving on Apple Silicon for systems engineers: build a tiny vLLM + Qwen.☆4,425Updated this week
- 从零复现 minimind👉minimind-v☆371Dec 24, 2025Updated 7 months ago
- 《汇编语言一发入魂》配套代码☆15May 30, 2020Updated 6 years ago
- 强化学习原理 + 强化学习代码实现 + 强化学习框架 + 强化学习论文☆32Jul 19, 2026Updated last week
- This is a repository to practice multi-thread programming in C++☆31Feb 21, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- 轻量级大语言模型MiniMind的源码解读,包含tokenizer、RoPE、MoE、KV Cache、pretraining、SFT、LoRA、DPO等完整流程☆1,126Jun 16, 2025Updated last year
- Co-Reinforcement Learning for Unified Multimodal Understanding and Generation☆48Jul 22, 2025Updated last year
- mini project for nanorllm☆64Mar 31, 2026Updated 3 months ago
- Mini_RWKV_V7_LM Only 34.2M params (also have RWKV7s architecture [deep embedding]/[deep embedding attention) with Full Training code & da…☆90Jan 26, 2026Updated 6 months ago
- Official Implementation of SAM-Decoding: Speculative Decoding via Suffix Automaton☆52May 12, 2026Updated 2 months ago
- some notes for opensource llm technical reorts☆20Mar 6, 2026Updated 4 months ago
- Flash Attention from Scratch on CUDA Ampere☆188Sep 1, 2025Updated 10 months ago