This project aims to replicate mainstream open-source model architectures with limited computational resources, implementing mini models with 100-200M parameters.
☆292Aug 24, 2026Updated 2 weeks ago
Alternatives and similar repositories for Mini-LLM
Users that are interested in Mini-LLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.☆4,301Updated this week
- 从零开始学大模型Transformer、GPT2、BERT pre-training and fine-tuning from scratch☆41Jul 1, 2024Updated 2 years ago
- ☆15May 10, 2026Updated 3 months ago
- A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.☆4,990May 17, 2026Updated 3 months ago
- 👀 Train a 65M-parameter VLM from scratch in just 2h!☆8,580Aug 6, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The AI-Native SDLC Playbook — Anthropic Claude Academy course as EPUB ebook☆63Updated this week
- 🧠 Train a 64M-parameter LLM from scratch in just 2h!☆59,506Updated this week
- An LLM training framework built from the ground up, featuring a custom BumbleBee architecture and end-to-end support for multiple open-so…☆101Aug 24, 2026Updated 2 weeks ago
- 手撕transformer并完成一个简单的机器翻译。☆22Feb 19, 2025Updated last year
- 从零构建大模型:从预训练到RLHF的完整实践☆2,691Aug 21, 2026Updated 2 weeks ago
- Nano vLLM☆15,333Apr 26, 2026Updated 4 months ago
- GEMV implementation with CUTLASS☆21Aug 21, 2025Updated last year
- 🎓从0开始训练一个大模型Minimind项目的超详细解析,包括但不限于用到的架构,算法,以及大模型面试经验☆1,698May 25, 2026Updated 3 months ago
- 复现大模型相关算法及一些学习记录☆3,517Jul 2, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- CAD - Memory Efficient Convolutional Adapter for Segment Anything☆12Oct 4, 2024Updated last year
- slime is an LLM post-training framework for RL Scaling.☆8,404Updated this week
- My implementation of Stanford CS336 assignments.☆246Mar 15, 2026Updated 5 months ago
- ☆48Nov 1, 2025Updated 10 months ago
- Based on Nano-vLLM, a simple replication of vLLM with self-contained paged attention and flash attention implementation☆1,017Aug 29, 2026Updated last week
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,654Updated this week
- 这是一个从头训练大语言模型的项目,包括预训练、微调和直接偏好优化,模型拥有1B参数,支持中英文。☆872Feb 18, 2025Updated last year
- This is the public repository of AAAI 2024 paper "Is a Large Language Model a Good Annotator for Event Extraction"☆10Feb 16, 2024Updated 2 years ago
- A minimal PyTorch re-implementation of Qwen 3.8☆444Aug 30, 2026Updated last week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆150Jan 18, 2026Updated 7 months ago
- ☆56Nov 22, 2025Updated 9 months ago
- 一份面向实践者的 verl 框架使用教程。verl 是字节跳动开源的大语言模型强化学习训练框架,支持 PPO、GRPO 等多种算法,以及分布式训练、AgentRL 等场景。☆133Jul 2, 2026Updated 2 months ago
- This repository presents the original implementation of Pretraining Data Detection for Large Language Models: A Divergence-based Calibrat…☆23May 21, 2025Updated last year
- 《汇编语言一发入魂》配套代码☆15May 30, 2020Updated 6 years ago
- 从零复现 minimind👉minimind-v☆388Dec 24, 2025Updated 8 months ago
- 强化学习原理 + 强化学习代码实现 + 强化学习框架 + 强化学习论文☆35Updated this week
- This is a repository to practice multi-thread programming in C++☆32Feb 21, 2024Updated 2 years ago
- 轻量级大语言模型MiniMind的源码解读,包含tokenizer、RoPE、MoE、KV Cache、pretraining、SFT、LoRA、DPO等完整流程☆1,179Jun 16, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Co-Reinforcement Learning for Unified Multimodal Understanding and Generation☆48Jul 22, 2025Updated last year
- Mini_RWKV_V7_LM Only 34.2M params (also have RWKV7s architecture [deep embedding]/[deep embedding attention) with Full Training code & da…☆95Jan 26, 2026Updated 7 months ago
- Official Implementation of SAM-Decoding: Speculative Decoding via Suffix Automaton☆53May 12, 2026Updated 3 months ago
- 使用 cutlass 实现 flash-attention 精简版,具有教学意义☆59Aug 12, 2024Updated 2 years ago
- some notes for opensource llm technical reorts☆20Mar 6, 2026Updated 6 months ago
- MEME: Multi-Entity & Evolving Memory Evaluation — reference implementation (companion to arXiv preprint)☆24May 11, 2026Updated 3 months ago
- Retriever-0.1B☆94Jun 6, 2024Updated 2 years ago