This project aims to replicate mainstream open-source model architectures with limited computational resources, implementing mini models with 100-200M parameters.
☆293Aug 4, 2026Updated 2 weeks ago
Alternatives and similar repositories for Mini-LLM
Users that are interested in Mini-LLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.☆4,003Updated this week
- ☆15May 10, 2026Updated 3 months ago
- A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.☆4,767May 17, 2026Updated 3 months ago
- 👀 Train a 65M-parameter VLM from scratch in just 2h!☆8,470Aug 6, 2026Updated last week
- 🧠 Train a 64M-parameter LLM from scratch in just 2h!☆54,808Aug 6, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An LLM training framework built from the ground up, featuring a custom BumbleBee architecture and end-to-end support for multiple open-so…☆100Apr 26, 2026Updated 3 months ago
- 手撕transformer并完成一个简单的机器翻译。☆22Feb 19, 2025Updated last year
- 从零构建大模型:从预训练到RLHF的完整实践☆2,684May 20, 2026Updated 2 months ago
- Nano vLLM☆15,046Apr 26, 2026Updated 3 months ago
- [CVPR 2026 Highlight] Official implementation of Log-linear Sparse Attention (LLSA).☆93May 1, 2026Updated 3 months ago
- GEMV implementation with CUTLASS☆21Aug 21, 2025Updated 11 months ago
- 🎓从0开始训练一个大模型Minimind项目的超详细解析,包括但不限于用到的架构,算法,以及大模型面试经验☆1,624May 25, 2026Updated 2 months ago
- ☆21Apr 21, 2026Updated 3 months ago
- 复现大模型相关算法及一些学习记录☆3,500Jul 2, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- CAD - Memory Efficient Convolutional Adapter for Segment Anything☆12Oct 4, 2024Updated last year
- My implementation of Stanford CS336 assignments.☆246Mar 15, 2026Updated 5 months ago
- ☆48Nov 1, 2025Updated 9 months ago
- Based on Nano-vLLM, a simple replication of vLLM with self-contained paged attention and flash attention implementation☆981Jul 22, 2026Updated 3 weeks ago
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,614Updated this week
- 这是一个从头训练大语言模型的项目,包括预训练、微调和直接偏好优化,模型拥有1B参数,支持中英文。☆864Feb 18, 2025Updated last year
- 我在梦河边浣洗星星。☆43Updated this week
- This is the public repository of AAAI 2024 paper "Is a Large Language Model a Good Annotator for Event Extraction"☆10Feb 16, 2024Updated 2 years ago
- ☆146Jan 18, 2026Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A minimal PyTorch re-implementation of Qwen 3.8☆441Updated this week
- 一份面向实践者的 verl 框架使用教程。verl 是字节跳动开源的大语言模型强化学习训练框架,支持 PPO、GRPO 等多种算法,以及分布式训练、AgentRL 等场景。☆124Jul 2, 2026Updated last month
- ☆55Nov 22, 2025Updated 8 months ago
- A from-scratch Prefill/Decode disaggregation inference engine for LLMs☆164May 10, 2026Updated 3 months ago
- A practical bilingual guide to staying safe and prepared at conferences in Brazil / 巴西参会实用攻略与自救指南☆17Apr 22, 2026Updated 3 months ago
- Agentic RL最详细入门☆286Aug 6, 2026Updated last week
- This repository presents the original implementation of Pretraining Data Detection for Large Language Models: A Divergence-based Calibrat…☆23May 21, 2025Updated last year
- learn LLM inference system on Apple Silicon for systems engineers: build a tiny vLLM + Qwen☆4,501Updated this week
- 从零复现 minimind👉minimind-v☆375Dec 24, 2025Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 强化学习原理 + 强化学习代码实现 + 强化学习框架 + 强化学习论文☆32Updated this week
- This is a repository to practice multi-thread programming in C++☆32Feb 21, 2024Updated 2 years ago
- 轻量级大语言模型MiniMind的源码解读,包含tokenizer、RoPE、MoE、KV Cache、pretraining、SFT、LoRA、DPO等完整流程☆1,146Jun 16, 2025Updated last year
- Co-Reinforcement Learning for Unified Multimodal Understanding and Generation☆48Jul 22, 2025Updated last year
- mini project for nanorllm☆63Mar 31, 2026Updated 4 months ago
- Mini_RWKV_V7_LM Only 34.2M params (also have RWKV7s architecture [deep embedding]/[deep embedding attention) with Full Training code & da…☆92Jan 26, 2026Updated 6 months ago
- Official Implementation of SAM-Decoding: Speculative Decoding via Suffix Automaton☆52May 12, 2026Updated 3 months ago