从零复现 minimind👉minimind-v
☆384Dec 24, 2025Updated 8 months ago
Alternatives and similar repositories for minimind-learn
Users that are interested in minimind-learn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 轻量级大语言模型MiniMind的源码解读,包含tokenizer、RoPE、MoE、KV Cache、pretraining、SFT、LoRA、DPO等完整流程☆1,173Jun 16, 2025Updated last year
- 🎓从0开始训练一个大模型Minimind项目的超详细解析,包括但不限于用到的架构,算法,以及大模型面试经验☆1,683May 25, 2026Updated 3 months ago
- 📖 从零基础到面试通关 —— 22节课彻底搞懂大语言模型 | Learn MiniMind: 系统化学习LLM训练全流程☆544Apr 1, 2026Updated 5 months ago
- 👀 Train a 65M-parameter VLM from scratch in just 2h!☆8,545Aug 6, 2026Updated 3 weeks ago
- MiniMind-V 多模态面试学习指南 - 20节课程 + 278道面试题 + STAR面试稿 + 哆啦A梦漫画☆148Apr 2, 2026Updated 5 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- 🚀 [从零构建 LLM] 极简大模型训练原理与实践指南。包含 Transformer, Pretraining, SFT 核心代码与对照实验。 | A minimal, principle-first guide to understanding and building…☆182Jun 4, 2026Updated 3 months ago
- ☆172Mar 18, 2026Updated 5 months ago
- 一个完整的 LLM 训练的基本流程笔记 (Tokenizer -> PreTraining -> SFT -> DPO -> GRPO)☆27Feb 23, 2026Updated 6 months ago
- ☆44Nov 22, 2025Updated 9 months ago
- 三元三小时手敲大模型☆590Mar 12, 2026Updated 5 months ago
- MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training Pipeline. 训练医疗大模型,实现了包括增量预训练(PT)、有监督微调(SFT)、RLHF、DPO、ORPO、GRPO。☆5,775Jun 3, 2026Updated 3 months ago
- ☆19Dec 3, 2025Updated 9 months ago
- Implementation of KDR-Agent, the AAAI 2025 accepted paper, focusing on knowledge-driven reasoning for autonomous agents.☆24Nov 24, 2025Updated 9 months ago
- 🚀 2026届大模型算法岗实习面经 | 包含 DeepSeek/Qwen 技术报告解析、手撕 PPO/RoPE/Transformer、RLHF 核心与八股文 | 持续更新中...☆665Mar 28, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 一个生产级的深度研究 Agent 系统,从零构建多智能体编排、Red-Blue 对抗降噪、 语义级上下文压缩、跨 Agent 共享记忆四大核心能力,配套 165 次独立实验 + Bootstrap 统计显著性检验的完整评测体系。☆128May 11, 2026Updated 3 months ago
- Use interactive notebook to break down MiniMind code and learn from scratch.☆157Jan 7, 2026Updated 7 months ago
- 主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题☆15,020Jun 14, 2026Updated 2 months ago
- ☆108Jul 24, 2025Updated last year
- ☆56Nov 22, 2025Updated 9 months ago
- 🔍大模型应用开发实战一:RAG 技术全栈指南,在线阅读地址:https://datawhalechina.github.io/all-in-rag/☆10,790Updated this week
- AI拆解论文,人人都能读懂前沿研究☆24Jul 10, 2026Updated last month
- Personal Project: MPP-Qwen14B & MPP-Qwen-Next(Multimodal Pipeline Parallel based on Qwen-LM). Support [video/image/multi-image] {sft/conv…☆686Mar 10, 2025Updated last year
- This project aims to replicate mainstream open-source model architectures with limited computational resources, implementing mini models …☆292Aug 24, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- CS336 面试导向学习指南 - Stanford Language Modeling from Scratch☆188Apr 2, 2026Updated 5 months ago
- 记录我在cs336学习时的笔记和作业☆1,106May 2, 2026Updated 4 months ago
- [EMNLP 2026] Official implementation for paper "Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive R…☆45Aug 23, 2026Updated last week
- Extending HSTU with Semantic IDs: reproducible MovieLens experiments and enriched MovieLens metadata☆19Sep 16, 2025Updated 11 months ago
- 📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程☆76,784Aug 18, 2026Updated 2 weeks ago
- 2025年腾讯广告算法大赛,“调教大师”队伍初赛方案。初赛排名150+☆16Mar 16, 2026Updated 5 months ago
- Mini_RWKV_V7_LM Only 34.2M params (also have RWKV7s architecture [deep embedding]/[deep embedding attention) with Full Training code & da…☆94Jan 26, 2026Updated 7 months ago
- Research Agent for multi-hop reasoning QA — 3rd place🥉 in 2026 Alibaba Cloud Data+AI Global Competition☆31Mar 16, 2026Updated 5 months ago
- A comparison of deepseek grpo and qwen gspo on Qwen2.5-1.5B-Instruct fine tunning.☆170Mar 28, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 新手友好的基于 Qwen2 的生成式推荐系统,通过大模型理解用户偏好生成候选物品,融合 TF-IDF 关键词召回、热门物品召回构建多源策略,平衡相关性与多样性。采用 LightGBM 排序模型精准打分,内置召回率、NDCG 等评估指标量化效果。通过 Flask 封装 API…☆54Dec 6, 2025Updated 8 months ago
- 《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程☆32,050Aug 26, 2026Updated last week
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"☆17Feb 15, 2026Updated 6 months ago
- 2025腾讯广告算法大赛-分数0.0978代码☆102Sep 17, 2025Updated 11 months ago
- 一份面向实践者的 verl 框架使用教程。verl 是字节跳动开源的大语言模型强化学习训练框架,支持 PPO、GRPO 等多种算法,以及 分布式训练、AgentRL 等场景。☆132Jul 2, 2026Updated 2 months ago
- 将SmolVLM2的视觉头与Qwen3-0.6B模型进行了拼接微调☆608Sep 8, 2025Updated 11 months ago
- Minimal reproduction of OneRec☆1,783May 14, 2026Updated 3 months ago