从零复现 minimind👉minimind-v
☆400Dec 24, 2025Updated 9 months ago
Alternatives and similar repositories for minimind-learn
Users that are interested in minimind-learn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 轻量级大语言模型MiniMind的源码解读,包含tokenizer、RoPE、MoE、KV Cache、pretraining、SFT、LoRA、DPO等完整流程☆1,205Jun 16, 2025Updated last year
- 🎓从0开始训练一个大模型Minimind项目的超详细解析,包括但不限于用到的架构,算法,以及大模型面试经验☆1,767May 25, 2026Updated 3 months ago
- 📖 从零基础到面试通关 —— 22节课彻底搞懂大语言模型 | Learn MiniMind: 系统化学习LLM训练全流程☆590Apr 1, 2026Updated 5 months ago
- 👀 Train a 65M-parameter VLM from scratch in just 2h!☆8,678Updated this week
- MiniMind-V 多模态面试学习指南 - 20节课程 + 278道面试题 + STAR面试稿 + 哆啦A梦漫画☆156Apr 2, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 一个完整的 LLM 训练的基本流程笔记 (Tokenizer -> PreTraining -> SFT -> DPO -> GRPO)☆27Feb 23, 2026Updated 7 months ago
- ☆44Nov 22, 2025Updated 10 months ago
- 三元三小时手敲大模型☆616Mar 12, 2026Updated 6 months ago
- 🎙️ A 0.1B Omni model trained from scratch, capable of listening, speaking, and seeing!☆2,574Updated this week
- Implementation of KDR-Agent, the AAAI 2025 accepted paper, focusing on knowledge-driven reasoning for autonomous agents.☆24Nov 24, 2025Updated 10 months ago
- 基于轻量级 LLM 与 Qwen2.5-1.5B 两条主线,完成从数据处理、模型训练、参数高效微调,到评测验证与服务部署的端到端闭环。☆229Apr 21, 2026Updated 5 months ago
- 🚀 2026届大模型算法岗实习面经 | 包含 DeepSeek/Qwen 技术报告解析、手撕 PPO/RoPE/Transformer、RLHF 核心与八股文 | 持续更新中...☆690Mar 28, 2026Updated 5 months ago
- Use interactive notebook to break down MiniMind code and learn from scratch.☆159Jan 7, 2026Updated 8 months ago
- 一个生产级的深度研究 Agent 系统,从零构建多智能体编排、Red-Blue 对抗降噪、 语义级上下文压缩、跨 Agent 共享记忆四大核心能力,配套 165 次独立实验 + Bootstrap 统计显著性检验的完整评测体系。☆153May 11, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题☆15,153Jun 14, 2026Updated 3 months ago
- ☆109Jul 24, 2025Updated last year
- ☆56Nov 22, 2025Updated 10 months ago
- 🔍大模型应用开发实战一:RAG 技术全栈指南,在线阅读地址:https://datawhalechina.github.io/all-in-rag/☆11,293Sep 4, 2026Updated 2 weeks ago
- Personal Project: MPP-Qwen14B & MPP-Qwen-Next(Multimodal Pipeline Parallel based on Qwen-LM). Support [video/image/multi-image] {sft/conv…☆691Mar 10, 2025Updated last year
- This project aims to replicate mainstream open-source model architectures with limited computational resources, implementing mini models …☆292Aug 24, 2026Updated last month
- 基于 Qwen/Qwen3-0.6B 的医疗问答微调与推理示例项目,记录健康数据、生成建议与报告☆16Sep 30, 2025Updated 11 months ago
- [ICLR2026] The first W4A4KV4 quantized + 50% sparse LLMs!☆35Jan 26, 2026Updated 7 months ago
- [EMNLP 2026] Official implementation for paper "Demystifying Reinforcement Learning for Long-Horizon Tool-Using Agents: A Comprehensive R…☆53Aug 23, 2026Updated last month
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- 记录我在cs336学习时的笔记和作业☆1,150May 2, 2026Updated 4 months ago
- CS336 面试导向学习指南 - Stanford Language Modeling from Scratch☆239Apr 2, 2026Updated 5 months ago
- The Quadratic Primal-Dual Optimizer☆14May 13, 2024Updated 2 years ago
- 📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程☆80,633Updated this week
- 2025年腾讯广告算法大赛,“调教大师”队伍初赛方案。初赛排名150+☆16Mar 16, 2026Updated 6 months ago
- Research Agent for multi-hop reasoning QA — 3rd place🥉 in 2026 Alibaba Cloud Data+AI Global Competition☆32Mar 16, 2026Updated 6 months ago
- Multi-Modal-AI-Orchestrator (Reset version),AI Full-modal Full-agent:Text → Image → Music → Lights → Video, Includes "Scenario Director,…☆102Nov 5, 2025Updated 10 months ago
- 《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程☆32,301Sep 12, 2026Updated last week
- Fixing GRPO training collapse in long-horizon multi-tool agents. A lightweight PRM-Lite + LATA joint approach achieves +37% over vanilla …☆242Jun 27, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Bilibili东川路第一可爱猫猫虫的AI笔记☆310May 2, 2026Updated 4 months ago
- 2025腾讯广告算法大赛-分数0.0978代码☆103Sep 17, 2025Updated last year
- 一份面向实践者的 verl 框架使用教程。verl 是字节跳动开源的大语言模型强化学习训练框架,支持 PPO、GRPO 等多种算法,以及分布式训练、AgentRL 等场景。☆143Jul 2, 2026Updated 2 months ago
- 企业级多Agent知识管理系统:4个AI Agent协作完成文档解析→知识抽取→智能问答→增量更新,Python/Java/Go三语言实现☆150Apr 5, 2026Updated 5 months ago
- [CVPR25 Highlight] A ChatGPT-Prompted Visual hallucination Evaluation Dataset, featuring over 100,000 data samples and four advanced eval…☆32Apr 16, 2025Updated last year
- 将SmolVLM2的视觉头与Qwen3-0.6B模型进行了拼接微调☆612Sep 8, 2025Updated last year
- Minimal reproduction of OneRec☆1,839Sep 7, 2026Updated 2 weeks ago