三元三小时手敲大模型
☆576Mar 12, 2026Updated 5 months ago
Alternatives and similar repositories for MokioMind
Users that are interested in MokioMind are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 现代化python入门教程☆49Nov 3, 2025Updated 9 months ago
- 🧠 Train a 64M-parameter LLM from scratch in just 2h!☆54,808Aug 6, 2026Updated last week
- 🎓从0开始训练一个大模型Minimind项目的超详细解析,包括但不限于用到的架构,算法,以及大模型面试经验☆1,624May 25, 2026Updated 2 months ago
- 轻量级大语言模型MiniMind的源码解读,包含tokenizer、RoPE、MoE、KV Cache、pretraining、SFT、LoRA、DPO等完整流程☆1,146Jun 16, 2025Updated last year
- 从0到1学习大模型。☆23Mar 23, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 👀 Train a 65M-parameter VLM from scratch in just 2h!☆8,470Aug 6, 2026Updated last week
- 从零复现 minimind👉minimind-v☆375Dec 24, 2025Updated 7 months ago
- MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training Pipeline. 训练医疗大模型,实现了包括增量预训练(PT)、有监督微调(SFT)、RLHF、DPO、ORPO、GRPO。☆5,728Jun 3, 2026Updated 2 months ago
- ☆107Jul 24, 2025Updated last year
- ☆449Aug 3, 2025Updated last year
- 从无名小卒到大模型(LLM)大英雄~ 欢迎关注后续!!!☆2,275Updated this week
- 🏥 从零基础到面试通关:20节课彻底搞懂MedicalGPT医疗大模型训练全流程 | PT/SFT/LoRA/RLHF/DPO/GRPO | 100+面试高频考点☆210Apr 1, 2026Updated 4 months ago
- ☆800Jul 26, 2026Updated 3 weeks ago
- 🚀 2026届大模型算法岗实习面经 | 包含 DeepSeek/Qwen 技术报告解析、手撕 PPO/RoPE/Transformer、RLHF 核心与八股文 | 持续更新中...☆643Mar 28, 2026Updated 4 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 一些大模型相关的开源网站☆101Mar 18, 2026Updated 5 months ago
- 复现大模型相关算法及一些学习记录☆3,500Jul 2, 2026Updated last month
- ☆33Sep 4, 2025Updated 11 months ago
- 主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题☆14,922Jun 14, 2026Updated 2 months ago
- 📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程☆73,595Updated this week
- 🚀 [从零构建 LLM] 极简大模型训练原理与实践指南。包含 Transformer, Pretraining, SFT 核心代码与对照实验。 | A minimal, principle-first guide to understanding and building…☆174Jun 4, 2026Updated 2 months ago
- 🎙️ A 0.1B Omni model trained from scratch, capable of listening, speaking, and seeing!☆2,336Aug 6, 2026Updated last week
- ☆18Jun 24, 2024Updated 2 years ago
- MiniMind-V 多模态面试学习指南 - 20节课程 + 278道面试题 + STAR面试稿 + 哆啦A梦漫画☆140Apr 2, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 推荐系统八股160问☆366Mar 30, 2025Updated last year
- 🔍大模型应用开发实战一:RAG 技术全栈指南,在线阅读地址:https://datawhalechina.github.io/all-in-rag/☆10,443Jul 29, 2026Updated 3 weeks ago
- - 【LLM面经】大模型实习面试指南。手撕代码、面经经验、思考题等。初学者学习ing......欢迎指正错误☆35Nov 11, 2025Updated 9 months ago
- 🎓 系统性大语言模型构建课程|🛠️ 覆盖预训练数据工程、Tokenizer、Transformer、MoE、GPU 编程 (CUDA/Triton)、分布式训练、Scaling Laws、推理优化及对齐 (SFT/RLHF/GRPO)|🚀 6 个渐进式作业 + 代码驱…☆1,215Updated this week
- skeleton-based action recognition☆19Jan 12, 2022Updated 4 years ago
- ☆146Jan 18, 2026Updated 7 months ago
- Local-first interview recording review reports with a Codex skill and CLI.☆79May 16, 2026Updated 3 months ago
- 《大模型白盒子构建指南》:一个全手搓的Tiny-Universe☆5,014Feb 12, 2026Updated 6 months ago
- Bilibili东川路第一可爱猫猫虫的AI笔记☆291May 2, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 利用简单的代码完成deepseek基于medical-o1-sft数据集的lora微调☆17Feb 25, 2025Updated last year
- mini project for nanorllm☆63Mar 31, 2026Updated 4 months ago
- [AAAI 2026] RSVG-ZeroOV: Exploring a Training-Free Framework for Zero-Shot Open-Vocabulary Visual Grounding in Remote Sensing Images.☆26Nov 11, 2025Updated 9 months ago
- Use interactive notebook to break down MiniMind code and learn from scratch.☆155Jan 7, 2026Updated 7 months ago
- ☆54Mar 26, 2026Updated 4 months ago
- ☆174Jul 26, 2026Updated 3 weeks ago
- Based on Nano-vLLM, a simple replication of vLLM with self-contained paged attention and flash attention implementation☆981Jul 22, 2026Updated 3 weeks ago