Bilibili东川路第一可爱猫猫虫的AI笔记
☆301May 2, 2026Updated 4 months ago
Alternatives and similar repositories for AI-Notes
Users that are interested in AI-Notes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 🎓从0开始训练一个大模型Minimind项目的超详细解析,包括但不限于用到的架构,算法,以及大模型面试经验☆1,683May 25, 2026Updated 3 months ago
- solution for cs336-assignment1,2,5 , including colab code link and blog link.☆16Feb 20, 2026Updated 6 months ago
- ☆55Mar 26, 2026Updated 5 months ago
- ☆831Jul 26, 2026Updated last month
- PyTorch implementation for Cross-modal Retrieval with Noisy Correspondence via Consistency Refining and Mining (TIP 2024)☆22Mar 25, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pretrain、Posttrain、RAG、Agent等大模型相关的基础项目合集☆39Dec 7, 2025Updated 8 months ago
- The 'missing header' for Chisel☆24Feb 5, 2026Updated 6 months ago
- 🧠 Train a 64M-parameter LLM from scratch in just 2h!☆58,236Updated this week
- A user-friendly open-source project for recommendation systems.☆58Apr 26, 2026Updated 4 months ago
- [AAAI'26 - Oral] Realism Control One-step Diffusion for Real-World Image Super-Resolution☆16Nov 15, 2025Updated 9 months ago
- OpenMOSS presents a collection of our research on LLMs, supported by SII, Fudan and Mosi.☆31Updated this week
- 🎯 Build a winning recommendation system with this effective generative framework, advancing to the finals of the 2025 Tencent Advertisin…☆27Updated this week
- 本项目存放RankMixer复现相关代码☆73Apr 3, 2026Updated 5 months ago
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,363Nov 13, 2025Updated 9 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training Pipeline. 训练医疗大模型,实现了包括增量预训练(PT)、有监督微调(SFT)、RLHF、DPO、ORPO、GRPO。☆5,775Jun 3, 2026Updated 3 months ago
- An open source SDR SDRAM controller based on the AXI4 bus and verified by FPGA and tapeout. It can support memory particles of different …☆22May 12, 2025Updated last year
- ☆335Jul 22, 2026Updated last month
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,258Updated this week
- Pytorch implementation of the paper 'Gaussian Mixture Proposals with Pull-Push Learning Scheme to Capture Diverse Events for Weakly Super…☆19Jan 19, 2024Updated 2 years ago
- 包含了LLM的一些手撕代码,如强化学习。可以帮助从代码层面深入理解原理,以及有助于准备大模型面试可能出现的手撕。后续会更新Transformer等更多手撕☆129Mar 15, 2026Updated 5 months ago
- Trust Region Preference Approximation: A simple and stable reinforcement learning algorithm for LLM reasoning☆14Jun 28, 2025Updated last year
- Implementation for NeurIPS 2024 oral paper: Divide-and-Conquer Meets Consensus: Unleashing the Power of Functions in Code Generation☆16Jan 27, 2025Updated last year
- My implementation of Stanford CS336 assignments.☆246Mar 15, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Based on Nano-vLLM, a simple replication of vLLM with self-contained paged attention and flash attention implementation☆1,011Updated this week
- A cross-platform UI automation framework based on Airtest and StarRailCopilot☆16Jan 29, 2026Updated 7 months ago
- A Flexible Framework for Generative Recommendation☆49Apr 9, 2026Updated 4 months ago
- 一份面向实践者的 verl 框架使用教程。verl 是字节跳动开源的大语言模型强化学习训练框架,支持 PPO、GRPO 等多种算法,以及分布式训练、AgentRL 等场景。☆132Jul 2, 2026Updated 2 months ago
- 三元三小时手敲大模型☆590Mar 12, 2026Updated 5 months ago
- 中文翻译的 Hands-On-Large-Language-Models (hands-on-llms),动手学习大模型☆2,868Jul 19, 2026Updated last month
- Few-Shot Preference Optimization (FSPO) personalizes LLMs by reframing reward modeling as a meta-learning problem, enabling rapid adaptat…☆17Feb 27, 2025Updated last year
- Local-first interview recording review reports with a Codex skill and CLI.☆80May 16, 2026Updated 3 months ago
- Make math learning simpler, starting with Nano Math plus , too!☆45Jul 16, 2026Updated last month
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [AAAI 2025] Grounded Multi-Hop VideoQA in Long-Form Egocentric Videos☆39May 27, 2025Updated last year
- 从零预训练LLM、SFT、RLHF、DPO笔记整理+面试问题☆21Sep 2, 2024Updated 2 years ago
- 基于天池新闻推荐赛数据集实现的新闻推荐☆53Dec 17, 2024Updated last year
- ☆21Jan 28, 2026Updated 7 months ago
- This is the official baseline code for TAAC2025(parquet format).☆115Apr 24, 2026Updated 4 months ago
- 🚀 2026届大模型算法岗实习面经 | 包含 DeepSeek/Qwen 技术报告解析、手撕 PPO/RoPE/Transformer、RLHF 核心与八股文 | 持续更新中...☆665Mar 28, 2026Updated 5 months ago
- Reinforcing Long-Term Performance in Recommender Systems with User-Oriented Exploration Policy (SIGIR 2024)☆14Oct 6, 2024Updated last year