Build LLM from scratch
☆133Sep 25, 2026Updated 2 weeks ago
Alternatives and similar repositories for llm-from-scratch
Users that are interested in llm-from-scratch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tiny-R2: A hybrid architecture integrating SWA, CSA, HCA, mHC, and DSMoE under the DeepSeek V4 design paradigm, enabling single-GPU OPD p…☆50May 30, 2026Updated 4 months ago
- Nex Venus Communication Library☆77Nov 17, 2025Updated 10 months ago
- ☆16Mar 24, 2026Updated 6 months ago
- A simple but instructive implementation of DP, TP, FSDP, FSDP+TP using pytorch distributed primitives☆24Apr 12, 2026Updated 5 months ago
- Make math learning simpler, starting with Nano Math plus , too!☆46Jul 16, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- GEMV implementation with CUTLASS☆21Aug 21, 2025Updated last year
- 从零构建大模型:从预训练到RLHF的完整实践☆2,701Aug 21, 2026Updated last month
- DeepSeek-V4 Lecture☆31Aug 10, 2026Updated last month
- Metrics for evaluating biological sequence design☆16Sep 1, 2026Updated last month
- ☆32Jul 2, 2025Updated last year
- KFunca: A minimalist, high-performance GPU-based automatic differentiation framework☆31Aug 14, 2025Updated last year
- 基于RISC-V的仿Linux操作系统内核☆13Aug 21, 2022Updated 4 years ago
- alphafold FAPE loss☆10Sep 28, 2021Updated 5 years ago
- ☆26May 27, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A NCCL extension library, designed to efficiently offload GPU memory allocated by the NCCL communication library.☆119Dec 17, 2025Updated 9 months ago
- ☆11Apr 17, 2025Updated last year
- Benchmark tests supporting the TiledCUDA library.☆19Nov 19, 2024Updated last year
- GPU Functional Descriptor for memory access☆35May 24, 2026Updated 4 months ago
- ☆18Jan 16, 2026Updated 8 months ago
- Distributed MoE in a Single Kernel [NeurIPS '25]☆300May 5, 2026Updated 5 months ago
- ☆57Jul 12, 2025Updated last year
- [NeurIPS 2025] ClusterFusion: Expanding Operator Fusion Scope for LLM Inference via Cluster-Level Collective Primitive☆79Aug 8, 2026Updated 2 months ago
- ☆17Jan 31, 2026Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆37Aug 7, 2025Updated last year
- TensorRT for RefineNet Segmentation☆12Apr 27, 2021Updated 5 years ago
- ☆26Jul 7, 2024Updated 2 years ago
- ☆14Apr 7, 2022Updated 4 years ago
- Dataset and code for “Multi-Interactive Attention Network for Fine-grained Feature Learning in CTR Prediction”☆16Aug 23, 2021Updated 5 years ago
- Virtual Decoupled Cores: Composable Programming Framework and Runtime for Async GPUs☆23Sep 29, 2026Updated last week
- Simple intermediate representation language for learning and research.☆22Mar 27, 2020Updated 6 years ago
- Cute layout visualization☆45Jan 18, 2026Updated 8 months ago
- Compose computational workflows from Python☆11Oct 9, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- let coding agents use ncu skills analysis cuda program automatically!☆129May 25, 2026Updated 4 months ago
- ☆98Mar 26, 2025Updated last year
- Codes for our paper "AgentMonitor: A Plug-and-Play Framework for Predictive and Secure Multi-Agent Systems"☆14Dec 13, 2024Updated last year
- This is a mirror repository of gpu4pyscf. Please refer to https://github.com/pyscf/gpu4pyscf for the latest development.☆17Oct 2, 2026Updated last week
- Pytorch refinenet for segmentation and pytorch -> onnx -> tensorrt☆10Aug 25, 2020Updated 6 years ago
- Optimize GEMM with tensorcore step by step☆40Dec 17, 2023Updated 2 years ago
- ☆29Jan 5, 2025Updated last year