Student version of Assignment 2 for Stanford CS336 - Language Modeling From Scratch
☆282May 1, 2026Updated 2 months ago
Alternatives and similar repositories for assignment2-systems
Users that are interested in assignment2-systems are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆84Jun 9, 2026Updated last month
- Student version of Assignment 1 for Stanford CS336 - Language Modeling From Scratch☆2,454Apr 7, 2026Updated 3 months ago
- ☆3,530May 28, 2026Updated last month
- My implementation of Stanford CS336 assignments.☆246Mar 15, 2026Updated 4 months ago
- 记录我在cs336学习时的笔记和作业☆1,018May 2, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Assignment 1 for Stanford CS336 - Language Modeling From Scratch☆77Jul 7, 2025Updated last year
- ☆18Nov 22, 2025Updated 8 months ago
- Official Repository of Paper "Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs"☆15Sep 25, 2025Updated 10 months ago
- ☆143Jan 18, 2026Updated 6 months ago
- CS33作业 2 的代码和飞书 qa, 这个作业太恶心了, 绝对是所有作业里面花的最久的☆24Jul 17, 2025Updated last year
- Implementation of my CS336 assignment1☆47Dec 23, 2025Updated 7 months ago
- ☆15Sep 29, 2022Updated 3 years ago
- Implementation of Stanford CS336 (Spring 2025) assignments☆35Mar 8, 2026Updated 4 months ago
- (包含完整代码和坑点记录)Student version of Assignment 1 for Stanford CS336 - Language Modeling From Scratch☆41Jan 22, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆101May 4, 2026Updated 2 months ago
- My Solution and Notes for the Stanford CS336: LLM from scratch☆264Mar 23, 2026Updated 4 months ago
- ☆45Nov 22, 2025Updated 8 months ago
- Official Code for What Makes and Breaks Safety Fine-tuning? A Mechanistic Study (NeurIPS 2024)☆12Oct 31, 2024Updated last year
- cs336作业 1 实现, 我把 qa 问题也放在飞书链接里面了, 仅供参考☆35Jul 3, 2025Updated last year
- Reproducing and studying RL algorithms for LLM agents, including PPO, GRPO, GSPO, DAPO, OPD and beyond.☆60Updated this week
- ☆36Jul 5, 2023Updated 3 years ago
- Implementing scalable LLMs in pure JAX (no third-party libraries)☆51Jun 11, 2026Updated last month
- A private repo for learning CS336☆38Sep 2, 2025Updated 10 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The Automated LLM Speedrunning Benchmark measures how well LLM agents can reproduce previous innovations and discover new ones in languag…☆145May 6, 2026Updated 2 months ago
- ☆26Feb 20, 2026Updated 5 months ago
- The Full Spectrum of Deepnet Hessians at Scale: Dynamics with SGD Training and Sample Size☆19May 19, 2019Updated 7 years ago
- 2024CCF国际AIOps挑战赛-赛道二(GLM4):基于检索增强的运维知识问答挑战赛解决方案分享。☆14Jul 5, 2024Updated 2 years ago
- Nano vLLM☆14,635Apr 26, 2026Updated 3 months ago
- 6,080-param transformer achieving 100% accuracy on 10-digit addition. Trained from scratch in 10 minutes.☆22Feb 19, 2026Updated 5 months ago
- Stanford "Language Modeling from Scratch" CS336 Assignment1 - 斯坦福大学 CS336 课程作业1 个人实现,仅供参考☆45Jun 15, 2025Updated last year
- 📚LeetCUDA: Modern CUDA Learn Notes with PyTorch for Beginners🐑, 200+ CUDA Kernels, Tensor Cores, HGEMM, FA-2 MMA.🎉☆11,631Updated this week
- Tiny-DeepSpeed, a minimalistic re-implementation of the DeepSpeed library☆53Aug 20, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- TTRV: Test-Time Reinforcement Learning for Vision–Language Models (CVPR 2026)☆46Mar 8, 2026Updated 4 months ago
- The simplest, fastest repository for training/finetuning medium-sized GPTs.☆199Jan 19, 2026Updated 6 months ago
- Intro to using DSPy with Kuzu to enrich the data within the Nobel Laureate mentorship network☆16Sep 16, 2025Updated 10 months ago
- Curse-of-memory phenomenon of RNNs in sequence modelling☆19May 8, 2025Updated last year
- [NAACL 2025 Main Selected Oral] Repository for the paper: Prompt Compression for Large Language Models: A Survey☆36May 18, 2025Updated last year
- Student version of Assignment 1 for Stanford CS336 - Language Modeling From Scratch☆34Apr 27, 2025Updated last year
- Benchmarking Optimizers for LLM Pretraining☆60May 3, 2026Updated 2 months ago