A repository consisting of paper/architecture replications of classic/SOTA AI/ML papers in pytorch
☆428Nov 11, 2025Updated 10 months ago
Alternatives and similar repositories for Paper-Replications
Users that are interested in Paper-Replications are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repository of implementations of classic and sota rl algorithms from scratch in PyTorch☆229Aug 3, 2026Updated last month
- An educational distributed training and inference library for neural nets using local computing☆85Jun 10, 2026Updated 3 months ago
- learningggggggg 🐳☆639Apr 2, 2025Updated last year
- just me trying to implement deep learning concepts in code☆236Nov 8, 2025Updated 10 months ago
- ☆45May 4, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- small auto-grad engine inspired from Karpathy's micrograd and PyTorch☆276Nov 21, 2024Updated last year
- Implementations of Papers that I read, you can read my breakdown in my blog☆90Oct 23, 2025Updated 11 months ago
- PyTorch implementations of algorithms from "Reinforcement Learning: An Introduction by Sutton and Barto", along with various RL research …☆207Aug 14, 2025Updated last year
- Optimizing diffusion for production-ready speeds☆40Jan 10, 2026Updated 8 months ago
- This series will take you on a journey from the fundamentals of NLP and Computer Vision to the cutting edge of Vision-Language Models.☆1,181Jan 23, 2025Updated last year
- ☆45Mar 31, 2025Updated last year
- Course on Flash-attention in Triton☆104Feb 9, 2026Updated 7 months ago
- NanoGPT-speedrunning for the poor T4 enjoyers☆72Apr 22, 2025Updated last year
- Basically a repo containing architectures/algorithms/papers from scratch in pytorch☆30Feb 11, 2026Updated 7 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- a simple c++ inference engine for gpt based architecture☆41Dec 10, 2025Updated 9 months ago
- a minimal paged attention implementation☆21Jan 30, 2026Updated 7 months ago
- KernelBench v2: Can LLMs Write GPU Kernels? - Benchmark with Torch -> Triton (and more!) problems☆26Jul 4, 2025Updated last year
- ☆446Apr 10, 2025Updated last year
- working implimention of deepseek MLA☆44Jan 8, 2025Updated last year
- Composition of Multimodal Language Models From Scratch☆16Aug 16, 2024Updated 2 years ago
- rl from zero pretrain, can it be done? yes.☆296Sep 28, 2025Updated last year
- This Repo consists the python note books of IITM - Mathematical Foundations for Generative AI Course,☆412Feb 17, 2026Updated 7 months ago
- GPU Kernels☆229Apr 27, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆14Jan 22, 2026Updated 8 months ago
- ☆658Aug 28, 2025Updated last year
- ☆87Jan 24, 2026Updated 8 months ago
- Row-wise block scaling for fp8 quantization matrix multiplication. Solution to GPU mode AMD challenge.☆19Feb 9, 2026Updated 7 months ago
- Flash Attention from scratch, tiled CUDA forward kernel, online softmax with running max and correction factor, recomputation trick in ba…☆19Mar 6, 2026Updated 6 months ago
- KV Cache & LoRA for minGPT☆61Mar 4, 2026Updated 6 months ago
- My submission for the GPUMODE/AMD fp8 mm challenge☆29Jun 4, 2025Updated last year
- Clean, reusable paper implementations for trending papers on alphaXiv☆212Jul 21, 2026Updated 2 months ago
- This repository contains an exhaustive coverage of a hands on approach to PyTorch along side powerful tools to accelerate model tuning an…☆308May 16, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.☆108Jul 19, 2025Updated last year
- Minimal and annotated implementations of key ideas from modern deep learning research.☆1,356Jan 29, 2026Updated 7 months ago
- building a Large Language Model (LLM) from scratch.☆36Feb 4, 2025Updated last year
- ML from scratch☆2,448Aug 12, 2025Updated last year
- ☆47May 24, 2025Updated last year
- ☆17May 6, 2025Updated last year
- This repository contains a curated collection of 300+ case studies from over 80 companies, detailing practical applications and insights …☆11,008Aug 5, 2025Updated last year