A curated list of papers of interesting empirical study and insight on deep learning. Continually updating...
☆404Jul 21, 2026Updated last week
Alternatives and similar repositories for awesome-deep-phenomena
Users that are interested in awesome-deep-phenomena are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Welcome to the Awesome Feature Learning in Deep Learning Thoery Reading Group! This repository serves as a collaborative platform for sch…☆211Apr 13, 2026Updated 3 months ago
- [ICLR 2025 Spotlight] Code release for "Sharpness-Aware Minimization Efficiently Selects Flatter Minima Late In Training"☆19Feb 20, 2025Updated last year
- ☆78Dec 7, 2024Updated last year
- ☆26Feb 20, 2026Updated 5 months ago
- Welcome to the 'In Context Learning Theory' Reading Group☆31Nov 8, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆15Sep 29, 2022Updated 3 years ago
- [NeurIPS 2021] A Geometric Analysis of Neural Collapse with Unconstrained Features☆62Jul 19, 2022Updated 4 years ago
- ☆13Mar 29, 2024Updated 2 years ago
- Neural Tangent Kernel Papers☆123Jan 12, 2025Updated last year
- A curated list of awesome Deep Learning theories that shed light on the mysteries of DL☆11Jul 20, 2018Updated 8 years ago
- NeurIPS'24 - LLM Safety Landscape☆40Oct 21, 2025Updated 9 months ago
- [NeurIPS 2022] The official code for our NeurIPS 2022 paper "Inducing Neural Collapse in Imbalanced Learning: Do We Really Need a Learnab…☆46Oct 12, 2022Updated 3 years ago
- ☆275Dec 2, 2024Updated last year
- This is a curated list for Information Bottleneck Principle, in memory of Professor Naftali Tishby.☆400Feb 12, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A curated list of LLM Interpretability related material - Tutorial, Library, Survey, Paper, Blog, etc..☆307Jan 22, 2026Updated 6 months ago
- Code accompanying our paper "Feature Learning in Infinite-Width Neural Networks" (https://arxiv.org/abs/2011.14522)☆61May 11, 2021Updated 5 years ago
- Physics of Language Models: Part 4.2, Canon Layers at Scale where Synthetic Pretraining Resonates in Reality☆356May 20, 2026Updated 2 months ago
- This is a list of peer-reviewed representative papers on deep learning dynamics (optimization dynamics of neural networks). The success o…☆304Apr 10, 2024Updated 2 years ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- ☆23Jan 25, 2024Updated 2 years ago
- [IJCAI'22 Survey] Recent Advances on Neural Network Pruning at Initialization.☆59Oct 10, 2023Updated 2 years ago
- Sharpness-Aware Minimization Leads to Low-Rank Features [NeurIPS 2023]☆29Sep 22, 2023Updated 2 years ago
- The is the official implementation of ICCV 2023 paper "No Fear of Classifier Biases: Neural Collapse Inspired Federated Learning with Syn…☆30Oct 27, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A curated list of awesome papers on dataset distillation and related applications.☆1,970Jul 21, 2026Updated last week
- ☆17Feb 4, 2025Updated last year
- ☆34Jan 25, 2024Updated 2 years ago
- The loss landscape of Large Language Models resemble basin!☆41Jul 8, 2025Updated last year
- [NeurIPS 2025, Spotlight] An official implementation of the paper Quantization-Free Autoregressive Action Transformer☆12Mar 3, 2026Updated 4 months ago
- Source code of "Task arithmetic in the tangent space: Improved editing of pre-trained models".☆113Jun 8, 2023Updated 3 years ago
- [ICML 2025] Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction☆89May 26, 2025Updated last year
- [NeurIPS-2023] The PyTorch Implementation of MoSo. The algorithms are based on our paper: "Data Pruning via Moving-one-Sample-out". MoSo …☆10May 21, 2026Updated 2 months ago
- ☆245May 10, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A fusion of a linear layer and a cross entropy loss, written for pytorch in triton.☆75Aug 2, 2024Updated last year
- The official implementation of “Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training”☆1,003Jan 30, 2024Updated 2 years ago
- Muon is an optimizer for hidden layers in neural networks☆2,748May 24, 2026Updated 2 months ago
- ☆20Jun 27, 2026Updated last month
- [ICLR 2023] The official code for our ICLR 2023 (top25%) paper: "Neural Collapse Inspired Feature-Classifier Alignment for Few-Shot Class…☆94Aug 7, 2023Updated 2 years ago
- Visualization of mean field and neural tangent kernel regime☆23Jul 25, 2024Updated 2 years ago
- ☆37Feb 4, 2022Updated 4 years ago