A curated list of papers of interesting empirical study and insight on deep learning. Continually updating...
☆410Aug 18, 2026Updated last month
Alternatives and similar repositories for awesome-deep-phenomena
Users that are interested in awesome-deep-phenomena are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆28Feb 20, 2026Updated 7 months ago
- A curated list of trustworthy deep learning papers. Continually updating...☆392Aug 11, 2026Updated last month
- Welcome to the Awesome Feature Learning in Deep Learning Thoery Reading Group! This repository serves as a collaborative platform for sch…☆210Apr 13, 2026Updated 5 months ago
- [ICLR 2025 Spotlight] Code release for "Sharpness-Aware Minimization Efficiently Selects Flatter Minima Late In Training"☆20Feb 20, 2025Updated last year
- ☆78Dec 7, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- This paper list focuses on the theoretical and empirical analysis of language models, especially large language models (LLMs). The papers…☆103Jul 29, 2026Updated last month
- Welcome to the 'In Context Learning Theory' Reading Group☆31Nov 8, 2024Updated last year
- ☆15Sep 29, 2022Updated 3 years ago
- [NeurIPS 2021] A Geometric Analysis of Neural Collapse with Unconstrained Features☆60Jul 19, 2022Updated 4 years ago
- ☆13Mar 29, 2024Updated 2 years ago
- Official repo for Detecting, Explaining, and Mitigating Memorization in Diffusion Models (ICLR 2024)☆80Apr 3, 2024Updated 2 years ago
- Neural Tangent Kernel Papers☆122Jan 12, 2025Updated last year
- A curated list of awesome Deep Learning theories that shed light on the mysteries of DL☆12Jul 20, 2018Updated 8 years ago
- NeurIPS'24 - LLM Safety Landscape☆41Oct 21, 2025Updated 11 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- awesome papers in LLM interpretability☆628Aug 20, 2025Updated last year
- [NeurIPS 2022] The official code for our NeurIPS 2022 paper "Inducing Neural Collapse in Imbalanced Learning: Do We Really Need a Learnab…☆46Oct 12, 2022Updated 3 years ago
- This is a curated list for Information Bottleneck Principle, in memory of Professor Naftali Tishby.☆402Feb 12, 2026Updated 7 months ago
- A curated list of LLM Interpretability related material - Tutorial, Library, Survey, Paper, Blog, etc..☆306Jan 22, 2026Updated 8 months ago
- Code accompanying our paper "Feature Learning in Infinite-Width Neural Networks" (https://arxiv.org/abs/2011.14522)☆61May 11, 2021Updated 5 years ago
- Physics of Language Models: Part 4.2, Canon Layers at Scale where Synthetic Pretraining Resonates in Reality☆361Aug 21, 2026Updated last month
- This is a list of peer-reviewed representative papers on deep learning dynamics (optimization dynamics of neural networks). The success o…☆305Apr 10, 2024Updated 2 years ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- ☆23Jan 25, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The is the official implementation of ICCV 2023 paper "No Fear of Classifier Biases: Neural Collapse Inspired Federated Learning with Syn…☆30Oct 27, 2023Updated 2 years ago
- A curated list of awesome papers on dataset distillation and related applications.☆1,974Sep 3, 2026Updated 3 weeks ago
- ☆17Feb 4, 2025Updated last year
- ☆34Jan 25, 2024Updated 2 years ago
- SGD with large step sizes learns sparse features [ICML 2023]☆35Apr 24, 2023Updated 3 years ago
- The loss landscape of Large Language Models resemble basin!☆45Jul 8, 2025Updated last year
- [NeurIPS 2025, Spotlight] An official implementation of the paper Quantization-Free Autoregressive Action Transformer☆12Mar 3, 2026Updated 6 months ago
- Implementation of "RaanA: A Fast, Flexible, and Data-Efficient Post-Training Quantization Algorithm"☆18Apr 11, 2025Updated last year
- Source code of "Task arithmetic in the tangent space: Improved editing of pre-trained models".☆114Jun 8, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICML 2025] Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction☆91May 26, 2025Updated last year
- [NeurIPS-2023] The PyTorch Implementation of MoSo. The algorithms are based on our paper: "Data Pruning via Moving-one-Sample-out". MoSo …☆10May 21, 2026Updated 4 months ago
- (ICML 2023) Feature learning in deep classifiers through Intermediate Neural Collapse: Accompanying code☆16Jul 27, 2023Updated 3 years ago
- ☆247May 10, 2024Updated 2 years ago
- A fusion of a linear layer and a cross entropy loss, written for pytorch in triton.☆77Aug 2, 2024Updated 2 years ago
- The official implementation of “Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training”☆1,000Jan 30, 2024Updated 2 years ago
- Muon is an optimizer for hidden layers in neural networks☆2,858May 24, 2026Updated 4 months ago