A curated list of papers of interesting empirical study and insight on deep learning. Continually updating...
☆403Jul 21, 2026Updated 3 weeks ago
Alternatives and similar repositories for awesome-deep-phenomena
Users that are interested in awesome-deep-phenomena are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Welcome to the Awesome Feature Learning in Deep Learning Thoery Reading Group! This repository serves as a collaborative platform for sch…☆212Apr 13, 2026Updated 4 months ago
- ☆78Dec 7, 2024Updated last year
- ☆26Feb 20, 2026Updated 5 months ago
- Welcome to the 'In Context Learning Theory' Reading Group☆31Nov 8, 2024Updated last year
- [NeurIPS 2021] A Geometric Analysis of Neural Collapse with Unconstrained Features☆62Jul 19, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Pytorch routines for (Ker)nel (Mac)hines☆11Oct 10, 2025Updated 10 months ago
- ☆13Mar 29, 2024Updated 2 years ago
- Official repo for Detecting, Explaining, and Mitigating Memorization in Diffusion Models (ICLR 2024)☆80Apr 3, 2024Updated 2 years ago
- Neural Tangent Kernel Papers☆123Jan 12, 2025Updated last year
- A curated list of awesome Deep Learning theories that shed light on the mysteries of DL☆11Jul 20, 2018Updated 8 years ago
- NeurIPS'24 - LLM Safety Landscape☆40Oct 21, 2025Updated 9 months ago
- awesome papers in LLM interpretability☆626Aug 20, 2025Updated 11 months ago
- [NeurIPS 2023] Code release for "Going Beyond Linear Mode Connectivity: The Layerwise Linear Feature Connectivity"☆19Oct 19, 2023Updated 2 years ago
- [NeurIPS 2022] The official code for our NeurIPS 2022 paper "Inducing Neural Collapse in Imbalanced Learning: Do We Really Need a Learnab…☆46Oct 12, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆278Dec 2, 2024Updated last year
- This is a curated list for Information Bottleneck Principle, in memory of Professor Naftali Tishby.☆401Feb 12, 2026Updated 6 months ago
- A curated list of LLM Interpretability related material - Tutorial, Library, Survey, Paper, Blog, etc..☆308Jan 22, 2026Updated 6 months ago
- Code accompanying our paper "Feature Learning in Infinite-Width Neural Networks" (https://arxiv.org/abs/2011.14522)☆61May 11, 2021Updated 5 years ago
- Physics of Language Models: Part 4.2, Canon Layers at Scale where Synthetic Pretraining Resonates in Reality☆357May 20, 2026Updated 2 months ago
- This is a list of peer-reviewed representative papers on deep learning dynamics (optimization dynamics of neural networks). The success o…☆306Apr 10, 2024Updated 2 years ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- ☆23Jan 25, 2024Updated 2 years ago
- [IJCAI'22 Survey] Recent Advances on Neural Network Pruning at Initialization.☆59Oct 10, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The is the official implementation of ICCV 2023 paper "No Fear of Classifier Biases: Neural Collapse Inspired Federated Learning with Syn…☆30Oct 27, 2023Updated 2 years ago
- A curated list of awesome papers on dataset distillation and related applications.☆1,967Updated this week
- ☆17Feb 4, 2025Updated last year
- ☆34Jan 25, 2024Updated 2 years ago
- SGD with large step sizes learns sparse features [ICML 2023]☆34Apr 24, 2023Updated 3 years ago
- The loss landscape of Large Language Models resemble basin!☆44Jul 8, 2025Updated last year
- Implementation of "RaanA: A Fast, Flexible, and Data-Efficient Post-Training Quantization Algorithm"☆18Apr 11, 2025Updated last year
- Source code of "Task arithmetic in the tangent space: Improved editing of pre-trained models".☆112Jun 8, 2023Updated 3 years ago
- [ICML 2025] Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction☆91May 26, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆16May 10, 2024Updated 2 years ago
- [NeurIPS-2023] The PyTorch Implementation of MoSo. The algorithms are based on our paper: "Data Pruning via Moving-one-Sample-out". MoSo …☆10May 21, 2026Updated 2 months ago
- ☆246May 10, 2024Updated 2 years ago
- A fusion of a linear layer and a cross entropy loss, written for pytorch in triton.☆75Aug 2, 2024Updated 2 years ago
- The official implementation of “Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training”☆1,003Jan 30, 2024Updated 2 years ago
- Muon is an optimizer for hidden layers in neural networks☆2,787May 24, 2026Updated 2 months ago
- [ICLR 2023] The official code for our ICLR 2023 (top25%) paper: "Neural Collapse Inspired Feature-Classifier Alignment for Few-Shot Class…☆94Aug 7, 2023Updated 3 years ago