This is a list of peer-reviewed representative papers on deep learning dynamics (optimization dynamics of neural networks). The success of deep learning attributes to both network architecture and stochastic optimization. Thus, deep learning dynamics play an essentially important role in theoretical foundation of deep learning.
☆306Apr 10, 2024Updated 2 years ago
Alternatives and similar repositories for deep-learning-dynamics-paper-list
Users that are interested in deep-learning-dynamics-paper-list are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2022, Oral] The PyTorch Implementation of Adaptive Inertia Methods. The algorithms are based on our paper: "Adaptive Inertia: Dise…☆150Feb 17, 2023Updated 3 years ago
- [ICML 2021] The official PyTorch Implementations of Positive-Negative Momentum Optimizers.☆28Aug 30, 2022Updated 3 years ago
- [Neural Computation, MIT Press] The PyTorch Implementation of Variable Optimizers/ Neural Variable Risk Minimization proposed in our Neur…☆33Aug 3, 2021Updated 5 years ago
- Welcome to the Awesome Feature Learning in Deep Learning Thoery Reading Group! This repository serves as a collaborative platform for sch…☆212Apr 13, 2026Updated 4 months ago
- Neural Tangent Kernel Papers☆123Jan 12, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Welcome to the 'In Context Learning Theory' Reading Group☆31Nov 8, 2024Updated last year
- This framework implements key experiments on the sparse double descent phenomenon (ICML 2022).☆15Dec 13, 2022Updated 3 years ago
- ☆21Jan 4, 2023Updated 3 years ago
- PyHessian is a Pytorch library for second-order based analysis and training of Neural Networks☆793Jul 10, 2025Updated last year
- Codebase for the paper "A Gradient Flow Framework for Analyzing Network Pruning"☆20Jan 31, 2021Updated 5 years ago
- Code for the paper: Why Transformers Need Adam: A Hessian Perspective☆65Mar 11, 2025Updated last year
- ☆14Oct 18, 2021Updated 4 years ago
- An implementation of the penalty-based bilevel gradient descent (PBGD) algorithm and the iterative differentiation (ITD/RHG) methods.☆19Feb 13, 2023Updated 3 years ago
- ☆15May 2, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A curated list of papers of interesting empirical study and insight on deep learning. Continually updating...☆403Jul 21, 2026Updated 3 weeks ago
- ☆29Jun 12, 2025Updated last year
- ☆13Jul 2, 2025Updated last year
- Code to implement Restormer-Plus, the Runner-up Solution to the GT-RAIN Challenge (CVPR 2023 UG2+ Track 3)☆15Oct 11, 2024Updated last year
- A curated list of awesome papers on dataset reduction, including dataset distillation (dataset condensation) and dataset pruning (coreset…☆61Jan 14, 2025Updated last year
- finding new ramsey bounds through scaling autoresearch☆48May 13, 2026Updated 3 months ago
- Visualization of mean field and neural tangent kernel regime☆23Jul 25, 2024Updated 2 years ago
- Efficient empirical NTKs in PyTorch☆22Jun 13, 2022Updated 4 years ago
- Official implementation for the paper "Controlled Sparsity via Constrained Optimization"☆12Aug 10, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NeurIPS 2023 Spotlight] Temperature Balancing, Layer-wise Weight Analysis, and Neural Network Training☆37Apr 7, 2025Updated last year
- ☆35Dec 5, 2022Updated 3 years ago
- Quantification of Uncertainties in Neural Networks☆11Feb 25, 2026Updated 5 months ago
- ☆94Jul 18, 2023Updated 3 years ago
- Awesome papers in machine learning theory☆10Feb 12, 2022Updated 4 years ago
- Repository for the paper "Interpreting Temporal Graph Neural Networks with Koopman Theory"☆12Apr 7, 2026Updated 4 months ago
- Experiments on trade-off among optimization, generalization and conflict aversion in multi-objective learning (MOL), and introducing MoDo…☆15Oct 21, 2023Updated 2 years ago
- Implementation of Beyond Neural Scaling beating power laws for deep models and prototype-based models☆35Oct 30, 2025Updated 9 months ago
- Code for experiments in my blog post on the Neural Tangent Kernel: https://eigentales.com/NTK☆174Nov 10, 2019Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆27Feb 2, 2023Updated 3 years ago
- ☆23Nov 1, 2022Updated 3 years ago
- ☆36Jun 13, 2023Updated 3 years ago
- Code for Paper (Policy Optimization in RLHF: The Impact of Out-of-preference Data)☆29Dec 19, 2023Updated 2 years ago
- Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers☆44Jul 1, 2026Updated last month
- Finetune Google's pre-trained ViT models from HuggingFace's model hub.☆19Apr 4, 2021Updated 5 years ago
- Library for computing the Finite-time Lyapunov Exponents of 2D flows using xarray☆10Apr 25, 2022Updated 4 years ago