This is a list of peer-reviewed representative papers on deep learning dynamics (optimization dynamics of neural networks). The success of deep learning attributes to both network architecture and stochastic optimization. Thus, deep learning dynamics play an essentially important role in theoretical foundation of deep learning.
☆305Apr 10, 2024Updated 2 years ago
Alternatives and similar repositories for deep-learning-dynamics-paper-list
Users that are interested in deep-learning-dynamics-paper-list are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2022, Oral] The PyTorch Implementation of Adaptive Inertia Methods. The algorithms are based on our paper: "Adaptive Inertia: Dise…☆150Feb 17, 2023Updated 3 years ago
- [Neural Computation, MIT Press] The PyTorch Implementation of Variable Optimizers/ Neural Variable Risk Minimization proposed in our Neur…☆33Aug 3, 2021Updated 5 years ago
- [NeurIPS 2023] The PyTorch Implementation of Scheduled (Stable) Weight Decay.☆61Feb 3, 2024Updated 2 years ago
- Welcome to the Awesome Feature Learning in Deep Learning Thoery Reading Group! This repository serves as a collaborative platform for sch…☆210Apr 13, 2026Updated 4 months ago
- Neural Tangent Kernel Papers☆122Jan 12, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Welcome to the 'In Context Learning Theory' Reading Group☆31Nov 8, 2024Updated last year
- This framework implements key experiments on the sparse double descent phenomenon (ICML 2022).☆15Dec 13, 2022Updated 3 years ago
- PyHessian is a Pytorch library for second-order based analysis and training of Neural Networks☆795Jul 10, 2025Updated last year
- ☆10Dec 17, 2019Updated 6 years ago
- ☆14Oct 18, 2021Updated 4 years ago
- An implementation of the penalty-based bilevel gradient descent (PBGD) algorithm and the iterative differentiation (ITD/RHG) methods.☆19Feb 13, 2023Updated 3 years ago
- A curated list of papers of interesting empirical study and insight on deep learning. Continually updating...☆409Aug 18, 2026Updated 3 weeks ago
- ☆29Jun 12, 2025Updated last year
- ☆13Jul 2, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Visualization of mean field and neural tangent kernel regime☆23Jul 25, 2024Updated 2 years ago
- ScalingOpt - Optimization Community☆107Jun 1, 2026Updated 3 months ago
- Efficient empirical NTKs in PyTorch☆22Jun 13, 2022Updated 4 years ago
- Official implementation for the paper "Controlled Sparsity via Constrained Optimization"☆12Aug 10, 2022Updated 4 years ago
- [NeurIPS 2023 Spotlight] Temperature Balancing, Layer-wise Weight Analysis, and Neural Network Training☆37Apr 7, 2025Updated last year
- Quantification of Uncertainties in Neural Networks☆11Aug 29, 2026Updated last week
- ☆95Jul 18, 2023Updated 3 years ago
- Awesome papers in machine learning theory☆10Feb 12, 2022Updated 4 years ago
- Repository for the paper "Interpreting Temporal Graph Neural Networks with Koopman Theory"☆12Apr 7, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Experiments on trade-off among optimization, generalization and conflict aversion in multi-objective learning (MOL), and introducing MoDo…☆15Oct 21, 2023Updated 2 years ago
- [NeurIPS 2021] code for "Taxonomizing local versus global structure in neural network loss landscapes" https://arxiv.org/abs/2107.11228☆20Jan 7, 2022Updated 4 years ago
- Implementation of Beyond Neural Scaling beating power laws for deep models and prototype-based models☆35Oct 30, 2025Updated 10 months ago
- Code for experiments in my blog post on the Neural Tangent Kernel: https://eigentales.com/NTK☆174Nov 10, 2019Updated 6 years ago
- ☆27Feb 2, 2023Updated 3 years ago
- ☆23Nov 1, 2022Updated 3 years ago
- Deep Learning Theory and Practice☆25Dec 5, 2023Updated 2 years ago
- ☆36Jun 13, 2023Updated 3 years ago
- Code for Paper (Policy Optimization in RLHF: The Impact of Out-of-preference Data)☆29Dec 19, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers☆44Jul 1, 2026Updated 2 months ago
- SE-PINN: Solving the Schrödinger Equation via Physics-Informed Machine Learning☆11Dec 17, 2025Updated 8 months ago
- Finetune Google's pre-trained ViT models from HuggingFace's model hub.☆19Apr 4, 2021Updated 5 years ago
- Library for computing the Finite-time Lyapunov Exponents of 2D flows using xarray☆10Apr 25, 2022Updated 4 years ago
- Sort out the researchers in the field of AI for Science☆21Apr 5, 2023Updated 3 years ago
- ☆15Mar 4, 2022Updated 4 years ago
- Code for 'Periodic Activation Functions Induce Stationarity' (NeurIPS 2021)☆21Oct 27, 2021Updated 4 years ago