Distributed K-FAC preconditioner for PyTorch
☆98Jul 14, 2026Updated last week
Alternatives and similar repositories for kfac-pytorch
Users that are interested in kfac-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pytorch implementation of KFAC and E-KFAC (Natural Gradient).☆134Jul 2, 2019Updated 7 years ago
- A Chainer extension for K-FAC☆20Jun 16, 2019Updated 7 years ago
- Pytorch implementation of KFAC - this is a port of https://github.com/tensorflow/kfac/☆32Jun 6, 2024Updated 2 years ago
- ADAHESSIAN: An Adaptive Second Order Optimizer for Machine Learning☆287Feb 27, 2023Updated 3 years ago
- ☆137Oct 23, 2017Updated 8 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- An implementation of KFAC for TensorFlow☆198Feb 11, 2022Updated 4 years ago
- SKFAC Preconditioner for MindSpore☆12Jul 2, 2021Updated 5 years ago
- {KFAC,EKFAC,Diagonal,Implicit} Fisher Matrices and finite width NTKs in PyTorch☆224Updated this week
- Regularization, Neural Network Training Dynamics☆14Jan 13, 2020Updated 6 years ago
- PyTorch-SSO: Scalable Second-Order methods in PyTorch☆150Oct 1, 2023Updated 2 years ago
- Large-batch Training, Neural Network Optimization☆10Nov 8, 2019Updated 6 years ago
- PyHessian is a Pytorch library for second-order based analysis and training of Neural Networks☆789Jul 10, 2025Updated last year
- ☆31Feb 11, 2021Updated 5 years ago
- Efficient reference implementations of the static & dynamic M-FAC algorithms (for pruning and optimization)☆17Feb 23, 2022Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆14Sep 14, 2021Updated 4 years ago
- Artifact for IPDPS'21: DSXplore: Optimizing Convolutional Neural Networks via Sliding-Channel Convolutions.☆13Apr 6, 2021Updated 5 years ago
- ☆10Apr 29, 2023Updated 3 years ago
- Limitations of the Empirical Fisher Approximation☆48Mar 3, 2025Updated last year
- In this project, we propose to study Vision Transformers trained using the Barlow Twins self-supervised method, and compare the results w…☆17Oct 3, 2023Updated 2 years ago
- ☆28Jul 11, 2021Updated 5 years ago
- The official implementation of “Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training”☆1,003Jan 30, 2024Updated 2 years ago
- ASDL: Automatic Second-order Differentiation Library for PyTorch☆192Dec 5, 2024Updated last year
- Advanced optimizer with Gradient-Centralization☆21Aug 26, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Randomized algorithm class at CU☆17Jul 8, 2025Updated last year
- ☆46Nov 18, 2022Updated 3 years ago
- Computing gradients and Hessians of feed-forward networks with GPU acceleration☆20Feb 14, 2024Updated 2 years ago
- Numerical Experiments☆15Jan 21, 2018Updated 8 years ago
- Look for arXiv papers in a Zotero library and find available DOIs of published versions.☆14Apr 11, 2024Updated 2 years ago
- Collection of algorithms for approximating Fisher Information Matrix for Natural Gradient (and second order method in general)☆142May 26, 2019Updated 7 years ago
- This repository contains the results for the paper: "Descending through a Crowded Valley - Benchmarking Deep Learning Optimizers"☆183Jul 17, 2021Updated 5 years ago
- Implementation of Influence Function approximations for differently sized ML models, using PyTorch☆18Sep 15, 2023Updated 2 years ago
- This package is dedicated to high-order optimization methods. All the methods can be used similarly to standard PyTorch optimizers.☆30Jun 17, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆33Dec 3, 2019Updated 6 years ago
- Pytorch implementation of preconditioned stochastic gradient descent (Kron and affine preconditioner, low-rank approximation precondition…☆198May 30, 2026Updated last month
- A LARS implementation in PyTorch☆353Feb 21, 2020Updated 6 years ago
- ☆13Feb 24, 2020Updated 6 years ago
- Parallel framework for training and fine-tuning deep neural networks☆74Apr 28, 2026Updated 2 months ago
- Advanced data flow management for distributed Python applications☆38Updated this week
- ☆15Feb 12, 2021Updated 5 years ago