Delta Orthogonal Initialization for PyTorch
☆18Jun 27, 2018Updated 8 years ago
Alternatives and similar repositories for delta_orthogonal_init_pytorch
Users that are interested in delta_orthogonal_init_pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- paper lists and information on mean-field theory of deep learning☆79Mar 25, 2019Updated 7 years ago
- Accelerating Transfer Learning with Robust Neural Nets☆11Oct 2, 2020Updated 5 years ago
- This repository provides code source used in the paper: A Mean Field Theory of Quantized Deep Networks: The Quantization-Depth Trade-Off☆13May 30, 2019Updated 7 years ago
- Interpolation between Residual and Non-Residual Networks, ICML 2020. https://arxiv.org/abs/2006.05749☆26Aug 16, 2020Updated 5 years ago
- ☆12Sep 26, 2019Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for "Aggregated Momentum: Stability Through Passive Damping", Lucas et al. 2018☆37Nov 6, 2018Updated 7 years ago
- ☆13Jul 25, 2024Updated 2 years ago
- Code for reproducing the results in "How Well do Sparse Imagenet Models Transfer?", presented at CVPR 2022☆10Jun 3, 2022Updated 4 years ago
- Large-batch Training, Neural Network Optimization☆10Nov 8, 2019Updated 6 years ago
- DropNet: Reducing Neural Network Complexity via Iterative Pruning (ICML 2020)☆16Aug 24, 2020Updated 5 years ago
- ☆19Mar 18, 2021Updated 5 years ago
- Code for CVPR2021 paper: MOOD: Multi-level Out-of-distribution Detection☆38Sep 4, 2023Updated 2 years ago
- Codebase for the paper "Beyond BatchNorm: Towards a Unified Understanding of Normalization in Deep Learning"☆17Jul 12, 2021Updated 5 years ago
- Partially Adaptive Momentum Estimation method in the paper "Closing the Generalization Gap of Adaptive Gradient Methods in Training Deep …☆40Apr 13, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICLR 2021] Beyond Categorical Label Representations for Image Classification☆26Dec 12, 2021Updated 4 years ago
- [ECCV 2022] SuperTickets: Drawing Task-Agnostic Lottery Tickets from Supernets via Jointly Architecture Searching and Parameter Pruning☆20Jul 7, 2022Updated 4 years ago
- bilevel_augment is a ServiceNow Research project that was started at Element AI.☆27Jun 23, 2022Updated 4 years ago
- Experiments for the paper "Exponential expressivity in deep neural networks through transient chaos"☆74Jun 9, 2016Updated 10 years ago
- Deep Learning Research☆16Nov 13, 2019Updated 6 years ago
- Fast Axiomatic Attribution for Neural Networks (NeurIPS*2021)☆15Feb 24, 2026Updated 5 months ago
- A simple and efficient baseline for data attribution☆11Nov 10, 2023Updated 2 years ago
- Accepted by AAAI2022☆21Apr 10, 2022Updated 4 years ago
- TensorFlow implementation of (Momentum) Stochastic Variance-Adapted Gradient.☆45May 11, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official Implementation of "Transferring Inductive Biases Through Knowledge Distillation"☆15Jun 3, 2020Updated 6 years ago
- [IJCAI'19] Nostalgic Adam: Weighting more of the past gradients when designing the adaptive learning rate☆12Dec 3, 2019Updated 6 years ago
- Lookahead: A Far-sighted Alternative of Magnitude-based Pruning (ICLR 2020)☆32Oct 25, 2020Updated 5 years ago
- Linear-chain LSTM-CRFs and Convolutional CRFs in PyTorch.☆22Aug 11, 2017Updated 8 years ago
- A Closer Look at Accuracy vs. Robustness☆88May 17, 2021Updated 5 years ago
- ☆19Apr 28, 2021Updated 5 years ago
- Code for Self-Tuning Networks (ICLR 2019) https://arxiv.org/abs/1903.03088☆62Jun 18, 2019Updated 7 years ago
- Implementation for <Regularizing Neural Networks via Minimizing Hyperspherical Energy> in CVPR'20.☆24Jun 23, 2020Updated 6 years ago
- ☆19Jan 27, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- can calculate the Hessian matrix and/or its spectrum for simple neural nets☆11May 7, 2018Updated 8 years ago
- Pytorch optimizers implementing Hilbert Constrained Gradient Descent☆19May 9, 2019Updated 7 years ago
- Echo Noise Channel for Exact Mutual Information Calculation☆17Jul 17, 2020Updated 6 years ago
- ☆33May 21, 2020Updated 6 years ago
- Contains code for the NeurIPS 2020 paper by Pan et al., "Continual Deep Learning by FunctionalRegularisation of Memorable Past"☆44Nov 10, 2020Updated 5 years ago
- Riemannian approach to batch normalization☆18Nov 17, 2017Updated 8 years ago
- ☆21Sep 17, 2022Updated 3 years ago