Practical low-rank gradient compression for distributed optimization: https://arxiv.org/abs/1905.13727
☆151Oct 29, 2024Updated last year
Alternatives and similar repositories for powersgd
Users that are interested in powersgd are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Atomo: Communication-efficient Learning via Atomic Sparsification☆29Dec 9, 2018Updated 7 years ago
- GRACE - GRAdient ComprEssion for distributed deep learning☆141Jul 23, 2024Updated last year
- QSGD-TF☆21May 15, 2019Updated 7 years ago
- Ok-Topk is a scheme for distributed training with sparse gradients. Ok-Topk integrates a novel sparse allreduce algorithm (less than 6k c…☆27Dec 10, 2022Updated 3 years ago
- [ICLR 2018] Deep Gradient Compression: Reducing the Communication Bandwidth for Distributed Training☆226Jul 10, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Stochastic Gradient Push for Distributed Deep Learning☆172Apr 5, 2023Updated 3 years ago
- It is implementation of Research paper "DEEP GRADIENT COMPRESSION: REDUCING THE COMMUNICATION BANDWIDTH FOR DISTRIBUTED TRAINING". Deep g…☆18Aug 14, 2019Updated 6 years ago
- Decentralized SGD and Consensus with Communication Compression: https://arxiv.org/abs/1907.09356☆74Sep 10, 2020Updated 5 years ago