SGD with compressed gradients and error-feedback: https://arxiv.org/abs/1901.09847
☆31Jul 25, 2024Updated 2 years ago
Alternatives and similar repositories for error-feedback-SGD
Users that are interested in error-feedback-SGD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- QSGD-TF☆21May 15, 2019Updated 7 years ago
- Code for the signSGD paper☆95Jan 12, 2021Updated 5 years ago
- Atomo: Communication-efficient Learning via Atomic Sparsification☆29Dec 9, 2018Updated 7 years ago
- ☆32Dec 3, 2019Updated 6 years ago
- ☆77Jun 7, 2019Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Simple Hierarchical Count Sketch in Python☆21Jun 3, 2021Updated 5 years ago
- Decentralized SGD and Consensus with Communication Compression: https://arxiv.org/abs/1907.09356☆74Sep 10, 2020Updated 5 years ago
- MISSION: Ultra Large-Scale Feature Selection using Count-Sketches☆13Oct 6, 2019Updated 6 years ago
- YALL1: Your ALgorithms for L1☆13Jan 28, 2018Updated 8 years ago
- Adaptive gradient sparsification for efficient federated learning: an online learning approach☆18Oct 31, 2020Updated 5 years ago
- Understanding Top-k Sparsification in Distributed Deep Learning☆24Nov 15, 2019Updated 6 years ago
- Layer-wise Sparsification of Distributed Deep Learning☆10Jul 6, 2020Updated 6 years ago
- Stochastic Gradient Push for Distributed Deep Learning☆172Apr 5, 2023Updated 3 years ago
- [ICLR 2018] Deep Gradient Compression: Reducing the Communication Bandwidth for Distributed Training☆226Jul 10, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A compressed adaptive optimizer for training large-scale deep learning models using PyTorch☆25Nov 26, 2019Updated 6 years ago
- ☆14Nov 3, 2019Updated 6 years ago
- The code for the paper "QuAFL: Federated Averaging Can Be Both Asynchronous and Communication-Efficient"☆17Mar 26, 2023Updated 3 years ago
- It is implementation of Research paper "DEEP GRADIENT COMPRESSION: REDUCING THE COMMUNICATION BANDWIDTH FOR DISTRIBUTED TRAINING". Deep g…☆18Aug 14, 2019Updated 6 years ago
- Sketched SGD☆29Jul 4, 2020Updated 6 years ago
- Code and Results for Master Thesis Project on Fixed-point Quantization of Convolutional Neural Networks for Quantized Inference on Embedd…☆13Feb 7, 2021Updated 5 years ago
- ☆14Mar 13, 2023Updated 3 years ago
- Code related to ’Beyond spectral gap: The role of the topology in decentralized learning‘.☆14Jun 7, 2022Updated 4 years ago
- A Sparse-tensor Communication Framework for Distributed Deep Learning☆13Nov 1, 2021Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Audio Keyword Search☆12May 5, 2019Updated 7 years ago
- PyTorch implementations of neural network models for keyword spotting☆11Oct 19, 2020Updated 5 years ago
- Artifacts of VLDB'22 paper "COMET: A Novel Memory-Efficient Deep Learning TrainingFramework by Using Error-Bounded Lossy Compression"☆10Aug 2, 2022Updated 4 years ago
- Partial implementation of paper "DEEP GRADIENT COMPRESSION: REDUCING THE COMMUNICATION BANDWIDTH FOR DISTRIBUTED TRAINING"☆33Nov 20, 2020Updated 5 years ago
- Code for paper "Learning a Code: Machine Learning for Approximate Non-Linear Coded-Computation"☆10Dec 21, 2020Updated 5 years ago
- ☆10May 4, 2018Updated 8 years ago
- A LaTeX template for note☆10May 4, 2023Updated 3 years ago
- Zeroth-order Min-max Optimization☆13Jun 28, 2020Updated 6 years ago
- An attempt to replicate the paper "Multi-shot Pedestrian Re-identification via Sequential Decision Making (CVPR2018)"☆10Nov 16, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Coordinate Descent Fuzzy Twin Support Vector Machine for Classification☆11Jan 13, 2018Updated 8 years ago
- Certifying Some Distributional Robustness with Principled Adversarial Training (https://arxiv.org/abs/1710.10571)☆45May 1, 2018Updated 8 years ago
- Prof. S. Boyd's LaTeX Templates☆13Dec 18, 2018Updated 7 years ago
- This is an implementation of ResNet-34 in TensorFlow2.0 using the Imperative API (subclassing tensorflow.keras.Model)☆12Dec 11, 2020Updated 5 years ago
- Application of topic models for topic extraction and similarity search☆15Sep 1, 2020Updated 5 years ago
- Implementation of the SuRP algorithm by the authors of the AISTATS 2022 paper "An Information-Theoretic Justification for Model Pruning".…☆13May 4, 2022Updated 4 years ago
- 🎓Automatically Update Distributed Learning Papers Daily using Github Actions (Update Every 12th hours)☆48Updated this week