☆27Aug 25, 2023Updated 3 years ago
Alternatives and similar repositories for CocktailSGD
Users that are interested in CocktailSGD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deferred Continuous Batching in Resource-Efficient Large Language Model Serving (EuroMLSys 2024)☆19May 28, 2024Updated 2 years ago
- This repository is the official implementation of 'EDEN: Communication-Efficient and Robust Distributed Mean Estimation for Federated Lea…☆18May 5, 2026Updated 5 months ago
- Code for "Practical Low-Rank Communication Compression in Decentralized Deep Learning"☆17Aug 4, 2020Updated 6 years ago
- Utilities for Training Very Large Models☆58Sep 25, 2024Updated 2 years ago
- The official implementation of TinyTrain [ICML '24]☆28Jul 19, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Practical low-rank gradient compression for distributed optimization: https://arxiv.org/abs/1905.13727☆152Oct 29, 2024Updated last year
- ☆10Apr 29, 2024Updated 2 years ago
- summer school materials☆46Aug 4, 2023Updated 3 years ago
- ☆12Mar 31, 2020Updated 6 years ago
- ☆14May 25, 2022Updated 4 years ago
- Associated codebase for Byzantine-resilient distributed / decentralized machine learning papers from INSPIRE Lab☆14Oct 11, 2021Updated 4 years ago
- DETOX: A Redundancy-based Framework for Faster and More Robust Gradient Aggregation☆16Jul 13, 2020Updated 6 years ago
- ☆19May 4, 2023Updated 3 years ago
- FaceGrabber is introduced in the following paper: D. Merget, T. Eckl, M. Schwörer, P. Tiefenbacher, and G. Rigoll, “Capturing Facial Vide…☆11Sep 7, 2016Updated 10 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for paper "Byzantine-Resilient Decentralized Stochastic Optimization with Robust Aggregation Rules"☆21Apr 19, 2024Updated 2 years ago
- List Flower resources☆12Feb 4, 2022Updated 4 years ago
- PolarGrad: A Class of Matrix-Gradient Optimizers from a Unifying Preconditioning Perspective☆18Aug 10, 2026Updated last month
- ☆13Oct 20, 2021Updated 4 years ago
- ☆10Jun 19, 2023Updated 3 years ago
- Collaborative inference of latent diffusion via hivemind☆12May 29, 2023Updated 3 years ago
- ☆152Jun 2, 2023Updated 3 years ago
- Source codes for paper "BlockFFN: Towards End-Side Acceleration-Friendly Mixture-of-Experts with Chunk-Level Activation Sparsity".☆19Jan 10, 2026Updated 9 months ago
- Artifact evaluation for HPCA'24 paper Lightening-Transformer: A Dynamically-operated Optically-interconnected Photonic Transformer Accele…☆11Mar 3, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆13Jun 8, 2021Updated 5 years ago
- ☆21Oct 23, 2024Updated last year
- Inducing Point Operator Transformer: A Flexible and Scalable Architecture for Solving PDEs (AAAI 2024)☆16Jul 30, 2024Updated 2 years ago
- Blog post☆17Feb 16, 2024Updated 2 years ago
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"☆17Jun 30, 2025Updated last year
- Data Poisoning in Deep Learning: A Survey☆28May 26, 2026Updated 4 months ago
- Word Embeddings for Low Resource Languages: The Case of Buryat☆10Mar 12, 2025Updated last year
- ☆15Sep 24, 2023Updated 3 years ago
- Grams: Gradient Descent with Adaptive Momentum Scaling (ICLR 2025 Workshop)☆17Mar 6, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Resources regarding evML (edge verified machine learning)☆24Jan 4, 2025Updated last year
- [ICLR24] Better Neural PDE Solvers Through Data-Free Mesh Movers☆18Mar 20, 2024Updated 2 years ago
- ☆15Nov 7, 2024Updated last year
- Code related to ’Beyond spectral gap: The role of the topology in decentralized learning‘.☆14Jun 7, 2022Updated 4 years ago
- PyTorch code for our paper "AdaSVD: Adaptive Singular Value Decomposition for Large Language Models"☆16Mar 9, 2025Updated last year
- Simple script to re-rank images using OpenAI's CLIP https://github.com/openai/CLIP.☆15May 3, 2021Updated 5 years ago
- Code for the paper "Secure Distributed Training at Scale" (ICML 2022)☆16Feb 4, 2025Updated last year