☆27Aug 25, 2023Updated 3 years ago
Alternatives and similar repositories for CocktailSGD
Users that are interested in CocktailSGD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deferred Continuous Batching in Resource-Efficient Large Language Model Serving (EuroMLSys 2024)☆19May 28, 2024Updated 2 years ago
- This repository is the official implementation of 'EDEN: Communication-Efficient and Robust Distributed Mean Estimation for Federated Lea…☆18May 5, 2026Updated 4 months ago
- ☆14May 4, 2026Updated 4 months ago
- [ICDCS 2023] Evaluation and Optimization of Gradient Compression for Distributed Deep Learning☆10Apr 28, 2023Updated 3 years ago
- This repository contains code for the MicroAdam paper.☆21Dec 14, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for "Practical Low-Rank Communication Compression in Decentralized Deep Learning"☆17Aug 4, 2020Updated 6 years ago
- Utilities for Training Very Large Models☆58Sep 25, 2024Updated last year
- The official implementation of TinyTrain [ICML '24]☆28Jul 19, 2024Updated 2 years ago
- Practical low-rank gradient compression for distributed optimization: https://arxiv.org/abs/1905.13727☆151Oct 29, 2024Updated last year
- summer school materials☆46Aug 4, 2023Updated 3 years ago
- Associated codebase for Byzantine-resilient distributed / decentralized machine learning papers from INSPIRE Lab☆14Oct 11, 2021Updated 4 years ago
- crystalnet -- a mini core AI library (being refactored, see https://github.com/lgarithm/stdnn-ops)☆17Oct 1, 2019Updated 6 years ago
- Implementation of the FedPM framework by the authors of the ICLR 2023 paper "Sparse Random Networks for Communication-Efficient Federated…☆30Feb 10, 2023Updated 3 years ago
- ☆19May 4, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- FaceGrabber is introduced in the following paper: D. Merget, T. Eckl, M. Schwörer, P. Tiefenbacher, and G. Rigoll, “Capturing Facial Vide…☆11Sep 7, 2016Updated 10 years ago
- Code for paper "Byzantine-Resilient Decentralized Stochastic Optimization with Robust Aggregation Rules"☆20Apr 19, 2024Updated 2 years ago
- Test scripts for exploring PyTorch JIT and quantization capability☆11Mar 8, 2021Updated 5 years ago
- List Flower resources☆12Feb 4, 2022Updated 4 years ago
- PolarGrad: A Class of Matrix-Gradient Optimizers from a Unifying Preconditioning Perspective☆18Aug 10, 2026Updated last month
- ☆10Jun 19, 2023Updated 3 years ago
- HALO: Hadamard-Assisted Low-Precision Optimization and Training method for finetuning LLMs. 🚀 The official implementation of https://arx…☆32Feb 17, 2025Updated last year
- An iOS app that integrates a Large Language Model (LLM) to process audio recordings for transcription and summarization.☆17Nov 29, 2024Updated last year
- Collaborative inference of latent diffusion via hivemind☆12May 29, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [WACV 2024] Meta-Learned Kernel For Blind Super-Resolution Kernel Estimation☆14Jul 11, 2024Updated 2 years ago
- ☆152Jun 2, 2023Updated 3 years ago
- Source codes for paper "BlockFFN: Towards End-Side Acceleration-Friendly Mixture-of-Experts with Chunk-Level Activation Sparsity".☆19Jan 10, 2026Updated 8 months ago
- ☆13Jun 8, 2021Updated 5 years ago
- ☆21Oct 23, 2024Updated last year
- ☆12Dec 9, 2020Updated 5 years ago
- [ICML 2026] Less Is More: Training-Free Sparse Attention with Global Locality for Efficient Reasoning☆36Sep 12, 2025Updated last year
- Blog post☆17Feb 16, 2024Updated 2 years ago
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"☆17Jun 30, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Data Poisoning in Deep Learning: A Survey☆26May 26, 2026Updated 3 months ago
- Word Embeddings for Low Resource Languages: The Case of Buryat☆10Mar 12, 2025Updated last year
- ☆15Sep 24, 2023Updated 2 years ago
- Shortcuts for AWS EC2 Instance Control from the command-line: list, start, stop and ssh☆16Jun 10, 2018Updated 8 years ago
- 🎬 3.7× faster video generation E2E 🖼️ 1.6× faster image generation E2E ⚡ ColumnSparseAttn 9.3× vs FlashAttn‑3 💨 ColumnSparseGEMM 2.5× …☆113Sep 8, 2025Updated last year
- Solve ciphers with python☆10Oct 24, 2018Updated 7 years ago
- Grams: Gradient Descent with Adaptive Momentum Scaling (ICLR 2025 Workshop)☆17Mar 6, 2025Updated last year