torch implementation of diloco
☆26Jul 17, 2026Updated 3 weeks ago
Alternatives and similar repositories for diloco_simple
Users that are interested in diloco_simple are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆51Jan 18, 2024Updated 2 years ago
- ☆14Apr 24, 2024Updated 2 years ago
- ☆10Aug 18, 2016Updated 9 years ago
- Parallel Stable Sort☆15Oct 11, 2015Updated 10 years ago
- implementation of nori the raytracer☆10Aug 12, 2019Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- TileGraph is an experimental DNN compiler that utilizes static code generation and kernel fusion techniques.☆11Sep 18, 2024Updated last year
- Parallel Self-Adjusting Computation☆18Jul 5, 2021Updated 5 years ago
- Transformer in Chemical Language Model sometimes misunderstands chirality☆13Apr 19, 2024Updated 2 years ago
- Large-batch Training, Neural Network Optimization☆10Nov 8, 2019Updated 6 years ago
- ☆10Jun 19, 2023Updated 3 years ago
- implementation of https://arxiv.org/pdf/2312.09299☆21Jul 3, 2024Updated 2 years ago
- Advanced Programming for Computer Design Problems☆17Aug 28, 2021Updated 4 years ago
- ☆14Mar 2, 2025Updated last year
- Single-header logger with pretty console output☆20Mar 23, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Python package for rematerialization-aware gradient checkpointing☆27Oct 31, 2023Updated 2 years ago
- ☆15Sep 24, 2023Updated 2 years ago
- [Poster; ICLR 2026] [Oral; Neurips OPT2024] μLO: Compute-Efficient Meta-Generalization of Learned Optimizers☆16Apr 15, 2026Updated 3 months ago
- Grams: Gradient Descent with Adaptive Momentum Scaling (ICLR 2025 Workshop)☆17Mar 6, 2025Updated last year
- Resources regarding evML (edge verified machine learning)☆24Jan 4, 2025Updated last year
- ☆10Apr 23, 2021Updated 5 years ago
- Parallel processing with sequential output, respecting order of input☆10Feb 20, 2023Updated 3 years ago
- Code related to ’Beyond spectral gap: The role of the topology in decentralized learning‘.☆14Jun 7, 2022Updated 4 years ago
- Evaluate state-of-the-art sparse embedding models on the LIMIT dataset (`limit-small` and `limit`) from google's paper `On the Theoretica…☆16Sep 4, 2025Updated 11 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A unified tree search framework for molecular generation.☆20Jul 21, 2026Updated 2 weeks ago
- Rust port of pandoc-types☆16Feb 11, 2023Updated 3 years ago
- ☆17Apr 28, 2020Updated 6 years ago
- Code for the paper "Understanding the Role of Momentum in Stochastic Gradient Methods"☆14Oct 27, 2019Updated 6 years ago
- Code and results accompanying our paper titled Leveraging Unlabeled Data to Predict Out-of-Distribution Performance at ICLR 2022☆11Dec 8, 2022Updated 3 years ago
- Fork of NACA from Google Code☆13Feb 25, 2010Updated 16 years ago
- Stochastic Weight Averaging Tutorials using pytorch.☆33Oct 23, 2020Updated 5 years ago
- Github Repo for ICML 2022 paper: Communication-Efficient Adaptive Federated Learning☆10Nov 18, 2022Updated 3 years ago
- Thread pool which supports c++20 coroutine. 一个支持c++20协程的线程池。☆24Mar 3, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Mille Crepe Bench: layer-wise performance analysis for deep learning frameworks.☆18Oct 22, 2019Updated 6 years ago
- Examples to control the Opal C1 from within python.☆18May 7, 2023Updated 3 years ago
- The code for paper Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models.☆13Apr 10, 2024Updated 2 years ago
- Implementation of Gradient Information Optimization (GIO) for effective and scalable training data selection☆14Jun 22, 2023Updated 3 years ago
- Specific implementation (based on the public rbuilder) of a block builder to be used on a TDX context.☆18Sep 30, 2025Updated 10 months ago
- SuperCLUE高考作文机器自动阅卷系统☆19Jun 8, 2023Updated 3 years ago
- ænet-PyTorch: a GPU-supported implementation for machine learning atomic potentials training☆17Mar 18, 2024Updated 2 years ago