ACCO: An optimization algorithm for sharded distributed LLM training.
☆13May 22, 2025Updated last year
Alternatives and similar repositories for acco
Users that are interested in acco are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV'25] ViLU: Learning Vision-Language Uncertainties for Failure Prediction☆16Jul 16, 2025Updated last year
- ENMA: Tokenwise Autoregression for Generative Neural PDE Operators☆18Mar 19, 2026Updated 4 months ago
- This is the official implementation of the ICML 2023 paper - Can Forward Gradient Match Backpropagation ?☆13May 31, 2023Updated 3 years ago
- Implementation of the ICLR 2022 paper "Phase Collapse in Neural Networks."☆10Mar 21, 2022Updated 4 years ago
- ☆28Jul 17, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A simple reference implementation of the single-worker MuLoCo optimizer in Jax & PyTorch. MuLoCo-1 has been shown to outperfrom Muon and …☆32Feb 27, 2026Updated 5 months ago
- Common tools for data processing☆22Dec 8, 2025Updated 8 months ago
- Scala Native 3 bindings for SFML library☆15Jul 9, 2023Updated 3 years ago
- [ICML-2025] We introduce Lie group Relative position Encodings (LieRE) that goes beyond RoPE in supporting n-dimensional inputs.☆37Aug 13, 2025Updated 11 months ago
- Wavelet Phase Harmonics in PyTorch for Images☆20Jan 28, 2025Updated last year
- ScatNetLight for fast classifications of signals via Scattering Networks☆15Apr 20, 2017Updated 9 years ago
- Yet another implementation of Python decorators in ruby☆10Aug 24, 2017Updated 8 years ago
- ☆11Jun 11, 2025Updated last year
- h-Shap provides an exact, fast, hierarchical implementation of Shapley coefficients for image explanations☆16Nov 20, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- My dotfiles☆11Mar 31, 2023Updated 3 years ago
- [Poster; ICLR 2026] [Oral; Neurips OPT2024] μLO: Compute-Efficient Meta-Generalization of Learned Optimizers☆16Apr 15, 2026Updated 3 months ago
- ☆43Jul 22, 2024Updated 2 years ago
- Official implementation of DIP: Unsupervised Dense In-Context Post-training of Visual Representations