Explore training for quantized models
☆28Jul 12, 2025Updated last year
Alternatives and similar repositories for quantized-training
Users that are interested in quantized-training are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- High-performance tokenized language data-loader for Python C++ extension☆15Jul 22, 2024Updated 2 years ago
- Zeta implementation of a reusable and plug in and play feedforward from the paper "Exponentially Faster Language Modeling"☆16Nov 11, 2024Updated last year
- Inline PTX Assembly in CUDA example☆15May 7, 2022Updated 4 years ago
- 🔀 yet another mixture of experts☆23Jun 5, 2026Updated 2 months ago
- ☆14Feb 17, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [PACT'24] GraNNDis. A fast and unified distributed graph neural network (GNN) training framework for both full-batch (full-graph) and min…☆10Aug 13, 2024Updated 2 years ago
- Weakly Supervised Object Localization via Class RE-Activation Mapping☆12Sep 19, 2022Updated 3 years ago
- small c99 blas inspired routines for finite field algebra☆12Jan 29, 2021Updated 5 years ago
- BPE modification that implements removing of the intermediate tokens during tokenizer training.☆27Nov 25, 2024Updated last year
- HALO: Hadamard-Assisted Low-Precision Optimization and Training method for finetuning LLMs. 🚀 The official implementation of https://arx…☆31Feb 17, 2025Updated last year
- ☆16Dec 29, 2024Updated last year
- An extention to the GaLore paper, to perform Natural Gradient Descent in low rank subspace☆19Oct 21, 2024Updated last year
- 1st Place Team Crane: @aswinkumar1999 @rathull @kyolebu☆32Sep 8, 2025Updated 11 months ago
- Learn CUDA with PyTorch☆366Jun 1, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Real-time (ish) data messaging over WebSockets, inspired by RTMFP☆19Feb 12, 2026Updated 6 months ago
- ☆14May 23, 2023Updated 3 years ago
- simple grpo☆12May 28, 2025Updated last year
- ☆14May 4, 2026Updated 3 months ago
- ☆92Feb 29, 2024Updated 2 years ago
- Write a fast kernel and see how you compare against the best humans and AI on gpumode.com☆109Jul 29, 2026Updated 2 weeks ago
- ☆18Nov 10, 2025Updated 9 months ago
- A framework for fast exploration of the depth-first scheduling space for DNN accelerators☆43Feb 8, 2023Updated 3 years ago
- ☆17Sep 6, 2025Updated 11 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Simple C++ wrapper of the SRT protocol for building Server/Client transport solutions☆20May 7, 2024Updated 2 years ago
- ☆11Aug 2, 2024Updated 2 years ago
- Official code for Class Tokens Infusion for Weakly Supervised Semantic Segmentation, CVPR2024☆23Oct 26, 2024Updated last year
- https://x.com/BlinkDL_AI/status/1884768989743882276☆28May 4, 2025Updated last year
- ☆51May 20, 2025Updated last year
- See https://youtube-dl.org/☆10Oct 24, 2020Updated 5 years ago
- ☆20Apr 16, 2025Updated last year
- This repository contains the experimental PyTorch native float8 training UX☆226Aug 1, 2024Updated 2 years ago
- Python scripts to facilitate easy working☆11Mar 23, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆22Jul 30, 2024Updated 2 years ago
- ☆11Jul 28, 2026Updated 2 weeks ago
- Model Predictive Path Integral Control (MPPI) with PyTorch☆18Jan 26, 2024Updated 2 years ago
- Python implementation of Efficient Graph-Based Image Segmentation☆25Sep 26, 2020Updated 5 years ago
- Community maintained hardware plugin for vLLM on Intel Gaudi☆51Updated this week
- SPAA'21: Efficient Stepping Algorithms and Implementations for Parallel Shortest Paths☆21Aug 10, 2024Updated 2 years ago
- Reinforcement learning modular with pytorch☆11Jan 18, 2021Updated 5 years ago