Efficient 2:4 sparse training algorithms and implementations
☆63Dec 8, 2024Updated last year
Alternatives and similar repositories for 2by4-pretrain
Users that are interested in 2by4-pretrain are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for "Accelerating Transformer Pre-training with 2:4 Sparsity"☆28Dec 8, 2024Updated last year
- Official implementation for "Pruning Large Language Models with Semi-Structural Adaptive Sparse Training" (AAAI 2025)☆20Jul 1, 2025Updated last year
- A selective knowledge distillation algorithm for efficient speculative decoders