[EMNLP 2024] Official implementation of "Rethinking Pruning Large Language Models: Benefits and Pitfalls of Reconstruction Error Minimization"
☆29Feb 21, 2025Updated last year
Alternatives and similar repositories for rethinking-LLM-pruning
Users that are interested in rethinking-LLM-pruning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Nov 10, 2025Updated 10 months ago
- [UAI 2025] Official implementation of "Critical Influence of Overparameterization on Sharpness-aware Minimization"☆21May 14, 2025Updated last year
- [ICML 2025] Official Pytorch code for "SASSHA: Sharpness-aware Adaptive Second-order Optimization With Stable Hessian Approximation"☆26Aug 11, 2025Updated last year
- 🔍 speek lets you peek into slurm resource info such as GPU avaiability, usage per user, job status, and more.☆20Apr 13, 2026Updated 5 months ago
- CharFormer(Tay et al., 2022; Gradient-based Subword Tokenizer + T5) model implementation for Huggingface Transformers☆19Oct 14, 2024Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆30Jul 22, 2024Updated 2 years ago
- Official implementation of "Multi-armed Bandit Algorithm against Strategic Replication"☆13May 17, 2022Updated 4 years ago
- [ICLR 2025] Official implementation of paper "Dynamic Low-Rank Sparse Adaptation for Large Language Models".☆25Mar 16, 2025Updated last year
- Official Repo for SparseLLM: Global Pruning of LLMs (NeurIPS 2024)☆71Mar 27, 2025Updated last year
- [ICML 2024] Official Implementation of SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks☆44Feb 4, 2025Updated last year
- Official PyTorch implementation of "Robust Deep Learning from Crowds with Belief Propagation"☆18Mar 22, 2022Updated 4 years ago
- Context-Robust Knowledge Editing for Language Models (ACL 2025 Findings)☆15Nov 3, 2025Updated 11 months ago
- ☆35May 24, 2024Updated 2 years ago
- Implementation of "Federated Full-Parameter Tuning of Billion-Sized Language Models with Communication Cost under 18 Kilobytes" (https://…☆13May 6, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Asynchronous pipeline parallel optimization☆23Feb 2, 2026Updated 8 months ago
- [ECCV24] MixDQ: Memory-Efficient Few-Step Text-to-Image Diffusion Models with Metric-Decoupled Mixed Precision Quantization☆14Nov 27, 2024Updated last year
- Official PyTorch implementation of "A Rotated Hyperbolic Wrapped Normal Distribution for Hierarchical Representation Learning"