[ICLR 2024] The Need for Speed: Pruning Transformers with One Recipe
☆28Sep 2, 2024Updated 2 years ago
Alternatives and similar repositories for optin-transformer-pruning
Users that are interested in optin-transformer-pruning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2024] SparseRefine: Sparse Refinement for Efficient High-Resolution Semantic Segmentation☆16Jan 10, 2025Updated last year
- [ICCV 2023] DataDAM: Efficient Dataset Distillation with Attention Matching☆34Jun 20, 2024Updated 2 years ago
- A repo for publishing solution to 3DCoMPaT++ challenge on an improved large-scale 3D vision dataset for compositional recognition☆14Jun 22, 2023Updated 3 years ago
- Loss Function Search for Face Recognition☆41Jan 9, 2021Updated 5 years ago
- Official code for paper "OpenCIL: Benchmarking Out-of-Distribution Detection in Class-Incremental Learning"☆13Jun 19, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Code for CVPR24 Paper - Resource-Efficient Transformer Pruning for Finetuning of Large Models☆12Oct 31, 2025Updated 10 months ago
- UniQL official repository (ICLR 2026)☆19Jan 27, 2026Updated 7 months ago
- [WACV 2025] Official code release for Transientangelo: Few-Viewpoint Surface Reconstruction Using Single-Photon Lidar☆24Oct 29, 2024Updated last year
- [CVPR'24] Once for Both: Single Stage of Importance and Sparsity Search for Vision Transformer Compression☆16Jul 1, 2024Updated 2 years ago
- [ICML 2025] SparseLoRA: Accelerating LLM Fine-Tuning with Contextual Sparsity☆78Mar 10, 2026Updated 5 months ago
- model-compression-and-acceleration-4-DNN☆21Nov 29, 2018Updated 7 years ago
- PyTorch implementation of "Learning from Students: Online Contrastive Distillation Network for General Continual Learning" (IJCAI 2022)☆11Dec 29, 2022Updated 3 years ago
- ☆15May 28, 2024Updated 2 years ago
- PyTorch Implementation of Self-Supervised Learning models☆13Apr 25, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repo is re-produce for Channel_pruning☆11May 17, 2018Updated 8 years ago
- mycloudhome is a cli tool for Western Digital MY CLOUD HOME☆22Feb 24, 2022Updated 4 years ago
- A repository to keep track of literature on catastrophic forgetting☆36Mar 10, 2020Updated 6 years ago
- Code for Improving Task-free Continual Learning by Distributionally Robust Memory Evolution (ICML 2022)☆11Aug 20, 2022Updated 4 years ago
- This is the code for CVPR2022 paper "Modeling Motion with Multi-Modal Features for Text-Based Video Segmentation"☆19Feb 19, 2023Updated 3 years ago
- Curse-of-memory phenomenon of RNNs in sequence modelling☆19May 8, 2025Updated last year
- Miro[ACM MobiCom '23] Cost-effective On-device Continual Learning over Memory Hierarchy with Miro☆16Feb 1, 2024Updated 2 years ago
- The official implementation of "Helen: Optimizing CTR Prediction Models with Frequency-wise Hessian Eigenvalue Regularization"☆16Mar 14, 2024Updated 2 years ago
- PyTorch code for our CoLLAs-2022 paper "Online Continual Learning for Embedded Devices"☆13Aug 4, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Examples for mixed-precision training for utilizing TensorCores in NVIDIA Volta GPUs☆16Jun 21, 2018Updated 8 years ago
- Official PyTorch implementation of "LGViT: Dynamic Early Exiting for Accelerating Vision Transformer" (ACM MM 2023)☆16Nov 18, 2024Updated last year
- Code release for "Understanding Bias in Large-Scale Visual Datasets"☆25Dec 4, 2024Updated last year
- [NeurIPS‘2021] "MEST: Accurate and Fast Memory-Economic Sparse Training Framework on the Edge", Geng Yuan, Xiaolong Ma, Yanzhi Wang et al…☆18Mar 16, 2022Updated 4 years ago
- Code for NeurIPS 2021 paper "Flattening Sharpness for Dynamic Gradient Projection Memory Benefits Continual Learning".☆16Oct 18, 2021Updated 4 years ago
- Official code release for Delta Activations: A Representation for Finetuned Large Language Models☆21Sep 5, 2025Updated last year
- ☆18Apr 8, 2025Updated last year
- RepoZero: Can LLMs Generate a Code Repository from Scratch? (https://arxiv.org/abs/2605.07122)☆31Jun 4, 2026Updated 3 months ago
- Post processing library used to analyze memory snapshots☆37May 29, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for JPTS:Enhancing Deep Learning Performance of Massive MIMO CSI Feedback☆17Jan 18, 2023Updated 3 years ago
- 🎓Automatically Update circult-eda-mlsys-tinyml Papers Daily using Github Actions (Update Every 8th hours)☆10Aug 31, 2026Updated last week
- ☆22Apr 24, 2025Updated last year
- Create tiny ML systems for on-device learning.☆19Jul 14, 2021Updated 5 years ago
- Source code for "Gradient Based Memory Editing for Task-Free Continual Learning", 4th Lifelong ML Workshop@ICML 2020☆17Dec 8, 2022Updated 3 years ago
- JAX implementation of configurable LLM distillation training☆24Nov 15, 2025Updated 9 months ago
- Code release for "Generative Modeling of Weights: Generalization or Memorization?"☆23Apr 9, 2026Updated 4 months ago