Parameter Efficient Transfer Learning with Diff Pruning
☆74Feb 3, 2021Updated 5 years ago
Alternatives and similar repositories for DiffPruning
Users that are interested in DiffPruning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for "Training Neural Networks with Fixed Sparse Masks" (NeurIPS 2021).☆59Jan 14, 2022Updated 4 years ago
- [NAACL 2022] "Learning to Win Lottery Tickets in BERT Transfer via Task-agnostic Mask Training", Yuanxin Liu, Fandong Meng, Zheng Lin, Pe…☆15Oct 18, 2022Updated 3 years ago
- Sequence-Level Mixed Sample Data Augmentation☆23Mar 7, 2021Updated 5 years ago
- Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models☆143Sep 4, 2022Updated 3 years ago
- Source code for "Importance-based Neuron Allocation for Multilingual Neural Machine Translation"☆12Sep 15, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official code repository for AAAI2021 paper Finding Sparse Structures for Domain Specific Neural Machine Translation☆11Apr 1, 2021Updated 5 years ago
- [NeurIPS 2023] Make Your Pre-trained Model Reversible: From Parameter to Memory Efficient Fine-Tuning☆33Jun 2, 2023Updated 3 years ago
- [ACL 2022] Structured Pruning Learns Compact and Accurate Models https://arxiv.org/abs/2204.00408☆199May 9, 2023Updated 3 years ago
- Structured Pruning Adapters in PyTorch☆19Aug 30, 2023Updated 2 years ago
- ☆54May 8, 2023Updated 3 years ago
- Implementation of paper "Towards a Unified View of Parameter-Efficient Transfer Learning" (ICLR 2022)☆542Mar 24, 2022Updated 4 years ago
- ☆11Mar 25, 2022Updated 4 years ago
- Embedding Recycling for Language models☆38Jul 11, 2023Updated 3 years ago
- ☆32Dec 17, 2025Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆146Jul 21, 2024Updated 2 years ago
- c++ mosestokenizer☆18Mar 13, 2024Updated 2 years ago
- ☆131Aug 18, 2022Updated 3 years ago
- Less is More: Task-aware Layer-wise Distillation for Language Model Compression (ICML2023)☆40Aug 28, 2023Updated 2 years ago
- ☆31Apr 27, 2022Updated 4 years ago
- On the Effectiveness of Parameter-Efficient Fine-Tuning☆39Nov 4, 2023Updated 2 years ago
- Zero-shot Learning by Generating Task-specific Adapters☆14Apr 2, 2021Updated 5 years ago
- ☆27Dec 15, 2022Updated 3 years ago
- This is the official implementation for the paper "Learning to Scaffold: Optimizing Model Explanations for Teaching"☆20May 19, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official codebase for our paper "Joslim: Joint Widths and Weights Optimization for Slimmable Neural Networks"☆12Jun 30, 2021Updated 5 years ago
- ☆12Apr 18, 2019Updated 7 years ago
- MUX-PLMs: Pretraining LMs with Data Multiplexing☆15Jan 29, 2023Updated 3 years ago
- ☆20Dec 16, 2020Updated 5 years ago
- [ACL-IJCNLP 2021] "EarlyBERT: Efficient BERT Training via Early-bird Lottery Tickets" by Xiaohan Chen, Yu Cheng, Shuohang Wang, Zhe Gan, …☆18Dec 30, 2021Updated 4 years ago
- Improving Transformation Invariance in Contrastive Representation Learning☆12Mar 13, 2021Updated 5 years ago
- Combining encoder-based language models☆11Nov 11, 2021Updated 4 years ago
- ☆226Feb 21, 2023Updated 3 years ago
- ☆27Oct 23, 2018Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆12Jul 7, 2021Updated 5 years ago
- [ACL 2023] Counterspeeches up my sleeve! Intent Distribution Learning and Persistent Fusion for Intent-Conditioned Counterspeech Generati…☆10Sep 23, 2023Updated 2 years ago
- ☆10Jul 5, 2019Updated 7 years ago
- Training and evaluation codes for the BertGen paper (ACL-IJCNLP 2021)☆11Sep 17, 2023Updated 2 years ago
- Minimum viable code for the Decodable Information Bottleneck paper. Pytorch Implementation.☆12Oct 20, 2020Updated 5 years ago
- Implementation of EPSNet for panoptic segmentation on PyTorch☆14Sep 5, 2020Updated 5 years ago
- Code for "Can We Characterize Tasks Without Labels or Features?" (CVPR 2021)☆11Aug 31, 2021Updated 4 years ago