Parameter Efficient Transfer Learning with Diff Pruning
☆73Feb 3, 2021Updated 5 years ago
Alternatives and similar repositories for DiffPruning
Users that are interested in DiffPruning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for "Training Neural Networks with Fixed Sparse Masks" (NeurIPS 2021).☆59Jan 14, 2022Updated 4 years ago
- Sequence-Level Mixed Sample Data Augmentation☆22Mar 7, 2021Updated 5 years ago
- Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models☆143Sep 4, 2022Updated 4 years ago
- Source code for "Importance-based Neuron Allocation for Multilingual Neural Machine Translation"☆12Sep 15, 2021Updated 5 years ago
- Official code repository for AAAI2021 paper Finding Sparse Structures for Domain Specific Neural Machine Translation☆11Apr 1, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NeurIPS 2023] Make Your Pre-trained Model Reversible: From Parameter to Memory Efficient Fine-Tuning☆33Jun 2, 2023Updated 3 years ago
- [ACL 2022] Structured Pruning Learns Compact and Accurate Models https://arxiv.org/abs/2204.00408☆198May 9, 2023Updated 3 years ago
- ☆54May 8, 2023Updated 3 years ago
- Implementation of paper "Towards a Unified View of Parameter-Efficient Transfer Learning" (ICLR 2022)☆541Mar 24, 2022Updated 4 years ago
- ☆11Mar 25, 2022Updated 4 years ago
- A library for parameter-efficient and composable transfer learning for NLP with sparse fine-tunings.☆75Aug 9, 2024Updated 2 years ago
- Official codebase accompanying our ACL 2022 paper "RELiC: Retrieving Evidence for Literary Claims" (https://relic.cs.umass.edu).☆20May 14, 2022Updated 4 years ago
- ☆33Dec 17, 2025Updated 9 months ago
- ☆146Jul 21, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- c++ mosestokenizer☆18Mar 13, 2024Updated 2 years ago
- ☆131Aug 18, 2022Updated 4 years ago
- Less is More: Task-aware Layer-wise Distillation for Language Model Compression (ICML2023)☆41Aug 28, 2023Updated 3 years ago
- Implementation of paper: Extending and Analyzing Self-Supervised Learning Across Domains☆10Jan 10, 2021Updated 5 years ago
- ☆31Apr 27, 2022Updated 4 years ago
- On the Effectiveness of Parameter-Efficient Fine-Tuning☆39Nov 4, 2023Updated 2 years ago
- ☆27Dec 15, 2022Updated 3 years ago
- Is BERT Robust to Label Noise? A Study on Learning with Noisy Labels in Text Classification☆11May 31, 2022Updated 4 years ago
- This is the official implementation for the paper "Learning to Scaffold: Optimizing Model Explanations for Teaching"☆20May 19, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆12Apr 18, 2019Updated 7 years ago
- ☆15Apr 10, 2018Updated 8 years ago
- ☆20Dec 16, 2020Updated 5 years ago
- [ACL-IJCNLP 2021] "EarlyBERT: Efficient BERT Training via Early-bird Lottery Tickets" by Xiaohan Chen, Yu Cheng, Shuohang Wang, Zhe Gan, …☆18Dec 30, 2021Updated 4 years ago
- Code of "Visualizing and Understanding Object Detecor"☆20Jun 24, 2021Updated 5 years ago
- On Transferability of Prompt Tuning for Natural Language Processing☆98May 3, 2024Updated 2 years ago
- Combining encoder-based language models☆11Nov 11, 2021Updated 4 years ago
- ☆226Feb 21, 2023Updated 3 years ago
- ☆12Jul 7, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆10Jul 5, 2019Updated 7 years ago
- Training and evaluation codes for the BertGen paper (ACL-IJCNLP 2021)☆11Sep 17, 2023Updated 3 years ago
- Minimum viable code for the Decodable Information Bottleneck paper. Pytorch Implementation.☆12Oct 20, 2020Updated 5 years ago
- Official implementation for "Let Offline RL Flow: Training Conservative Agents in the Latent Space of Normalizing Flows", NeurIPS 2022, O…☆12Jan 31, 2023Updated 3 years ago
- [NAACL 2022] TreeMix: Compositional Constituency-based Data Augmentation for Natural Language Understanding☆10Jul 15, 2023Updated 3 years ago
- Code for ACL 2022 paper "Expanding Pretrained Models to Thousands More Languages via Lexicon-based Adaptation"☆29Apr 2, 2022Updated 4 years ago
- Code for EMNLP 2021 paper: Improving Sequence-to-Sequence Pre-training via Sequence Span Rewriting☆17Nov 30, 2021Updated 4 years ago