Parameter Efficient Transfer Learning with Diff Pruning
☆73Feb 3, 2021Updated 5 years ago
Alternatives and similar repositories for DiffPruning
Users that are interested in DiffPruning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for "Training Neural Networks with Fixed Sparse Masks" (NeurIPS 2021).☆59Jan 14, 2022Updated 4 years ago
- [NAACL 2022] "Learning to Win Lottery Tickets in BERT Transfer via Task-agnostic Mask Training", Yuanxin Liu, Fandong Meng, Zheng Lin, Pe…☆15Oct 18, 2022Updated 3 years ago
- Sequence-Level Mixed Sample Data Augmentation☆22Mar 7, 2021Updated 5 years ago
- Code for "Inducer-tuning: Connecting Prefix-tuning and Adapter-tuning" (EMNLP 2022) and "Empowering Parameter-Efficient Transfer Learning…☆11Feb 6, 2023Updated 3 years ago
- Source code for "Importance-based Neuron Allocation for Multilingual Neural Machine Translation"☆12Sep 15, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ACL 2022] Structured Pruning Learns Compact and Accurate Models https://arxiv.org/abs/2204.00408☆198May 9, 2023Updated 3 years ago
- Implementation of paper "Towards a Unified View of Parameter-Efficient Transfer Learning" (ICLR 2022)☆541Mar 24, 2022Updated 4 years ago
- Embedding Recycling for Language models☆38Jul 11, 2023Updated 3 years ago
- Official codebase accompanying our ACL 2022 paper "RELiC: Retrieving Evidence for Literary Claims" (https://relic.cs.umass.edu).☆20May 14, 2022Updated 4 years ago
- ☆26Nov 23, 2023Updated 2 years ago
- ☆33Dec 17, 2025Updated 8 months ago
- ☆146Jul 21, 2024Updated 2 years ago
- ☆131Aug 18, 2022Updated 4 years ago
- ☆49Jan 21, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [Neurips 2022] “ Back Razor: Memory-Efficient Transfer Learning by Self-Sparsified Backpropogation”, Ziyu Jiang*, Xuxi Chen*, Xueqin Huan…☆19Mar 14, 2023Updated 3 years ago
- Implementation of paper: Extending and Analyzing Self-Supervised Learning Across Domains☆10Jan 10, 2021Updated 5 years ago
- ☆31Apr 27, 2022Updated 4 years ago
- On the Effectiveness of Parameter-Efficient Fine-Tuning☆39Nov 4, 2023Updated 2 years ago
- ☆27Dec 15, 2022Updated 3 years ago
- Zero-shot Learning by Generating Task-specific Adapters☆14Apr 2, 2021Updated 5 years ago
- Is BERT Robust to Label Noise? A Study on Learning with Noisy Labels in Text Classification☆10May 31, 2022Updated 4 years ago
- Official codebase for our paper "Joslim: Joint Widths and Weights Optimization for Slimmable Neural Networks"☆12Jun 30, 2021Updated 5 years ago
- ☆15Apr 10, 2018Updated 8 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- MUX-PLMs: Pretraining LMs with Data Multiplexing☆15Jan 29, 2023Updated 3 years ago
- ☆20Dec 16, 2020Updated 5 years ago
- [ACL-IJCNLP 2021] "EarlyBERT: Efficient BERT Training via Early-bird Lottery Tickets" by Xiaohan Chen, Yu Cheng, Shuohang Wang, Zhe Gan, …☆18Dec 30, 2021Updated 4 years ago
- Improving Transformation Invariance in Contrastive Representation Learning☆12Mar 13, 2021Updated 5 years ago
- Code of "Visualizing and Understanding Object Detecor"☆20Jun 24, 2021Updated 5 years ago
- On Transferability of Prompt Tuning for Natural Language Processing☆98May 3, 2024Updated 2 years ago
- Combining encoder-based language models☆11Nov 11, 2021Updated 4 years ago
- ☆226Feb 21, 2023Updated 3 years ago
- ☆27Oct 23, 2018Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- CSS-LM: Contrastive Semi-supervised Fine-tuning of Pre-trained Language Models☆11Jul 1, 2023Updated 3 years ago
- ☆99Jun 4, 2024Updated 2 years ago
- ☆30Jul 22, 2024Updated 2 years ago
- ☆12Jul 7, 2021Updated 5 years ago
- [ACL 2023] Counterspeeches up my sleeve! Intent Distribution Learning and Persistent Fusion for Intent-Conditioned Counterspeech Generati…☆10Sep 23, 2023Updated 2 years ago
- ACL 2021: HiTransformer☆13May 29, 2021Updated 5 years ago
- Training and evaluation codes for the BertGen paper (ACL-IJCNLP 2021)☆11Sep 17, 2023Updated 2 years ago