Implementation for Variational Information Bottleneck for Effective Low-resource Fine-tuning, ICLR 2021
☆44May 10, 2021Updated 5 years ago
Alternatives and similar repositories for vibert
Users that are interested in vibert are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deep Variational Information Bottleneck (DVIB) in PyTorch.☆10Apr 25, 2020Updated 6 years ago
- Compressing Neural Networks using the Variational Information Bottleneck☆66Aug 3, 2022Updated 4 years ago
- Code for "Towards Robust k-Nearest-Neighbor Machine Translation" (EMNLP 2022)☆12Oct 18, 2022Updated 3 years ago
- an implementation of Deep Variational Informational Bottleneck in pytorch (https://arxiv.org/pdf/1612.00410.pdf)☆31Apr 26, 2018Updated 8 years ago
- Code for ACL2022 publication Transkimmer: Transformer Learns to Layer-wise Skim☆22Aug 21, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for TACL 2020 paper "An Empirical Study on Robustness to Spurious Correlations using Pre-trained Language Models"☆14Jul 31, 2020Updated 6 years ago
- Running massive simulations using RNNs on CPUs for building bots and all kinds of things.☆12Jun 13, 2021Updated 5 years ago
- Codebase for DualEnc (ACL-20)☆21Oct 3, 2023Updated 2 years ago
- ☆34Jan 8, 2021Updated 5 years ago
- The repository for our EMNLP'20 paper SeqMix: Augmenting Active Sequence Labeling via Sequence Mixup.☆43Sep 5, 2021Updated 5 years ago
- Pytorch implementation of Deep Variational Information Bottleneck☆213Mar 22, 2018Updated 8 years ago
- Code for the paper: Learning Adversarially Robust Representations via Worst-Case Mutual Information Maximization (https://arxiv.org/abs/2…☆23Nov 23, 2020Updated 5 years ago
- Implementation of ICLR 2020 paper "Revisiting Self-Training for Neural Sequence Generation"☆46Jun 30, 2022Updated 4 years ago
- ☆160Aug 24, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆13Mar 22, 2023Updated 3 years ago
- ☆19Sep 19, 2022Updated 4 years ago
- AG2E: A novel adaptive graph based multi-label learning framework for multi-label annotation, image retrieval, and other applications.☆22Oct 15, 2019Updated 6 years ago
- ☆131Aug 18, 2022Updated 4 years ago
- Open-source Human Feedback Library☆11Oct 25, 2023Updated 2 years ago
- Code for paper "Nearest Neighbor Knowledge Distillation for Neural Machine Translation" by Zhixian Yang, Renliang Sun, and Xiaojun Wan. T…☆32Jul 16, 2022Updated 4 years ago
- ☆14Oct 15, 2022Updated 3 years ago
- Pre-processing text and tokenization for UTH-BERT☆10Sep 30, 2020Updated 5 years ago
- The 4th rank system of the SemEval 2021 Task4.☆10May 7, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A Variational Information Bottleneck Approach to Multi-Omics Data Integration☆24May 11, 2021Updated 5 years ago
- ☆11Mar 19, 2023Updated 3 years ago
- Code for ACL 2022 paper "BERT Learns to Teach: Knowledge Distillation with Meta Learning".☆86Aug 4, 2022Updated 4 years ago
- The Python crash course of the Summer Institute in Computational Social Science 2022!☆10Nov 19, 2022Updated 3 years ago
- ☆19Feb 1, 2021Updated 5 years ago
- EMNLP 2020: Filtering before Iteratively Referring for Knowledge-Grounded Response Selection in Retrieval-Based Chatbots☆12Dec 15, 2020Updated 5 years ago
- ☆23Sep 3, 2020Updated 6 years ago
- Theory and PyTorch implementation of Deep Variational Information Bottleneck☆35Jul 17, 2020Updated 6 years ago
- Project for SNARE benchmark☆11Jun 5, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Code for testing DCT plus Sparse (DCTpS) networks☆14Jun 15, 2021Updated 5 years ago
- Code for paper "Can contrastive learning avoid shortcut solutions?" NeurIPS 2021.☆47Mar 29, 2022Updated 4 years ago
- Code for the EMNLP 2021 Oral paper "Are Gender-Neutral Queries Really Gender-Neutral? Mitigating Gender Bias in Image Search" https://arx…☆12Feb 6, 2023Updated 3 years ago
- Spectral Graph Attention Network with Fast Eigen-approximation☆11Dec 24, 2021Updated 4 years ago
- ☆21Mar 7, 2024Updated 2 years ago
- ☆26Jul 18, 2019Updated 7 years ago
- In this paper, we propose Filter Gradient Decent (FGD), an efficient stochastic optimization algorithm that makes a consistent estimation…☆12May 18, 2021Updated 5 years ago