LV-BERT: Exploiting Layer Variety for BERT (Findings of ACL 2021)
☆19May 10, 2023Updated 3 years ago
Alternatives and similar repositories for LV-BERT
Users that are interested in LV-BERT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A cuda & mkl implementation of closed-form matting☆15Jul 23, 2018Updated 8 years ago
- Official Implementation of CVPR2021 paper: Continual Learning via Bit-Level Information Preserving☆39Jan 24, 2023Updated 3 years ago
- ☆110Sep 15, 2021Updated 5 years ago
- ☆30Apr 6, 2021Updated 5 years ago
- Official Implementation of CVPR 2022 paper: "Mimicking the Oracle: An Initial Phase Decorrelation Approach for Class Incremental Learning…☆35Feb 10, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆21Dec 5, 2022Updated 3 years ago
- [ICLR2023] Towards Understanding and Mitigating Dimensional Collapse in Heterogeneous Federated Learning (https://arxiv.org/abs/2210.0022…☆40Jan 30, 2023Updated 3 years ago
- Temporary remove unused tokens during training to save ram and speed.☆23Jun 15, 2025Updated last year
- Official implementation of "Recovering the Unbiased Scene Graphs from the Biased Ones" (ACMMM 2021)☆78Sep 4, 2022Updated 4 years ago
- MLP-Like Vision Permutator for Visual Recognition (PyTorch)☆192Mar 31, 2022Updated 4 years ago
- ☆141Dec 18, 2021Updated 4 years ago
- Get_dr_feature python api. You can recover mesh by: https://github.com/QianyiWu/get_mesh_py_API☆19Apr 24, 2019Updated 7 years ago
- CIFS: Improving Adversarial Robustness of CNNs via Channel-wise Importance-based Feature Selection☆21Oct 12, 2021Updated 4 years ago
- ☆11Mar 17, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A list of papers regarding generalization in (deep) reinforcement learning☆156Aug 12, 2023Updated 3 years ago
- ☆74Dec 8, 2022Updated 3 years ago
- GPT-4V(ision) as A Social Media Analysis Engine☆39Dec 20, 2024Updated last year
- (ACL-IJCNLP 2021) Convolutions and Self-Attention: Re-interpreting Relative Positions in Pre-trained Language Models.☆21Jul 13, 2022Updated 4 years ago
- awesome unsupervised learning paper list☆12Jan 4, 2018Updated 8 years ago
- A Multilingual Keyboard Layout-Based Typo Generator☆17Nov 23, 2025Updated 10 months ago
- [NAACL 2024] A Framework aims to wisely initialize unseen subword embeddings in PLMs for efficient large-scale continued pretraining☆18Nov 26, 2023Updated 2 years ago
- On Robustness of Neural Ordinary Differential Equations☆11Oct 12, 2021Updated 4 years ago
- ☆11Jan 13, 2019Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official code for the paper "Why Do Self-Supervised Models Transfer? Investigating the Impact of Invariance on Downstream Tasks".☆16Dec 7, 2021Updated 4 years ago
- Code for Policy Consolidation for Continual Reinforcement Learning☆10May 12, 2019Updated 7 years ago
- ☆12Dec 7, 2024Updated last year
- PyTorch implementation of the estimator proposed in the paper "Estimating Differential Entropy under Gaussian Convolutions"☆13Oct 22, 2020Updated 5 years ago
- AOT: Appearance Optimal Transport Based Identity Swapping for Forgery Detection (NeurIPS 2020)☆39Nov 9, 2020Updated 5 years ago
- Rectified Convolution☆45Oct 16, 2022Updated 3 years ago
- Code for SaGe subword tokenizer (EACL 2023)☆28Nov 30, 2024Updated last year
- Our implementation of Shampoo optimizer based on https://arxiv.org/pdf/1802.09568.pdf☆13Dec 23, 2019Updated 6 years ago
- 🚀🤗 A collection of templates for Hugging Face Spaces☆35Oct 9, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Lightweight Transformer for Multi-modal Tasks☆16Dec 9, 2022Updated 3 years ago
- A framework for adversarial attacks against token classification models☆33Nov 6, 2021Updated 4 years ago
- using opencv-python cv2.grabCut to cut image interactively☆10Jun 10, 2018Updated 8 years ago
- A CLIP conditioned Decision Transformer.☆22Jul 14, 2021Updated 5 years ago
- HyPe: Better Pre-trained Language Model Fine-tuning with Hidden Representation Perturbation [ACL 2023]☆14Jul 11, 2023Updated 3 years ago
- Basic implementation of variational autoencoders in Torch☆10Apr 16, 2016Updated 10 years ago
- This is the official PyTorch implementation for "Mesa: A Memory-saving Training Framework for Transformers".☆119Dec 12, 2021Updated 4 years ago