Multi-Task Deep Neural Networks for Natural Language Understanding
☆167Jun 12, 2023Updated 3 years ago
Alternatives and similar repositories for MT-DNN
Users that are interested in MT-DNN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-Task Deep Neural Networks for Natural Language Understanding☆2,257Mar 7, 2024Updated 2 years ago
- Virtual Adversarial Training (VAT) techniques in PyTorch☆17Jul 19, 2022Updated 4 years ago
- ☆45Oct 14, 2021Updated 4 years ago
- modification of official bert for downstream task☆32Mar 24, 2023Updated 3 years ago
- Code associated with the Don't Stop Pretraining ACL 2020 paper☆544Nov 15, 2021Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆69Oct 27, 2020Updated 5 years ago
- Lite Self-Training☆30Jul 25, 2023Updated 3 years ago
- XtremeDistil framework for distilling/compressing massive multilingual neural network models to tiny and efficient models for AI at scale☆157Dec 20, 2023Updated 2 years ago
- ☆31Jun 28, 2022Updated 4 years ago
- XTREME is a benchmark for the evaluation of the cross-lingual generalization ability of pre-trained multilingual models that covers 40 ty…☆651Jan 4, 2023Updated 3 years ago
- Boolean Question Answering with multi-task learning and uses large LM embeddings like BERT, RoBERTa☆18Aug 30, 2019Updated 7 years ago
- Re-rank n-best lists using additional features.☆29Jun 5, 2018Updated 8 years ago
- ☆39Jul 25, 2024Updated 2 years ago
- WMG agent☆35Oct 3, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Fixed-point scalar and matrix multiplication library for SectorLISP☆15Jan 23, 2022Updated 4 years ago
- ACL22 paper: Imputing Out-of-Vocabulary Embeddings with LOVE Makes Language Models Robust with Little Cost☆41Nov 15, 2023Updated 2 years ago
- Code to create pre-training data for a span selection pre-training task inspired by reading comprehension and an effort to avoid encoding…☆30May 20, 2022Updated 4 years ago
- The code of EMNLP 2019 paper "A Split-and-Recombine Approach for Follow-up Query Analysis"☆18Jul 20, 2023Updated 3 years ago
- Code and Data for ACL 2020 paper "Few-Shot NLG with Pre-Trained Language Model"☆188May 23, 2025Updated last year
- ☆16Feb 20, 2023Updated 3 years ago
- PyTorch Implementation of Zero-shot User Intent Detection via Capsule Neural Networks☆18Apr 3, 2019Updated 7 years ago
- Matching Natural Language Sentences with Hierarchical Sentence Factorization☆22Apr 26, 2018Updated 8 years ago
- BERT for Multitask Learning☆544Apr 12, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Pytorch implementation of "A Probabilistic Formulation of Unsupervised Text Style Transfer" by He. et. al. at ICLR 2020☆161Oct 19, 2022Updated 3 years ago
- Code and resources for papers "Generation-Augmented Retrieval for Open-Domain Question Answering" and "Reader-Guided Passage Reranking fo…☆74Feb 11, 2022Updated 4 years ago
- Implementation for our TOIS paper --- Attentive Long Short-Term Preference Modeling for Personalized Product Search.☆18Feb 14, 2020Updated 6 years ago
- Multiple Different Natural Language Processing Tasks in a Single Deep Model☆48Dec 5, 2018Updated 7 years ago
- The Tweets2013 Internet Archive collection☆10Aug 7, 2020Updated 6 years ago
- KLUE Benchmark 1st place (2021.12) solutions. (RE, MRC, NLI, STS, TC)☆25Apr 11, 2022Updated 4 years ago
- ABCD: A Graph Framework to Convert Complex Sentences to a Covering Set of Simple Sentences☆28May 9, 2023Updated 3 years ago
- PyTorch ObjectDetection Modules and ONNX ops☆18Jun 12, 2023Updated 3 years ago
- A Translation Task using TurboTransformers☆10Dec 17, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [EMNLP 2021] SimCSE: Simple Contrastive Learning of Sentence Embeddings https://arxiv.org/abs/2104.08821☆3,652Oct 16, 2024Updated last year
- The score code of FastBERT (ACL2020)☆606Oct 29, 2021Updated 4 years ago
- MPNet: Masked and Permuted Pre-training for Language Understanding https://arxiv.org/pdf/2004.09297.pdf☆300Sep 11, 2021Updated 5 years ago
- LGEB: Benchmark of Language Generation Evaluation☆16Oct 21, 2022Updated 3 years ago
- BERT-related papers☆2,031Aug 12, 2023Updated 3 years ago
- 🌊HMTL: Hierarchical Multi-Task Learning - A State-of-the-Art neural network model for several NLP tasks based on PyTorch and AllenNLP☆1,196Aug 1, 2023Updated 3 years ago
- EMNLP 2021 - Pre-training architectures for dense retrieval☆256Mar 18, 2022Updated 4 years ago