A simple recipe for training and inferencing Transformer architecture for Multi-Task Learning on custom datasets. You can find two approaches for achieving this in this repo.
☆102Jul 14, 2022Updated 4 years ago
Alternatives and similar repositories for multitask-learning-transformers
Users that are interested in multitask-learning-transformers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A simple project training 3 separate NLP tasks simultaneously using Multitask-Learning☆24Jun 12, 2023Updated 3 years ago
- Multi-task modelling extensions for huggingface transformers☆14Jul 8, 2025Updated last year
- Easy modernBERT fine-tuning and multi-task learning☆66Mar 13, 2026Updated 5 months ago
- ☆15Oct 19, 2020Updated 5 years ago
- BERT for Multitask Learning☆544Apr 12, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆12Jun 6, 2020Updated 6 years ago
- 📄 Evidence Retrieval and Claim Verification for the FEVER shared task using Transformer Networks☆12Feb 21, 2020Updated 6 years ago
- ☆10Jul 27, 2018Updated 8 years ago
- Task Compass: Scaling Multi-task Pre-training with Task Prefix (EMNLP 2022: Findings) (stay tuned & more will be updated)☆22Oct 17, 2022Updated 3 years ago
- Data and code for "Understanding Linearity of Cross-Lingual Word Embedding Mappings" (TMLR 2022)☆12Jun 8, 2022Updated 4 years ago
- Arabic News Stance Corpus☆11Feb 5, 2021Updated 5 years ago
- ☆11Aug 19, 2024Updated last year
- Token-free Language Modeling with ByGPT5 & Friends!☆12Jul 18, 2025Updated last year
- Framework for unified summarisation and evaluation of English documents using state-of-the-art models and measures.☆33May 13, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- COMIC: This is the code repo of our TMM2019 work titled "COMIC: Towards a Compact Image Captioning Model with Attention".☆15Jun 22, 2021Updated 5 years ago
- Course for Interpreting ML Models☆51Feb 16, 2023Updated 3 years ago
- MFAQ: a Multilingual FAQ Dataset☆18Aug 5, 2026Updated last week
- A neural text style transfer model☆12Jun 23, 2019Updated 7 years ago
- Comparing PyTorch, JIT and ONNX for inference with Transformers☆19Feb 22, 2021Updated 5 years ago
- Locally Linear Embedding for Regression - Journal of Chemometrics 2015☆11Aug 5, 2015Updated 11 years ago
- The NLPStatTest project☆12Mar 12, 2022Updated 4 years ago
- ☆64Nov 27, 2022Updated 3 years ago
- Multitask-learning of a BERT backbone. Allows to easily train a BERT model with state-of-the-art method such as PCGrad, Gradient Vaccine,…☆20Oct 8, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Abstractive and Extractive Text summarization using Transformers.☆86Jun 9, 2023Updated 3 years ago
- ☆16Dec 25, 2021Updated 4 years ago
- GERNERMED++ is a transfer-learning-based open neural NER model for medical entities designed for German data.☆10Oct 20, 2023Updated 2 years ago
- 🦮 Code and pretrained models for Findings of ACL 2022 paper "LaPraDoR: Unsupervised Pretrained Dense Retriever for Zero-Shot Text Retrie…☆49Apr 25, 2022Updated 4 years ago
- Fine-tuned BERT on SQuAd 2.0 Dataset. Applied Knowledge Distillation (KD) and fine-tuned DistilBERT (student) using BERT as the teacher m…☆26Feb 13, 2021Updated 5 years ago
- Code for various active transfer learning projects.☆10Feb 9, 2015Updated 11 years ago
- NERC-fr: Supervised Named Entity Recognition for French☆13Jul 10, 2015Updated 11 years ago
- multi_task_NLP is a utility toolkit enabling NLP developers to easily train and infer a single model for multiple tasks.☆375Nov 21, 2022Updated 3 years ago
- Multi-task feature learning☆15Jun 17, 2018Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- GPTNERMED is a language model-generated, synthetic dataset and an open neural NER model for medical entities designed for German data.☆15Oct 5, 2023Updated 2 years ago
- This is the Grammarly's Yahoo Answers Formality Corpus☆108Jul 7, 2025Updated last year
- XtremeDistil framework for distilling/compressing massive multilingual neural network models to tiny and efficient models for AI at scale☆157Dec 20, 2023Updated 2 years ago
- Code for gradient rollback, which explains predictions of neural matrix factorization models, as for example used for knowledge base comp…☆21Mar 16, 2021Updated 5 years ago
- Code for the paper: Saying No is An Art: Contextualized Fallback Responses for Unanswerable Dialogue Queries☆19Nov 29, 2021Updated 4 years ago
- A package for fine tuning of pretrained NLP transformers using Semi Supervised Learning☆14Oct 27, 2021Updated 4 years ago
- jiant is an nlp toolkit☆1,674Jul 6, 2023Updated 3 years ago