[EMNLP 2022] Continual Training of Language Models for Few-Shot Learning
☆44Feb 13, 2023Updated 3 years ago
Alternatives and similar repositories for CPT
Users that are interested in CPT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2022] Adapting a Language Model While Preserving its General Knowledge☆21Feb 12, 2023Updated 3 years ago
- PyContinual (An Easy and Extendible Framework for Continual Learning)☆323Jan 29, 2024Updated 2 years ago
- An Extensible Continual Learning Framework Focused on Language Models (LMs)☆295Jan 28, 2024Updated 2 years ago
- Code for ECML-PKDD 2022 Paper --- CMG: A Class-Mixed Generation Approach to Out-of-Distribution Detection☆13Oct 12, 2022Updated 3 years ago
- Official repo for "ProSec: Fortifying Code LLMs with Proactive Security Alignment"☆19Feb 26, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 👌[ICLR 2025] TFG-Flow: Training-free Guidance in Multimodal Generative Flow☆20Mar 4, 2025Updated last year
- VQ-VAE implementation pytorch☆11Mar 15, 2023Updated 3 years ago
- A extension of Transformers library to include T5ForSequenceClassification class.☆40Apr 17, 2023Updated 3 years ago
- [ICLR26] AI-based scaling law discovery☆36Jan 30, 2026Updated 7 months ago
- ☆14Nov 4, 2022Updated 3 years ago
- ✌[ICLR 2024] Class Incremental Learning via Likelihood Ratio Based Task Prediction☆31Oct 29, 2024Updated last year
- ☆11Jul 26, 2023Updated 3 years ago
- [ICLR 2022] Towards Continual Knowledge Learning of Language Models☆91Oct 11, 2022Updated 3 years ago
- ☆45Sep 12, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [SIGIR24] Pre-training with Bag-of-Word Prediction for Dense Passage Retrieval☆18Feb 29, 2024Updated 2 years ago
- Code for ECML/PKDD 2020 Paper --- Continual Learning with Knowledge Transfer for Sentiment Classification☆25Nov 19, 2021Updated 4 years ago
- ☆17Nov 3, 2022Updated 3 years ago
- ☆12Jul 6, 2023Updated 3 years ago
- The official implementation of the paper "Text Classification in the Wild: a Large-scale Long-tailed Name Normalization Dataset"(ICASSP 2…☆12Feb 19, 2023Updated 3 years ago
- A Universal Discriminator for Zero-Shot Generalization☆18Jun 21, 2023Updated 3 years ago
- Topic Model based on Pretrained Sentence Embeddings (with BERT)☆13Feb 8, 2023Updated 3 years ago
- Repo for "Zemi: Learning Zero-Shot Semi-Parametric Language Models from Multiple Tasks" ACL 2023 Findings☆15May 3, 2023Updated 3 years ago
- ☆20Jun 3, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- PyTorch reimplementation of the paper "SimCLS: A Simple Framework for Contrastive Learning of Abstractive Summarization"☆16Oct 17, 2021Updated 4 years ago
- Code and data for "Inferring Rewards from Language in Context" [ACL 2022].☆16May 22, 2022Updated 4 years ago
- ☆10Oct 15, 2020Updated 5 years ago
- Code for co-training large language models (e.g. T0) with smaller ones (e.g. BERT) to boost few-shot performance☆16Sep 23, 2022Updated 3 years ago
- Code associated with the paper: "Few-Shot Self-Rationalization with Natural Language Prompts"☆12Apr 27, 2022Updated 4 years ago
- 基于中心度的中文关键短语抽取工具☆11Sep 2, 2022Updated 4 years ago
- Serial Contrastive Knowledge Distillation for Continual Few-shot Relation Extraction, Findings of ACL 2023☆14May 12, 2023Updated 3 years ago
- Auxiliary tasks for task-oriented dialogue systems. Published in ICNLSP'22 and indexed in the ACL Anthology.☆17Feb 27, 2023Updated 3 years ago
- ☆19Nov 17, 2019Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- DQ-BART: Efficient Sequence-to-Sequence Model via Joint Distillation and Quantization (ACL 2022)☆49Jun 12, 2023Updated 3 years ago
- Codes for the paper: "Continual Learning for Text Classification with Information Disentanglement Based Regularization"☆45Feb 9, 2023Updated 3 years ago
- “Generate to Understand for Representation”☆14Apr 18, 2024Updated 2 years ago
- ☆59Oct 21, 2025Updated 10 months ago
- Survey on machine learning.☆14Nov 28, 2020Updated 5 years ago
- Adding new tasks to T0 without catastrophic forgetting☆33Oct 20, 2022Updated 3 years ago
- incremental symbol learning for natural language understanding☆10Jun 12, 2023Updated 3 years ago