Finetune CPM-2
☆80Mar 18, 2023Updated 3 years ago
Alternatives and similar repositories for CPM-2-Finetune
Users that are interested in CPM-2-Finetune are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for CPM-2 Pre-Train☆157Mar 18, 2023Updated 3 years ago
- Introduction to CPM☆166Sep 26, 2021Updated 5 years ago
- ☆36Jan 5, 2021Updated 5 years ago
- A plug-in of Microsoft DeepSpeed to fix the bug of DeepSpeed pipeline☆25Apr 16, 2021Updated 5 years ago
- Efficient Inference for Big Models☆585Jul 7, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Finetune CPM-1☆73Mar 18, 2023Updated 3 years ago
- Introduction to CPM☆17Jun 22, 2021Updated 5 years ago
- Finetune CPM-1☆24Jun 20, 2021Updated 5 years ago
- Chinese Pre-Trained Language Models (CPM-LM) Version-I☆1,579Mar 18, 2023Updated 3 years ago
- ☆54Apr 15, 2022Updated 4 years ago
- Finetune CPM-1 For Text Generation☆18Jul 9, 2021Updated 5 years ago
- Pretrain CPM-1☆53Apr 20, 2021Updated 5 years ago
- The codes and data for paper "Learning to Control the Fine-grained Sentiment for Story Ending Generation (ACL 2019)".☆25Sep 26, 2019Updated 7 years ago
- Inference framework for MoE layers based on TensorRT with Python binding☆40May 31, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Repository for the ACL'22 paper "So Different Yet So Alike! Constrained Unsupervised Text Style Transfer"☆16Jan 19, 2024Updated 2 years ago
- [Findings of ACL 2023] Communication Efficient Federated Learning for Multilingual Machine Translation with Adapter☆12Sep 4, 2023Updated 3 years ago
- Easy-to-use CPM for Chinese text generation(基于CPM的中文文本生成)☆530Apr 10, 2023Updated 3 years ago
- ☆34Jul 29, 2021Updated 5 years ago
- ☆23Mar 19, 2021Updated 5 years ago
- Algorithmic and AI MIDI Drums Generator Implementation☆13Apr 1, 2022Updated 4 years ago
- preprocessing of the MUC4 dataset☆11Aug 28, 2012Updated 14 years ago
- ☆220Dec 8, 2022Updated 3 years ago
- Source code for our EMNLP19 paper "Retrieval-guided Dialogue Response Generation via a Matching-to-Generation Framework"☆37Mar 20, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- using lear to do ner extraction☆28Mar 13, 2022Updated 4 years ago
- Code for CAET5☆23Jun 12, 2023Updated 3 years ago
- The code for ``STYLEDGPT: Stylized Response Generation with Pre-trained LanguageModels'' (Findings of EMNLP2020)☆21Nov 16, 2020Updated 5 years ago
- Visualisation & processing of MEG data and networks on meshes. A bunch of functions, examples and a gui.☆11Oct 11, 2019Updated 6 years ago
- The repo of "Improving Seq2Seq Grammatical Error Correction via Decoding Interventions"☆32Jan 22, 2024Updated 2 years ago
- [ACL 2023] Are Pre-trained Language Models Useful for Model Ensemble in Chinese Grammatical Error Correction?☆10Dec 15, 2025Updated 9 months ago
- ☆11Nov 1, 2017Updated 8 years ago
- ☆13Jun 21, 2021Updated 5 years ago
- ☆246Oct 21, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Implementation of Paper Deep Coupled ISTA Network for Multi-modal Image Super-Resolution☆11Sep 17, 2019Updated 7 years ago
- Implementation of Semantic Parsing with BERT and compositional pre-training on GeoQuery☆11Mar 20, 2019Updated 7 years ago
- [ACL 2021] LM-BFF: Better Few-shot Fine-tuning of Language Models https://arxiv.org/abs/2012.15723☆727Aug 29, 2022Updated 4 years ago
- Mike X Cohen lecturelets on Analyzing Neural Time Series Data: Theory and Practice http://mikexcohen.com/lectures.html☆15Dec 30, 2020Updated 5 years ago
- A hierarchical Bayesian model accounting for endmember variability and abrupt spectral changes to unmix multitemporal hyperspectral image…☆10Nov 20, 2020Updated 5 years ago
- ☆11Nov 13, 2020Updated 5 years ago
- CPT: A Pre-Trained Unbalanced Transformer for Both Chinese Language Understanding and Generation☆493Dec 30, 2022Updated 3 years ago