Finetune CPM-2
☆80Mar 18, 2023Updated 3 years ago
Alternatives and similar repositories for CPM-2-Finetune
Users that are interested in CPM-2-Finetune are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for CPM-2 Pre-Train☆157Mar 18, 2023Updated 3 years ago
- Introduction to CPM☆164Sep 26, 2021Updated 4 years ago
- ☆36Jan 5, 2021Updated 5 years ago
- A plug-in of Microsoft DeepSpeed to fix the bug of DeepSpeed pipeline☆25Apr 16, 2021Updated 5 years ago
- Efficient Inference for Big Models☆583Jul 7, 2026Updated last month
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Finetune CPM-1☆73Mar 18, 2023Updated 3 years ago
- Introduction to CPM☆17Jun 22, 2021Updated 5 years ago
- Finetune CPM-1☆24Jun 20, 2021Updated 5 years ago
- Clues Before Answers: Generation-Enhanced Multiple-Choice QA (NAACL 2022)☆28Oct 3, 2023Updated 2 years ago
- ☆54Apr 15, 2022Updated 4 years ago
- Finetune CPM-1 For Text Generation☆18Jul 9, 2021Updated 5 years ago
- Pretrain CPM-1☆53Apr 20, 2021Updated 5 years ago
- Inference framework for MoE layers based on TensorRT with Python binding☆40May 31, 2021Updated 5 years ago
- Repository for the ACL'22 paper "So Different Yet So Alike! Constrained Unsupervised Text Style Transfer"☆16Jan 19, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- In this work, we implement different cross-modal learning schemes such as Siamese Network, Correlational Network and Deep Cross-Modal Pro…☆11Aug 23, 2021Updated 4 years ago
- [Findings of ACL 2023] Communication Efficient Federated Learning for Multilingual Machine Translation with Adapter☆12Sep 4, 2023Updated 2 years ago
- Easy-to-use CPM for Chinese text generation(基于CPM的中文文本生成)☆530Apr 10, 2023Updated 3 years ago
- ☆34Jul 29, 2021Updated 5 years ago
- ☆22Mar 19, 2021Updated 5 years ago
- Vision Large Language Models trained on M3IT instruction tuning dataset☆17Aug 16, 2023Updated 3 years ago
- Implementation of Seq2SQL. Inspired by: https://github.com/xiaojunxu/SQLNet☆18Sep 18, 2018Updated 7 years ago
- MASKER: Masked Keyword Regularization for Reliable Text Classification (AAAI 2021)☆54Oct 4, 2023Updated 2 years ago
- ☆219Dec 8, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Source code for our EMNLP19 paper "Retrieval-guided Dialogue Response Generation via a Matching-to-Generation Framework"☆37Mar 20, 2020Updated 6 years ago
- Code for CAET5☆23Jun 12, 2023Updated 3 years ago
- The code for ``STYLEDGPT: Stylized Response Generation with Pre-trained LanguageModels'' (Findings of EMNLP2020)☆21Nov 16, 2020Updated 5 years ago
- investigating use of variational auto encoders with multinomial latent variables for unsupervised data.☆25Jun 12, 2017Updated 9 years ago
- The repo of "Improving Seq2Seq Grammatical Error Correction via Decoding Interventions"☆32Jan 22, 2024Updated 2 years ago
- [ACL 2023] Are Pre-trained Language Models Useful for Model Ensemble in Chinese Grammatical Error Correction?☆10Dec 15, 2025Updated 8 months ago
- 一个基于内容的图像检索系统☆14Aug 19, 2022Updated 3 years ago
- Latex Workshop (2024 Spring)☆11Oct 20, 2024Updated last year
- ☆11Nov 1, 2017Updated 8 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆13Jun 21, 2021Updated 5 years ago
- 2023新春版:React+Antd开发Chrome插件教程☆14Feb 6, 2023Updated 3 years ago
- Mike X Cohen lecturelets on Analyzing Neural Time Series Data: Theory and Practice http://mikexcohen.com/lectures.html☆14Dec 30, 2020Updated 5 years ago
- A hierarchical Bayesian model accounting for endmember variability and abrupt spectral changes to unmix multitemporal hyperspectral image…☆10Nov 20, 2020Updated 5 years ago
- ☆11Nov 13, 2020Updated 5 years ago
- CPT: A Pre-Trained Unbalanced Transformer for Both Chinese Language Understanding and Generation☆494Dec 30, 2022Updated 3 years ago
- ☆10Feb 16, 2025Updated last year