☆54Apr 15, 2022Updated 4 years ago
Alternatives and similar repositories for CUGE
Users that are interested in CUGE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code of the COLING22 paper "uChecker: Masked Pretrained Language Models as Unsupervised Chinese Spelling Checkers"☆19Aug 17, 2022Updated 4 years ago
- [EMNLP 2024 Tutorial] Language Agents: Foundations, Prospects, and Risks☆10Nov 27, 2024Updated last year
- ☆10Sep 27, 2021Updated 4 years ago
- Princeton NLP's pre-training library based on fairseq with DeepSpeed kernel integration 🚃☆117Oct 27, 2022Updated 3 years ago
- A plug-in of Microsoft DeepSpeed to fix the bug of DeepSpeed pipeline☆25Apr 16, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The codebase for "Group-wise Contrastive Learning for Neural Dialogue Generation" (Cai et al., Findings of EMNLP 2020)☆55Feb 24, 2021Updated 5 years ago
- ☆10Oct 17, 2021Updated 4 years ago
- Source code for the EMNLP 2020 long paper <Token-level Adaptive Training for Neural Machine Translation>.☆20Oct 28, 2022Updated 3 years ago
- [ACL 2020] Explicit Memory Tracker with Coarse-to-Fine Reasoning for Conversational Machine Reading☆39Dec 8, 2022Updated 3 years ago
- Code for CPM-2 Pre-Train☆157Mar 18, 2023Updated 3 years ago
- ☆34Mar 22, 2021Updated 5 years ago
- MATCH-TUNING☆15Aug 6, 2022Updated 4 years ago
- [NAACL 2021] Factual Probing Is [MASK]: Learning vs. Learning to Recall https://arxiv.org/abs/2104.05240☆167Oct 7, 2022Updated 3 years ago
- Code and data for COLING 2022 paper titled "Structural Bias For Aspect Sentiment Triplet Extraction"☆26May 28, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Scripts for KGIRNet model for ESWC☆10Jul 6, 2023Updated 3 years ago
- Introduction to CPM☆165Sep 26, 2021Updated 4 years ago
- ☆269Jul 26, 2024Updated 2 years ago
- EVA: Large-scale Pre-trained Chit-Chat Models☆304Mar 11, 2023Updated 3 years ago
- An (incomplete) overview of information extraction☆43Apr 28, 2022Updated 4 years ago
- 《自然语言处理概论》 张奇、桂韬、黄萱菁著☆123Sep 10, 2023Updated 3 years ago
- Data for paper "CC-Riddle: A Question Answering Dataset of Chinese Character Riddles": https://arxiv.org/abs/2206.13778☆20Aug 19, 2023Updated 3 years ago
- Template Filling with Generative Transformers☆22Jun 8, 2021Updated 5 years ago
- ☆98Jun 6, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for embedding and retrieval research.☆16Oct 24, 2023Updated 2 years ago
- ☆45Nov 20, 2025Updated 10 months ago
- Code for our paper Resources and Evaluations for Multi-Distribution Dense Information Retrieval☆17Jan 16, 2024Updated 2 years ago
- Finetune CPM-2☆80Mar 18, 2023Updated 3 years ago
- Implementation of AAAI 21 paper: Nested Named Entity Recognition with Partially Observed TreeCRFs☆50May 11, 2021Updated 5 years ago
- Open-Retrieval Conversational Machine Reading: A new setting & OR-ShARC dataset☆13Nov 19, 2022Updated 3 years ago
- baseline for MGTV competition 2022 PIR☆11Apr 11, 2022Updated 4 years ago
- A plug-and-play library for parameter-efficient-tuning (Delta Tuning)☆1,048Sep 19, 2024Updated 2 years ago
- Code for ACL 2023 paper titled "Lifting the Curse of Capacity Gap in Distilling Language Models"☆29Jul 14, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆224Sep 19, 2023Updated 3 years ago
- [NAACL'22] TaCL: Improving BERT Pre-training with Token-aware Contrastive Learning☆94Jun 8, 2022Updated 4 years ago
- List of Papers on Attack and Defense (AD) in AI Models☆29Mar 18, 2022Updated 4 years ago
- Implementation for paper "A Knowledge-Enhanced Pretraining Model for Commonsense Story Generation"☆102Dec 25, 2019Updated 6 years ago
- Official code of our work, Syntax-augmented Multilingual BERT for Cross-lingual Transfer [ACL 2021].☆16Dec 2, 2021Updated 4 years ago
- Anomaly Detection using SH-ESD☆10Feb 6, 2019Updated 7 years ago
- Chinese Pre-Trained Language Models (CPM-LM) Version-I☆1,579Mar 18, 2023Updated 3 years ago