AnchiBERT: A Pre-Trained Model for Ancient Chinese Language Understanding and Generation(古文预训练模型)
☆75Jul 16, 2021Updated 5 years ago
Alternatives and similar repositories for AnchiBERT
Users that are interested in AnchiBERT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 基于ChineseAlpaca微调的,专精与古汉语翻译、古汉语断句的大语言模型☆20Aug 20, 2023Updated 3 years ago
- Dataset for TALLIP2019 paper "Ancient-Modern Chinese Translation with a New Large Training Dataset"☆27Jul 8, 2022Updated 4 years ago
- GuwenModels: 古文自然语言处理模型合集, 收录互联网上的古文相关模型及资源. A collection of Classical Chinese natural language processing models, including Classical Ch…☆202Dec 11, 2023Updated 2 years ago
- 甲言,专注于古代汉语(古汉语/古文/文言文/文言)处理的NLP工具包,支持文言词库构建、分词、词性标注、断句和标点。Jiayan, the 1st NLP toolkit designed for Classical Chinese, supports lexicon co…☆678Nov 2, 2021Updated 4 years ago
- 本仓库是基于bert4keras实现的古文-现代文翻译模型。具体使用了基于掩码自注意力机制的UNILM(Li al., 2019)预训练模型作为翻译系统的backbone。我们首先使用了普通的中文(现代文)BERT、Roberta权重作为UNILM的初始权重以训练UNILM…☆54May 3, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Pretrained BERT for Ancient (Classical) Chinese, with an expanded vocabulary for rare characters.☆48Feb 20, 2023Updated 3 years ago
- 非常全的文言文(古文)-现代文平行语料☆1,488Apr 21, 2024Updated 2 years ago
- SikuBERT:四库全书的预训练语言模型(四库BERT) Pre-training Model of Siku Quanshu☆171Jul 30, 2023Updated 3 years ago
- ☆21Oct 6, 2023Updated 2 years ago
- 文言文翻译、古文翻译 语料数据 集☆55Oct 14, 2020Updated 5 years ago
- Ancient Chinese Corpus with Word Sense Annotation☆77May 29, 2024Updated 2 years ago
- Chinese Machine Reading 2021海华AI挑战赛·中文阅读理解·技术组·第三名☆20May 27, 2021Updated 5 years ago
- ☆23Jun 2, 2019Updated 7 years ago
- 開放漢語字典 - 現代漢語字音數據庫☆29Oct 31, 2020Updated 5 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Serves aggregated news from 13 local news publishers in Hong Kong☆11Jun 26, 2022Updated 4 years ago
- Implemention of NER model on chinese dataset.☆76Apr 8, 2023Updated 3 years ago
- Contextualised Word Representations for Lexical Semantic Change Analysis☆33Jul 17, 2020Updated 6 years ago
- Code repository for MMUGL: Multi-modal Graph Learning over UMLS Knowledge Graphs☆11Dec 7, 2023Updated 2 years ago
- 使用biaffine的中文命名实体识别☆10Jan 12, 2023Updated 3 years ago
- The implementation of RAGSynth: Synthetic Data for Robust and Faithful RAG Component Optimization☆21May 26, 2025Updated last year
- 文言文命名实体识别,基于BILSTM+CRF完成文言文的命名实体实体,识别实体包括人物、地点、机构、时间等。☆10Jan 19, 2021Updated 5 years ago
- ☆14Mar 7, 2022Updated 4 years ago
- A Benchmark for Classical Chinese Based on a Crowdsourcing System.☆60May 25, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- chinese few-shot ner☆16Aug 28, 2022Updated 4 years ago
- The current version of Data by Design, an interactive history of data visualization☆17Updated this week
- ☆11Jun 28, 2023Updated 3 years ago
- Neural ngram language model in PyTorch.☆10Sep 27, 2018Updated 8 years ago
- Evaluation metrics package☆10May 21, 2025Updated last year
- Global Greedy Dependency Parsing☆10Mar 16, 2021Updated 5 years ago
- 古汉语(文言文)字典-爬取文言文字典网,制作Kindle字典.☆70Jun 21, 2018Updated 8 years ago
- ☆12Aug 16, 2018Updated 8 years ago
- ☆23Jan 25, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- This repository is intended for people who are interesting in learning/reading Classical Chinese (文言文) but be at a loss what to do to. As…☆14Jul 1, 2021Updated 5 years ago
- Code and Dataset for Learning to Solve Complex Tasks by Talking to Agents☆24May 24, 2022Updated 4 years ago
- Playing around with an inverted index☆14Oct 29, 2015Updated 10 years ago
- [EMNLP 2018] Training for Diversity in Image Paragraph Captioning☆91Sep 12, 2019Updated 7 years ago
- ChineseDiachronicCorpus,中文历时语料库,横跨六十余年,包括腾讯历时新闻2000-2016,人民日报历时语料1946-2003,参考消息历时语料1957-2002。基于历时流通语料库,可用于历时语言变化计算、语言监测、社会文化变迁研究提供基础性的语料支…☆25Jan 10, 2021Updated 5 years ago
- Revisiting Cross-Lingual Summarization: A Corpus-based Study and A New Benchmark with Improved Annotation☆19Mar 23, 2024Updated 2 years ago
- ☆14Jan 6, 2025Updated last year