Workshop on Language Technologies for Historical and Ancient Languages (LT4HALA)
☆38May 19, 2026Updated 3 months ago
Alternatives and similar repositories for LT4HALA
Users that are interested in LT4HALA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pretrained BERT for Ancient (Classical) Chinese, with an expanded vocabulary for rare characters.☆48Feb 20, 2023Updated 3 years ago
- The official GitHub repository for AC-EVAL, an ancient Chinese evaluation suite for large language models (LLMs)☆17Nov 12, 2024Updated last year
- SikuBERT:四库全书的预训练语言模型(四库BERT) Pre-training Model of Siku Quanshu☆169Jul 30, 2023Updated 3 years ago
- CCL 2023 古汉语通假字语料库的构建及应用研究:通假字资源库☆33Sep 23, 2023Updated 2 years ago
- Tokenizer POS-tagger and Dependency-parser for Classical Chinese☆75Jun 10, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 古文语言理解测评基准 Classical Chinese Language Understanding Evaluation Benchmark: datasets, baselines, pre-trained models, corpus and leaderboard☆58Aug 23, 2023Updated 3 years ago
- ☆19Oct 6, 2023Updated 2 years ago
- A dataset used for NLP tasks.☆10Apr 17, 2021Updated 5 years ago
- A bunch of modules that use/extend CLTK in order to work with Greek and Latin corpora maintained by the Perseus DL☆12Oct 26, 2019Updated 6 years ago
- <数字人文教程>资源合集☆122May 28, 2024Updated 2 years ago
- Ancient Chinese Corpus with Word Sense Annotation☆77May 29, 2024Updated 2 years ago
- 文言文信息抽取(实体识别+关系抽取)☆10Feb 24, 2023Updated 3 years ago
- Tokenizer POS-tagger and Dependency-parser for Classical Chinese☆15Dec 30, 2025Updated 8 months ago
- GuwenBERT: 古文预训练语言模型(古文BERT) A Pre-trained Language Model for Classical Chinese (Literary Chinese)☆567Aug 31, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- a Corpus for Classical Chinese Language Event Extraction☆25Nov 11, 2025Updated 9 months ago
- An evaluation bentchmark for classical Chinese☆21Dec 13, 2023Updated 2 years ago
- ☆32May 6, 2026Updated 3 months ago
- A general-purpose NLP pipeline for Ancient Greek☆29Mar 26, 2024Updated 2 years ago
- An Ancient Greek Morphology Tagger☆28May 9, 2023Updated 3 years ago
- 渊 - A project for Classical Chinese☆112Feb 23, 2022Updated 4 years ago
- 🎉 Repo for Ancient-Agri-LLM.古农文大语言模型☆10Sep 13, 2024Updated last year
- 基于ChineseAlpaca微调的,专精与古汉语翻译、古汉语断句的大语言模型☆20Aug 20, 2023Updated 3 years ago
- Python 3 tool for generating (initially Biblical) Greek readers☆35Oct 12, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Dataset for TALLIP2019 paper "Ancient-Modern Chinese Translation with a New Large Training Dataset"☆27Jul 8, 2022Updated 4 years ago
- Evaluation of Natural Language Processing (NLP) tools for the Ancient Chinese language☆48Mar 15, 2026Updated 5 months ago
- Machine-corrected versions of selections of the Patrologia Latina.☆27Apr 2, 2019Updated 7 years ago
- A PyTorch implementation of a BiLSTM \ BERT \ Roberta (+ BiLSTM + CRF) model for Chinese Word Segmentation (中文分词) .☆216Jul 28, 2022Updated 4 years ago
- Using CRF++ for NER☆20Feb 28, 2019Updated 7 years ago
- 汉语古典文本资料库☆363Feb 3, 2018Updated 8 years ago
- Research Environment for Ancient Documents☆46Jan 24, 2026Updated 7 months ago
- An implementation of data augmentation methods for natural language processing tasks.☆13Jul 25, 2024Updated 2 years ago
- 斗破苍穹小说的新词发现☆13May 12, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Preprocessing scripts for ACE and ERE datasets☆15Jul 28, 2020Updated 6 years ago
- 粤港澳大湾区(黄埔)国际算法算例大赛-古籍文档图像识别与分析算法比赛 Alphx队源码☆46Mar 16, 2023Updated 3 years ago
- 古文现代文翻译平行语料库☆118Jan 12, 2022Updated 4 years ago
- ☆44Sep 26, 2021Updated 4 years ago
- Coupling Distant Annotation and Adversarial Training for Cross-Domain Chinese Word Segmentation☆23Sep 18, 2020Updated 5 years ago
- Official github repo for ACLUE, an evaluation benchmark focused on ancient Chinese language comprehension☆35Mar 20, 2024Updated 2 years ago
- Accelerate pretraining by pre-pretraining on formal languages!☆22Feb 13, 2026Updated 6 months ago