The spoken L1 corpus represents present-day spoken Chinese (Putonghua) used in mainland China, which is designed as a comparable corpus to the spoken L2 corpus. It comprises L1-L1 conversational interactions between L1 speakers of Chinese and a native Chinese speaker in informal settings. This corpus contains 228,306 words of transcribed intera…
☆23Aug 2, 2021Updated 4 years ago
Alternatives and similar repositories for The-spoken-L1-corpus
Users that are interested in The-spoken-L1-corpus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A simple desktop app development framework combining Python, Vue.js, Element Plus and Electron.☆11Feb 9, 2023Updated 3 years ago
- Providing a reactivity system similar to Vue.js for Python.☆16Sep 28, 2024Updated last year
- A collection of research papers related to Natural Language Reasoning☆10May 27, 2022Updated 4 years ago
- 基于多层级语言特征融合的中文文本可读性分级模型☆12Feb 27, 2024Updated 2 years ago
- Yet Another Chinese Learner Corpus☆77Jan 10, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆21Nov 19, 2023Updated 2 years ago
- VIMQA dataset☆15Jul 6, 2022Updated 4 years ago
- Master thesis: Exploring bias in German NLG (GPT-3 & GerPT-2). Applies regard classification and bias mitigation triggers.☆16Sep 25, 2024Updated last year
- Automated Essay Scoring Method for Chinese Second Language Writing☆33Mar 17, 2022Updated 4 years ago
- We provide benchmark datasets for evaluating Vietnamese processing models: UIT-ViQuAD, ViNewsQA, UIT-VSFC, UIT-ViIC, UIT-ViNames, UIT-VSM…☆22Jun 19, 2021Updated 5 years ago
- The Arborator software is aimed at collaboratively annotating dependency corpora.☆26Nov 5, 2019Updated 6 years ago
- Garnet - a graphical toolkit for Lisp☆21Dec 30, 2021Updated 4 years ago
- Concept dictionary☆41Apr 4, 2024Updated 2 years ago
- QuanSyn: A Python Package for Quantitative Syntax Analysis.☆40Jul 3, 2026Updated 2 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Promoting critical thinking through machine-generated prompts.☆19Sep 21, 2021Updated 4 years ago
- Workshop on Language Technologies for Historical and Ancient Languages (LT4HALA)☆38May 19, 2026Updated 2 months ago
- Xmixers: A collection of SOTA efficient token/channel mixers☆28Sep 4, 2025Updated 10 months ago
- A Supervised Word Alignment Method based on Cross-Language Span Prediction using Multilingual BERT☆26Jan 27, 2021Updated 5 years ago
- 这是一个中日 汉字 文字转换网站☆39Mar 4, 2021Updated 5 years ago
- Code for the paper: https://arxiv.org/pdf/2309.06979.pdf☆21Jul 29, 2024Updated last year
- Double-Sided Braille Image Dataset☆29Nov 18, 2020Updated 5 years ago
- Pretrained BERT for Ancient (Classical) Chinese, with an expanded vocabulary for rare characters.☆48Feb 20, 2023Updated 3 years ago
- This is a code example repo for the NLP course offered by the Institute of Chinese Information Processing of BNU.☆57May 2, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A curated list of resources dedicated to Natural Language Processing (NLP) of Cantonese | 粵語 NLP☆95Oct 17, 2021Updated 4 years ago
- Deep learning spelling patterns with a recurrent neural network☆11Jun 5, 2017Updated 9 years ago
- HTML Agent based on NexAU☆16Nov 20, 2025Updated 8 months ago
- Ancient Chinese Corpus with Word Sense Annotation☆73May 29, 2024Updated 2 years ago
- kgraph whitepaper☆19Nov 16, 2021Updated 4 years ago
- 中文自然语言处理数据集,平时做做实验的材料。欢迎补充提交合并。☆38Dec 3, 2021Updated 4 years ago
- Elm Set built on top of AnyDict☆10Aug 12, 2024Updated last year
- ☆14Jan 20, 2023Updated 3 years ago
- ☆57Oct 23, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Anki 复习数据处理与分析☆18May 18, 2021Updated 5 years ago
- Content Extraction via Text Density (SIGIR11)☆24Sep 21, 2015Updated 10 years ago
- ☆57May 4, 2026Updated 2 months ago
- For loops in const☆13Jul 3, 2026Updated 2 weeks ago
- Reference implementation of "Softmax Attention with Constant Cost per Token" (Heinsen, 2024)☆25Jun 6, 2024Updated 2 years ago
- A tool for ancient Chinese segmentation.☆54Apr 27, 2019Updated 7 years ago
- A C++ library implementing fast language models estimation using the 1-Sort algorithm.☆16May 18, 2023Updated 3 years ago