歴史資料の市民参加型翻刻プラットフォーム「みんなで翻刻」のテキストデータ置き場です。 / Transcription texts created on Minna de Honkoku (https://honkoku.org), a crowdsourced transcription platform for historical Japanese documents.
☆21Jul 26, 2026Updated this week
Alternatives and similar repositories for honkoku-data
Users that are interested in honkoku-data are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- NDL古典籍OCR学習用データセット(みんなで翻刻加工データ)☆22Mar 13, 2026Updated 4 months ago
- TEIガイドラインへの準拠の仕方を日本語で解説します。☆12Feb 15, 2021Updated 5 years ago
- NDL古典籍OCR-Liteのアプリケーションのリポジトリ(ソースコードを含む)☆178Jul 15, 2026Updated 2 weeks ago
- デジタル化資料から作成したOCRテキストデータのngram頻度統計情報のデータセット☆17Jan 10, 2023Updated 3 years ago
- 青空文庫テキストをより便利にする(機械可読性を高める)ためのプロジェクト☆26Jun 23, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆27May 6, 2026Updated 2 months ago
- Schema for modelling parliamentary debates☆23May 23, 2022Updated 4 years ago
- An OpenSeadragon selection plugin for the tiled viewer.☆51Jun 14, 2026Updated last month
- ☆14Jun 29, 2021Updated 5 years ago
- A kaomoji input method for macOS. ( ^ ▽ ^ )☆10Jun 4, 2026Updated last month
- デジタル化資料OCRテキスト化事業において作成されたOCR学習用データセット☆83Jun 26, 2024Updated 2 years ago
- WEB+DB PRESS Vol.92特集1「Web開発新人研修」で利用したサンプルコードを公開しています。☆15May 11, 2016Updated 10 years ago
- FIWARE 401: IDM - Managing Users and Organizations☆10May 15, 2026Updated 2 months ago
- Examples of integrating Mirador with modern frontend build systems☆10Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆15Jun 27, 2026Updated last month
- 解析が難しい日本の 住所のテストデータセット☆14Sep 25, 2023Updated 2 years ago
- Python Bluetooth low energy observer example for OMRON Environment Sensor (2JCIE-BL01)☆36May 13, 2021Updated 5 years ago
- ☆30Mar 6, 2026Updated 4 months ago
- Tokenizer POS-tagger Lemmatizer and Dependency-parser for modern and contemporary Japanese☆38Dec 29, 2025Updated 7 months ago
- ベイズ階層言語モデルによる教師なし形態素解析☆34Oct 16, 2023Updated 2 years ago
- Apple Vision Pro向けの日本語入力アプリ☆11Feb 8, 2024Updated 2 years ago
- A module for Omeka S that provides an API for the Neatline 3 single page application☆18Mar 26, 2023Updated 3 years ago
- A Firefox/Chrome extension to open IIIF manifest link in your favorite IIIF viewer.☆22Jul 9, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- translation tool☆11Apr 1, 2024Updated 2 years ago
- The Text Encoding Initiative Guidelines☆344Updated this week
- ☆13May 3, 2017Updated 9 years ago
- AMI Meeting Parallel Corpus☆13Dec 11, 2020Updated 5 years ago
- Core functions and components for RecogitoJS and Annotorious☆16Nov 9, 2023Updated 2 years ago
- safe and easy programming for you☆10Sep 1, 2022Updated 3 years ago
- Tokenizer POS-Tagger and Dependency-parser with BERT/RoBERTa/DeBERTa/GPT models for Japanese and other languages☆55Feb 28, 2026Updated 5 months ago
- 文法誤り訂正に関する日本語文献を収集・分類するためのリポジトリ☆14Apr 17, 2025Updated last year
- Inforex is a web system for text corpora construction.☆12Jun 24, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Linguistic Reconstruction with LingPy☆16Aug 5, 2024Updated last year
- [WIP] 青空文庫テキストのパーサ☆12Feb 23, 2019Updated 7 years ago
- IIIF Presentation API 3 Python Library☆41Jul 20, 2026Updated last week
- GTFS/GTFS-JP固定URLデータ 日付チェック☆11Updated this week
- 日本十進分類法のIME辞書☆11Dec 8, 2022Updated 3 years ago
- Word List by Semantic Principles (WLSP): “It is a collection of words classified and arranged by their meanings”☆69Aug 20, 2025Updated 11 months ago
- Viterbi-based accelerated tokenizer (Python wrapper)☆46May 30, 2026Updated last month