A large parallel corpus of English and Japanese
☆91Nov 1, 2017Updated 8 years ago
Alternatives and similar repositories for JESC
Users that are interested in JESC are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An example usage of JParaCrawl pre-trained Neural Machine Translation (NMT) models.☆104Apr 29, 2021Updated 5 years ago
- ☆22Aug 18, 2020Updated 6 years ago
- 50k English-Japanese Parallel Corpus for Machine Translation Benchmark.☆98Sep 11, 2019Updated 6 years ago
- ☆22Dec 20, 2019Updated 6 years ago
- Zunda: Japanese Enhanced Modality Analyzer client for Python.☆10Nov 30, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Coursera Corpus Mining and Multistage Fine-Tuning for Improving Lectures Translation☆15Aug 27, 2024Updated 2 years ago
- Scripts for creating a Japanese-English parallel corpus and training NMT models☆19Nov 9, 2021Updated 4 years ago
- The Business Scene Dialogue corpus☆76Nov 10, 2021Updated 4 years ago
- Decoding platform for machine translation research☆54Aug 24, 2019Updated 7 years ago
- A Neural Machine Translation implementation in Chainer☆46May 22, 2020Updated 6 years ago
- Tools for extracting parallel corpora from article titles across languages in Wikipedia☆74Feb 25, 2015Updated 11 years ago
- ☆43Sep 16, 2020Updated 5 years ago
- Practical example from Human-in-the-Loop Machine Learning book☆11Oct 28, 2021Updated 4 years ago
- ☆24Nov 29, 2017Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Efficient Markov Chain word alignment☆53Aug 1, 2021Updated 5 years ago
- Efficient teacher-student models and scripts to make them☆58Dec 16, 2023Updated 2 years ago
- My implementation of LASER architecture in Fairseq☆12Oct 6, 2020Updated 5 years ago
- 1.身份证识别,可以拍照或导入身份证图片进行识别☆13Sep 27, 2022Updated 3 years ago
- Bicleaner is a parallel corpus classifier/cleaner that aims at detecting noisy sentence pairs in a parallel corpus.☆159Jun 18, 2024Updated 2 years ago
- Korean Parallel Corpus☆148Feb 24, 2024Updated 2 years ago
- MT Evaluation in Many Languages via Zero-Shot Paraphrasing☆102Jul 25, 2024Updated 2 years ago
- Yet another sentence-level tokenizer for the Japanese text☆24Nov 27, 2025Updated 9 months ago
- Neural macine translation soft alignment visualisations for web and command line☆72Aug 19, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Lexically Constrained Neural Machine Translation with Levenshtein Transformer☆40Jul 14, 2020Updated 6 years ago
- CaboCha wrapper for Python3☆46Jul 5, 2018Updated 8 years ago
- Modified version of fairseq, including new implementations for criterions using reinforcement learning methods.☆11Aug 14, 2019Updated 7 years ago
- Hadoop-based tool for extraction of large scale synchronous grammars for paraphrasing and machine translation☆15Dec 2, 2016Updated 9 years ago
- ☆15Nov 5, 2020Updated 5 years ago
- 敬語変換タスクにおける評価用データセット☆21Nov 24, 2022Updated 3 years ago
- A Supervised Word Alignment Method based on Cross-Language Span Prediction using Multilingual BERT☆26Jan 27, 2021Updated 5 years ago
- Reference BLEU implementation that auto-downloads test sets and reports a version string to facilitate cross-lab comparisons☆1,259Aug 20, 2026Updated 2 weeks ago
- A summarizer for Japanese articles (but ChatGPT is better)☆10Aug 1, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- COrpus based Morphological Analyzer with INtegrated User dictionary☆21Mar 30, 2025Updated last year
- Calorie counter for the command-line with 8,000 food items (USDA)☆11Sep 21, 2017Updated 8 years ago
- AMI Meeting Parallel Corpus☆13Dec 11, 2020Updated 5 years ago
- Python bindings for Tobii Gaze SDK☆11Jul 16, 2015Updated 11 years ago
- Crawling engine that crawls a set of top-level domains looking for documents in a list of languages☆11Feb 6, 2024Updated 2 years ago
- A High-Quality Multilingual Dataset for Structured Documentation Translation☆39May 1, 2025Updated last year
- Data collection, alignment and TAUS repository☆24Nov 30, 2017Updated 8 years ago