Yada is a yet another double-array trie library aiming for fast search and compact data representation.
☆49Jun 7, 2026Updated 3 months ago
Alternatives and similar repositories for yada
Users that are interested in yada are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A tool for visualizing the internal structures of morphological analyzer Sudachi☆18Jun 9, 2022Updated 4 years ago
- Japanese tokenizer for rust☆39Nov 5, 2019Updated 6 years ago
- A Japanese Morphological Analyzer written in pure Rust☆26Oct 25, 2019Updated 6 years ago
- A Japanese law parser☆25Jan 25, 2024Updated 2 years ago
- A multi-language segmenter using high-order CRF.☆17Feb 27, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Japanese synonym library☆55Feb 7, 2022Updated 4 years ago
- Efficiently-updatable double-array trie in Rust (ported from cedar)☆107Sep 7, 2026Updated 2 weeks ago
- ☆34Updated this week
- Finding all pairs of similar documents time- and memory-efficiently☆62Mar 13, 2025Updated last year
- 🐎 A fast implementation of the Aho-Corasick algorithm using the compact double-array data structure in Rust.☆284Aug 18, 2026Updated last month
- A blend of the compact and sparse hash table implementations.☆15Aug 20, 2021Updated 5 years ago
- Testing tool to verify the search qualities of the Elasticsearch indices☆29May 22, 2026Updated 4 months ago
- At-a-glance overview diagrams of Apache Lucene's default PostingsFormat (inverted index binary format).☆82Mar 25, 2023Updated 3 years ago
- 🌳 A compressed rank/select dictionary exploiting approximate linearity and repetitiveness.☆15Jun 28, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Suika 🍉 is a Japanese morphological analyzer written in pure Ruby☆53Jun 18, 2026Updated 3 months ago
- Yosina is a transliteration library deals with the letters and symbols used in Japanese writing.☆27Apr 18, 2026Updated 5 months ago
- A multilingual morphological analysis library.☆672Updated this week
- Compact Japanese tokenizer☆16Aug 17, 2018Updated 8 years ago
- ⚡Japanese sentence splitting(日本語文境界判定器), 40–250× faster via a Rust-accelerated Python library with near-perfect API compatibility with …☆77Oct 14, 2025Updated 11 months ago
- 🦞 Rust library of natural language dictionaries using character-wise double-array tries.☆39Aug 12, 2026Updated last month
- ☆19Jan 17, 2023Updated 3 years ago
- Build URL of GCP Cloud Logging Logs Explorer☆16Jul 4, 2024Updated 2 years ago
- go-active-learning is a command line annotation tool for binary classification problem written in Go.☆15Apr 3, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Sudachi in Rust 🦀 and new generation of SudachiPy☆480Updated this week
- Automatically exported from code.google.com/p/esaxx☆17Jun 23, 2015Updated 11 years ago
- NLP Toolkit based on Deep Learning☆35Apr 23, 2017Updated 9 years ago
- CC-CEDICT-MeCab is a MeCab dictionary for Chinese (Mandarin) text segmentation☆13Apr 9, 2020Updated 6 years ago
- bqiam is an admin tool for managing BigQuery permissions☆12Updated this week
- 『機械学習による検索ランキング改善ガイド』のサンプルコードのリポジトリ☆23Aug 3, 2023Updated 3 years ago
- Lindera tokenizer for Tantivy.☆70Aug 7, 2026Updated last month
- Edit and create Kubernetes job from cronjob template using your EDITOR☆18Apr 8, 2025Updated last year
- Yet another sentence-level tokenizer for the Japanese text☆24Nov 27, 2025Updated 9 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Lexical Augmented Unified Retrieval Using Semantics☆28Updated this week
- ☆10Aug 26, 2021Updated 5 years ago
- 🦀 A Rust implementation of a RoBERTa classification model for the SNLI dataset☆13Sep 13, 2021Updated 5 years ago
- LBFGS optimization algorithm ported from liblbfgs☆12Nov 25, 2022Updated 3 years ago
- tokenize text and separate it into words for Japanese☆11Jan 5, 2020Updated 6 years ago
- Japanese language utility in Golang.☆15Sep 21, 2020Updated 6 years ago
- A clone of Darts (Double-ARray Trie System)☆161May 14, 2025Updated last year