Library for fast text representation and classification.
☆31Jan 9, 2024Updated 2 years ago
Alternatives and similar repositories for fasterText
Users that are interested in fasterText are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Aug 23, 2024Updated last year
- Targetted language identifier, based on FastText and Hunspell.☆38Sep 4, 2025Updated 10 months ago
- Bicleaner fork that uses neural networks☆40Feb 23, 2026Updated 4 months ago
- A library for data streaming and augmentation☆22May 5, 2025Updated last year
- A simple Rust library to retrieve data from https://api.carbonintensity.org.uk/☆11Apr 25, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A Python utility for indexing file lines. Best demo honourable mention at ECIR 2024.☆23Nov 9, 2025Updated 8 months ago
- ☆39Apr 17, 2024Updated 2 years ago
- Working repo to support the Alliance's Open Trusted Data Initiative☆15Jul 1, 2026Updated 2 weeks ago
- Efficient teacher-student models and scripts to make them☆57Dec 16, 2023Updated 2 years ago
- A collection of Zsh functions to augment Git☆19Dec 11, 2025Updated 7 months ago
- Transform TMX to text☆27Nov 23, 2022Updated 3 years ago
- Examples for using the Daft data engine☆16Jul 10, 2026Updated last week
- ☆16Feb 10, 2026Updated 5 months ago
- data related codebase for polyglot project☆19Mar 30, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Command Line Interactive Periodic Table of Elements in multiple languages☆17Mar 28, 2023Updated 3 years ago
- collaborative web tool to enrich content☆11Nov 13, 2011Updated 14 years ago
- Inference slice of marian for bergamot's tiny11 models. Faster to compile, and wield. Fewer model-archs than bergamot-translator.☆16Oct 24, 2024Updated last year
- ☆29Feb 11, 2026Updated 5 months ago
- Bicleaner is a parallel corpus classifier/cleaner that aims at detecting noisy sentence pairs in a parallel corpus.☆160Jun 18, 2024Updated 2 years ago
- Code for our project CROWN (Conversational Passage Ranking by Reasoning over Word Networks)☆10Jan 11, 2024Updated 2 years ago
- ☆146Jul 2, 2026Updated 2 weeks ago
- ☆38Mar 16, 2026Updated 4 months ago
- ☆30Apr 29, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- COMET for African languages☆11Jan 24, 2025Updated last year
- Programs written in BASIC☆16Mar 20, 2021Updated 5 years ago
- A simple semi-supervised approach for creating huggingface data script loaders and upload to the hub.☆11Jun 23, 2024Updated 2 years ago
- All code and content for my blog.☆15Sep 23, 2018Updated 7 years ago
- The pipeline for the OSCAR corpus☆178Nov 9, 2025Updated 8 months ago
- Micro-framework for publishing linked data☆11Aug 1, 2017Updated 8 years ago
- AfroLID, a powerful neural toolkit for African languages identification which covers 517 African languages.☆39Feb 5, 2026Updated 5 months ago
- Coursera Corpus Mining and Multistage Fine-Tuning for Improving Lectures Translation☆15Aug 27, 2024Updated last year
- ☆32Mar 30, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Medusa combo files, Hashcat rules and dictionaries, JRT rules☆14Oct 20, 2022Updated 3 years ago
- IAI Style Guide☆10Jun 27, 2025Updated last year
- Library and command line utility to do approximate string matching of a source against a bitext index and get matched source and target.☆54Apr 22, 2025Updated last year
- Efficient Low-Memory Aligner☆148Jan 15, 2025Updated last year
- PyTorch implementation of NAACL 2021 paper "Multi-view Subword Regularization"☆26Jun 2, 2021Updated 5 years ago
- Code and experiments for the COLING2020 paper "Conception: Multilingually-Enhanced, Human-Readable Concept Vector Representations".☆11Dec 9, 2020Updated 5 years ago
- Transformer Implementation for NMT using PyTorch Lightning (Korean to English)☆10Oct 19, 2020Updated 5 years ago