FlexiTokens
☆23Dec 27, 2025Updated 8 months ago
Alternatives and similar repositories for flexitokens
Users that are interested in flexitokens are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- BPE modification that implements removing of the intermediate tokens during tokenizer training.☆27Nov 25, 2024Updated last year
- 🍡 30x faster tokenization for every HuggingFace model☆50Updated this week
- TokEval: intrinsic quality metrics for tokenizers across natural language, code, and math☆53Updated this week
- A Multilingual Keyboard Layout-Based Typo Generator☆17Nov 23, 2025Updated 9 months ago
- Official code for the NeurIPS25 paper "RAT: Bridging RNN Efficiencyand Attention Accuracy in Language Modeling" (https://arxiv.org/abs/25…☆26Dec 10, 2025Updated 8 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Getting interpretable dimensions in word embedding spaces.☆15Jul 6, 2023Updated 3 years ago
- Code for SaGe subword tokenizer (EACL 2023)☆28Nov 30, 2024Updated last year
- 🚀🤗 A collection of templates for Hugging Face Spaces☆35Oct 9, 2023Updated 2 years ago
- PathPiece tokenizer☆14Nov 10, 2024Updated last year
- This repository contains data, code and models for contextual noncompliance.☆26Jul 18, 2024Updated 2 years ago
- ☆21Apr 3, 2026Updated 4 months ago
- Pre-train Static Word Embeddings☆111Jun 9, 2026Updated 2 months ago
- ☆16Aug 22, 2026Updated last week
- 2D chess pieces inspired by handmade wooden chess sets, featuring cuteness and simplicity☆15Feb 14, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for the paper "BPE stays on SCRIPT", "Which Pieces Does Unigram Tokenization Really Need?" and MinGram☆21Updated this week
- A code snippet that proves that there is no legal, potentially non-reachable chess position with more than 218 moves.☆23May 23, 2024Updated 2 years ago
- ANE accelerated embedding models!☆20Dec 11, 2024Updated last year
- Datamodels for hugging face tokenizers☆111Updated this week
- Nearly Inference Free Embeddings: make your RAG queries 500x faster☆85Apr 27, 2026Updated 4 months ago
- This repository includes the masking vocabulary used in the ICLR 2021 spotlight PMI-Masking paper☆14Aug 9, 2021Updated 5 years ago
- LV-BERT: Exploiting Layer Variety for BERT (Findings of ACL 2021)☆19May 10, 2023Updated 3 years ago
- ☆104Jul 4, 2025Updated last year
- Contextualized per-token embeddings☆41Aug 22, 2026Updated last week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Crowd-sourced lists of urls to help Common Crawl crawl under-resourced languages. See https://github.com/commoncrawl/web-languages-code/ …☆71Aug 24, 2026Updated last week
- The training codes of Jasper-Token-Compression-600M☆21Nov 19, 2025Updated 9 months ago
- Official Repository for Paper "BaichuanSEED: Sharing the Potential of ExtensivE Data Collection and Deduplication by Introducing a Compet…☆18Aug 28, 2024Updated 2 years ago
- Vocabulary Trimming (VT) is a model compression technique, which reduces a multilingual LM vocabulary to a target language by deleting ir…☆69Oct 25, 2024Updated last year
- Python library to use Pleias-RAG models☆72Jul 1, 2026Updated last month
- Official code release for "SuperBPE: Space Travel for Language Models"☆98May 28, 2026Updated 3 months ago
- [ACL'26 Findings] Recovered in Translation: Efficient Pipeline for Automated Translation of Benchmarks and Datasets☆20Jun 27, 2026Updated 2 months ago
- Trully flash implementation of DeBERTa disentangled attention mechanism.☆91Feb 10, 2026Updated 6 months ago
- Automatically exported from code.google.com/p/transducersaurus☆11Apr 1, 2015Updated 11 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This repository contains the code for the Form-Context Model and its Attentive Mimicking variant.☆30May 11, 2020Updated 6 years ago
- ☆58Dec 27, 2025Updated 8 months ago
- Nanoloop source files for the album "Prime 16"☆12Mar 7, 2026Updated 5 months ago
- Code for ACL 2023 Paper: ACLM: A Selective-Denoising based Generative Data Augmentation Approach for Low-Resource Complex NER☆22Jul 19, 2023Updated 3 years ago
- Trainable embedding transformation for confidence estimation, feature extraction, explainability and conversion from dense to sparse.☆28Jun 23, 2026Updated 2 months ago
- Metric Space Magnitude Computations☆15Jun 30, 2026Updated 2 months ago
- Ukrainian ELECTRA model☆12Mar 11, 2023Updated 3 years ago