Temporary remove unused tokens during training to save ram and speed.
☆23Jun 15, 2025Updated last year
Alternatives and similar repositories for transformer-smaller-training-vocab
Users that are interested in transformer-smaller-training-vocab are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LV-BERT: Exploiting Layer Variety for BERT (Findings of ACL 2021)☆19May 10, 2023Updated 3 years ago
- ☆28Apr 19, 2026Updated 4 months ago
- Code for the paper "Getting the most out of your tokenizer for pre-training and domain adaptation"☆22Feb 14, 2024Updated 2 years ago
- Compiled tools, datasets, and other resources for historical text normalization.