Code for Bolmo: Byteifying the Next Generation of Language Models
☆136Aug 28, 2026Updated last week
Alternatives and similar repositories for bolmo-core
Users that are interested in bolmo-core are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆11Mar 17, 2026Updated 5 months ago
- [EMNLP'23] Official Code for "FOCUS: Effective Embedding Initialization for Monolingual Specialization of Multilingual Models"☆37Jun 7, 2025Updated last year
- ☆16Aug 22, 2026Updated 2 weeks ago
- ♡☆23Aug 16, 2026Updated 3 weeks ago
- A library for language transfer methods and algorithms.☆16Feb 6, 2026Updated 7 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ACL'26 Findings] Recovered in Translation: Efficient Pipeline for Automated Translation of Benchmarks and Datasets☆20Jun 27, 2026Updated 2 months ago
- The official code of "PixelWorld: Towards Perceiving Everything as Pixels" [TMLR25]☆15Sep 12, 2025Updated 11 months ago
- LTG-Bert☆35Jan 8, 2024Updated 2 years ago
- Extending the Context of Pretrained LLMs by Dropping Their Positional Embedding☆222Jan 12, 2026Updated 7 months ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆17Apr 30, 2025Updated last year
- This is the repository for MorphScore, a tokenizer evaluation framework for morphological alignment.☆19Jul 10, 2025Updated last year
- Code for EMNLP2021 paper "Allocating Large Vocabulary Capacity for Cross-lingual Language Model Pre-training"☆20Nov 12, 2021Updated 4 years ago
- ☆16Aug 7, 2026Updated last month
- Semantic Regex☆20Nov 13, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- suffix array construction and searching algorithms for in-memory binary data.☆13Sep 10, 2022Updated 3 years ago
- Fluid Language Model Benchmarking☆29Sep 16, 2025Updated 11 months ago
- Official repo for BWLer: Barycentric Weight Layer☆31Aug 11, 2026Updated 3 weeks ago
- RePo: Language Models with Context Re-Positioning☆84Mar 30, 2026Updated 5 months ago
- Implementation of the MetaController proposed in "Emergent temporal abstractions in autoregressive models enable hierarchical reinforceme…☆106Aug 14, 2026Updated 3 weeks ago
- GPT-2 Metadata Pretraining Towards Instruction Finetuning for Ukrainian☆20Aug 6, 2023Updated 3 years ago
- Research into identifying and correcting incorrect labels in the CoNLL-2003 corpus.☆12May 11, 2021Updated 5 years ago
- Complete set of English dialect transformation rules and evaluation code☆15Jun 7, 2024Updated 2 years ago
- Codebase for FinePDFs☆191Jan 9, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆16Dec 9, 2023Updated 2 years ago
- ☆59Sep 26, 2025Updated 11 months ago
- ☆16May 14, 2024Updated 2 years ago
- Code for the paper "BPE stays on SCRIPT", "Which Pieces Does Unigram Tokenization Really Need?" and MinGram☆22Aug 27, 2026Updated last week
- Efficient non-uniform quantization with GPTQ for GGUF☆64Sep 17, 2025Updated 11 months ago
- Official Project Page for Deep Delta Learning (https://arxiv.org/abs/2601.00417)☆357Jul 27, 2026Updated last month
- Code for BLT research paper☆2,060Nov 3, 2025Updated 10 months ago
- [ICLR 2026] Official PyTorch Implementation of RLP: Reinforcement as a Pretraining Objective☆254Jan 26, 2026Updated 7 months ago
- ✂️ Sentence segmentation with wtpsplit's state-of-the-art Segment any Text (SaT) models☆41May 2, 2026Updated 4 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Master thesis: Exploring bias in German NLG (GPT-3 & GerPT-2). Applies regard classification and bias mitigation triggers.☆16Sep 25, 2024Updated last year
- WeDLM: The fastest diffusion language model with standard causal attention and native KV cache compatibility, delivering real speedups ov…☆655Mar 3, 2026Updated 6 months ago
- ☆17Sep 11, 2025Updated 11 months ago
- PathPiece tokenizer☆14Nov 10, 2024Updated last year
- advanced, scalable, no-code RAG☆330Aug 4, 2026Updated last month
- Trainable H-Net Package☆34Sep 3, 2025Updated last year
- ☆60Nov 18, 2025Updated 9 months ago