Fast bare-bones BPE for modern tokenizer training
☆180Jun 23, 2025Updated last year
Alternatives and similar repositories for bpeasy
Users that are interested in bpeasy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Simple Byte pair Encoding mechanism used for tokenization process . written purely in C☆152Nov 11, 2024Updated last year
- The official PyTorch implementation of Google's Gemma models☆5,719May 30, 2025Updated last year
- JAX implementation ViT-VQGAN☆66Jul 23, 2022Updated 4 years ago
- Google+ Blog☆15Oct 9, 2011Updated 14 years ago
- UNet diffusion model in pure CUDA☆663Jun 28, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A benchmark to evaluate language models on questions I've previously asked them to solve.☆1,065Apr 27, 2025Updated last year
- RuLES: a benchmark for evaluating rule-following in language models☆256Feb 24, 2025Updated last year
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated 2 years ago
- ####### ALERT! #########: my fork of the project has moved:☆18Dec 23, 2016Updated 9 years ago
- Minimal, clean code for the Byte Pair Encoding (BPE) algorithm commonly used in LLM tokenization.☆10,692Jul 1, 2024Updated 2 years ago
- Code for the paper "Getting the most out of your tokenizer for pre-training and domain adaptation"☆22Feb 14, 2024Updated 2 years ago
- A subset of PyTorch's neural network modules, written in Python using OpenAI's Triton.☆605Aug 14, 2026Updated 2 weeks ago
- Artificial Life simulator using canvas. Based on https://github.com/karpathy/scriptsbots☆79Jun 7, 2016Updated 10 years ago
- 0-Shot Tokenizer Transplant☆14May 16, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- BPE modification that implements removing of the intermediate tokens during tokenizer training.☆27Nov 25, 2024Updated last year
- ☆11Mar 17, 2026Updated 5 months ago
- Extracts plain text, language identification and more metadata from WARC records☆22Apr 16, 2026Updated 4 months ago
- Simple MPI implementation for prototyping or learning☆327Aug 6, 2025Updated last year
- Code for Zero-Shot Tokenizer Transfer☆147Jan 14, 2025Updated last year
- ☆16Apr 4, 2022Updated 4 years ago
- Multi-Word Probabilistic based supertokenizer☆15May 15, 2025Updated last year
- Supervoice diffusion enhance☆28Jul 15, 2024Updated 2 years ago
- Ruby Gem that makes sure that only a single instance of a code block is running.☆16Mar 13, 2013Updated 13 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- gpt-2 from scratch in mlx☆438Jun 12, 2024Updated 2 years ago
- Visualize multi-model embedding spaces. The first goal is to quickly get a lay of the land of any embedding space. Then be able to scroll…☆28May 16, 2024Updated 2 years ago
- Measuring if attention is explanation with ROAR☆22Mar 3, 2023Updated 3 years ago
- [ICLR 2025] Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling☆971Nov 16, 2025Updated 9 months ago
- Fine-tune mistral-7B on 3090s, a100s, h100s☆737Oct 11, 2023Updated 2 years ago
- A fusion of a linear layer and a cross entropy loss, written for pytorch in triton.☆76Aug 2, 2024Updated 2 years ago
- ☆12Jun 27, 2024Updated 2 years ago
- Code for co-training large language models (e.g. T0) with smaller ones (e.g. BERT) to boost few-shot performance☆16Sep 23, 2022Updated 3 years ago
- Community Implementation of the paper: "Multi-Head Mixture-of-Experts" In PyTorch☆31Aug 3, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Experimental CUDA kernel framework unifying typed dimensions, NVRTC JIT specialization, and ML‑guided tuning.☆47Feb 9, 2026Updated 6 months ago
- ☆20Apr 26, 2026Updated 4 months ago
- Using experimental methods to merge large language models☆11Jul 11, 2026Updated last month
- Training and evaluation code for the paper "Headless Language Models: Learning without Predicting with Contrastive Weight Tying" (https:/…☆29Apr 17, 2024Updated 2 years ago
- PyTorch Implementation of ViT-TTS (EMNLP'23)☆11Oct 20, 2023Updated 2 years ago
- Code and data for automatic paraphrase dataset augmentation.☆11Mar 8, 2021Updated 5 years ago
- decontamination☆38Mar 4, 2026Updated 5 months ago