The missing tiktoken training code
☆502Jan 3, 2026Updated 6 months ago
Alternatives and similar repositories for rustbpe
Users that are interested in rustbpe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Analyzing Hacker News discussions from a decade ago in hindsight with LLMs☆668Dec 10, 2025Updated 7 months ago
- Minimal Claude Code alternative. Single Python file, zero dependencies, ~250 lines.☆2,524Jan 14, 2026Updated 6 months ago
- The best ChatGPT that $100 can buy.☆56,732Jul 4, 2026Updated 3 weeks ago
- Render any git repo into a single static HTML page for humans or LLMs☆2,416Aug 21, 2025Updated 11 months ago
- Quick illustration of how one can easily read books together with LLMs. It's great and I highly recommend it.☆3,806Nov 18, 2025Updated 8 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Minimalistic 4D-parallelism distributed training framework for education purpose☆2,260Aug 26, 2025Updated 11 months ago
- 🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.☆250Jun 27, 2022Updated 4 years ago
- nice and effective super simple calorie counter web app☆160May 30, 2024Updated 2 years ago
- NanoGPT (124M) in 90 seconds☆5,600Updated this week
- Persistent dict, backed by sqlite3 and pickle, multithread-safe.☆49Feb 21, 2020Updated 6 years ago
- FlexAttention based, minimal vllm-style inference engine for fast Gemma 2 inference.☆356Nov 2, 2025Updated 8 months ago
- Automating research publications discovery and analysis. For example, ever wish your computer could automatically open papers that are mo…☆492Sep 1, 2023Updated 2 years ago
- A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API☆16,889Aug 8, 2024Updated last year
- Game making library for using Canvas element☆109Oct 17, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- The simplest, fastest repository for training/finetuning small-sized VLMs.☆4,969Oct 27, 2025Updated 9 months ago
- The simplest, fastest repository for training/finetuning medium-sized GPTs.☆61,632Nov 12, 2025Updated 8 months ago
- Our library for RL environments + evals☆4,410Updated this week
- A numeric optimization package for Torch.☆42Aug 19, 2021Updated 4 years ago
- Minimal, clean code for the Byte Pair Encoding (BPE) algorithm commonly used in LLM tokenization.☆10,647Jul 1, 2024Updated 2 years ago
- PyTorch Single Controller☆1,065Updated this week
- gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI☆20,266Updated this week
- my blog☆1,825Apr 10, 2026Updated 3 months ago
- LLM training in simple, raw C/CUDA☆30,663Jun 26, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆15Dec 4, 2025Updated 7 months ago
- A set of tools/pages that help explore academic literature☆88Aug 11, 2014Updated 11 years ago
- AI agents running research on single-GPU nanochat training automatically☆92,224Mar 26, 2026Updated 4 months ago
- LLM Council works together to answer your hardest questions☆23,300Nov 22, 2025Updated 8 months ago
- A PyTorch native platform for training generative AI models☆5,568Updated this week
- Implementation of the fast weight product key memory from Sakana AI☆19Apr 1, 2026Updated 3 months ago
- Efficient Triton Kernels for LLM Training☆6,537Updated this week
- UNet diffusion model in pure CUDA☆661Jun 28, 2024Updated 2 years ago
- Simple Byte pair Encoding mechanism used for tokenization process . written purely in C☆151Nov 11, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- CIFAR-10 speedrun: Trains to 94% accuracy in 1.98 seconds on a single NVIDIA A100 GPU.☆79Oct 17, 2025Updated 9 months ago
- Implementing DeepSeek R1's GRPO algorithm from scratch☆1,883Apr 18, 2025Updated last year
- Simple MPI implementation for prototyping or learning☆325Aug 6, 2025Updated 11 months ago
- Minimalistic large language model 3D-parallelism training☆2,768May 26, 2026Updated 2 months ago
- Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models☆4,568Jan 14, 2026Updated 6 months ago
- Simple & Scalable Pretraining for Neural Architecture Research☆340Mar 31, 2026Updated 3 months ago
- Agentic RL Training at Scale☆1,759Updated this week