The missing tiktoken training code
☆517Jan 3, 2026Updated 8 months ago
Alternatives and similar repositories for rustbpe
Users that are interested in rustbpe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Analyzing Hacker News discussions from a decade ago in hindsight with LLMs☆677Dec 10, 2025Updated 9 months ago
- Minimal Claude Code alternative. Single Python file, zero dependencies, ~250 lines.☆2,675Jan 14, 2026Updated 8 months ago
- The best ChatGPT that $100 can buy.☆58,164Sep 7, 2026Updated last week
- Minimalistic 4D-parallelism distributed training framework for education purpose☆2,305Aug 26, 2025Updated last year
- 🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.☆254Jun 27, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- nice and effective super simple calorie counter web app☆164May 30, 2024Updated 2 years ago
- NanoGPT (124M) in 90 seconds☆5,807Updated this week
- Persistent dict, backed by sqlite3 and pickle, multithread-safe.☆50Feb 21, 2020Updated 6 years ago
- A research tool for visually exploring Bureau of Labor Statistics Occupational Outlook Handbook data. This is not a report, a paper, or a…☆2,010Mar 16, 2026Updated 6 months ago
- FlexAttention based, minimal vllm-style inference engine for fast Gemma 2 inference.☆359Nov 2, 2025Updated 10 months ago
- Automating research publications discovery and analysis. For example, ever wish your computer could automatically open papers that are mo…☆505Sep 1, 2023Updated 3 years ago
- A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API☆17,591Aug 3, 2026Updated last month
- Game making library for using Canvas element☆109Oct 17, 2023Updated 2 years ago
- The simplest, fastest repository for training/finetuning small-sized VLMs.☆5,028Oct 27, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The simplest, fastest repository for training/finetuning medium-sized GPTs.☆63,254Nov 12, 2025Updated 10 months ago
- Our library for RL environments + evals☆4,635Updated this week
- A numeric optimization package for Torch.☆42Aug 19, 2021Updated 5 years ago
- Minimal, clean code for the Byte Pair Encoding (BPE) algorithm commonly used in LLM tokenization.☆10,728Jul 1, 2024Updated 2 years ago
- PyTorch Single Controller☆1,076Updated this week
- gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI☆20,420Jul 24, 2026Updated last month
- A set of tools/pages that help explore academic literature☆89Aug 11, 2014Updated 12 years ago
- LLM training in simple, raw C/CUDA☆31,030Jun 26, 2025Updated last year
- ☆15Aug 22, 2026Updated 3 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- my blog☆1,913Apr 10, 2026Updated 5 months ago
- AI agents running research on single-GPU nanochat training automatically☆96,417Mar 26, 2026Updated 5 months ago
- LLM Council works together to answer your hardest questions☆24,924Nov 22, 2025Updated 9 months ago
- Efficient implementation (and explorations) into polar coordinate positional embedding (PoPE) - from Gopalakrishnan et al. under Schmidhu…☆73Jun 21, 2026Updated 2 months ago
- Implementation of the fast weight product key memory from Sakana AI☆20Updated this week
- A PyTorch native platform for training generative AI models☆5,753Updated this week
- Efficient Triton Kernels for LLM Training☆6,620Updated this week
- UNet diffusion model in pure CUDA☆665Jun 28, 2024Updated 2 years ago
- Simple Byte pair Encoding mechanism used for tokenization process . written purely in C☆154Nov 11, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- CIFAR-10 speedrun: Trains to 94% accuracy in 1.98 seconds on a single NVIDIA A100 GPU.☆81Jul 30, 2026Updated last month
- Implementing DeepSeek R1's GRPO algorithm from scratch☆1,897Apr 18, 2025Updated last year
- Agentic RL Training at Scale☆2,059Updated this week
- Simple MPI implementation for prototyping or learning☆329Aug 6, 2025Updated last year
- Minimalistic large language model 3D-parallelism training☆2,825Updated this week
- Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models☆4,680Jan 14, 2026Updated 8 months ago
- Simple & Scalable Pretraining for Neural Architecture Research☆347Mar 31, 2026Updated 5 months ago