Fast Tokens
β140Sep 10, 2026Updated this week
Alternatives and similar repositories for fastokens
Users that are interested in fastokens are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LLM Kernel Library for Rustβ31Updated this week
- π¦ Rust library of natural language dictionaries using character-wise double-array tries.β39Aug 12, 2026Updated last month
- Model Express is a Rust-based component meant to be placed next to existing model inference systems to speed up their startup times and iβ¦β149Updated this week
- LBFGS optimization algorithm ported from liblbfgsβ12Nov 25, 2022Updated 3 years ago
- Implements kernels with RISC-V Vectorβ22Mar 24, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- High-performance GPU kernels for Ads and Recsys model training, independently implemented and optimized for real-world workloads and modeβ¦β42Aug 25, 2026Updated 2 weeks ago
- Includes a file with zstd compression in Rustβ14Feb 17, 2023Updated 3 years ago
- Completed codelabs from the gRPC project.β18Sep 2, 2026Updated last week
- A high-performance and light-weight router for vLLM large scale deploymentβ403Updated this week
- β345Updated this week
- IP Over Infiniband (IPoIB) CNI Pluginβ19Updated this week
- π A fast implementation of the Aho-Corasick algorithm using the compact double-array data structure in Rust.β283Aug 18, 2026Updated 3 weeks ago
- A better wrapper for using RDMA programming APIs in Rust flavorβ96Aug 13, 2026Updated last month
- TokenSpeed is a speed-of-light LLM inference engine.β2,116Updated this week
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for the examples presented in the talk "Training a Llama in your backyard: fine-tuning very large models on consumer hardware" givenβ¦β15Oct 16, 2023Updated 2 years ago
- Set-Encoder: Permutation-Invariant Inter-Passage Attention for Listwise Passage Re-Ranking with Cross-Encodersβ19May 23, 2025Updated last year
- Elixir library for Apple Containersβ15Mar 24, 2026Updated 5 months ago
- Learning High-Quality and General-Purpose Phrase Representations. Findings of EACL 2024β16Feb 29, 2024Updated 2 years ago
- SiMM: Scalable in-Memory Middlewareβ42Apr 20, 2026Updated 4 months ago
- An on-device, GPU-accelerated language model in Rust.β36Aug 18, 2026Updated 3 weeks ago
- Systematic and comprehensive benchmarks for LLM systems.β61Jan 28, 2026Updated 7 months ago
- Checkpoint-engine is a simple middleware to update model weights in LLM inference enginesβ1,005Sep 4, 2026Updated last week
- π₯ Vaporetto: Very accelerated pointwise prediction based tokenizerβ298Jul 20, 2026Updated last month
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- β96Updated this week
- High-performance KV cache storage for LLM inference β GPU offloading, SSD caching, and cross-node sharing via RDMA. Works with vLLM and Sβ¦β203Updated this week
- Communication patterns for AI, built on top of NCCL device and host APIsβ52Sep 3, 2026Updated last week
- Early-stage Rust drop-in alternative frontend for vLLMβ73May 22, 2026Updated 3 months ago
- CUDA Embedding Lookup Kernel Libraryβ50Jun 26, 2026Updated 2 months ago
- NVIDIA Inference Xfer Library (NIXL)β1,250Updated this week
- All things generative! Discord Botβ21Aug 13, 2023Updated 3 years ago
- A NCCL extension library, designed to efficiently offload GPU memory allocated by the NCCL communication library.β117Dec 17, 2025Updated 8 months ago
- bpf(2)-based ftrace(1)-like function graph tracer for golang processes.β33Aug 18, 2023Updated 3 years ago
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Kaggle AIMO2 solution with token-efficient reasoning LLM recipesβ51Aug 7, 2025Updated last year
- Proof of concept for running moshi/hibiki using webrtcβ21Feb 28, 2025Updated last year
- Unleash the full potential of exascale LLMs on consumer-class GPUs, proven by extensive benchmarks, with no long-term adjustments and minβ¦β27Nov 11, 2024Updated last year
- Official Code Repositiry for "RaDeR: Reasoning-aware Dense Retrieval Models" accepted at Main Conference EMNLP 2025β18Jun 23, 2025Updated last year
- IREE Tokenizer Python Bindingsβ21Mar 5, 2026Updated 6 months ago
- High-performance safetensors model loaderβ166Updated this week
- In this project, we propose to study Vision Transformers trained using the Barlow Twins self-supervised method, and compare the results wβ¦β17Oct 3, 2023Updated 2 years ago