The Bytepiece Tokenizer Implemented in Rust.
☆15Nov 28, 2023Updated 2 years ago
Alternatives and similar repositories for bytepiece-rs
Users that are interested in bytepiece-rs are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Manages vllm-nccl dependency☆19Jun 3, 2024Updated 2 years ago
- FLASHQuad_pytorch☆68Apr 1, 2022Updated 4 years ago
- Convert between pulldown parser events for various markup formats☆24Apr 9, 2026Updated 4 months ago
- Debug DeepSpeed-Chat step by step in IDE (在IDE里一步一步调试DeepSpeed-Chat)☆10Apr 17, 2023Updated 3 years ago
- bert-of-theseus via bert4keras☆31Jul 17, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Graph model execution API for Candle☆18Jul 27, 2025Updated last year
- 同花顺算法挑战平台:【9-10双月赛】跨领域迁移的文本语义匹配☆11Oct 28, 2021Updated 4 years ago
- 3位代码类目表;6位扩展代码表;疾病分类与代码(修订版);章节名称及代码☆11Aug 20, 2018Updated 7 years ago
- 更纯粹、更高压缩率的Tokenizer☆488Nov 27, 2024Updated last year
- Lower chisel memories to SRAM macros☆13Mar 25, 2024Updated 2 years ago
- Running LLaMA 3 with Rust.☆10May 21, 2024Updated 2 years ago
- A Rust crate offering similar functionality to the Python transformers package using Candle.☆15Nov 19, 2024Updated last year
- Sampling techniques for Candle.☆21Apr 3, 2024Updated 2 years ago
- 基于 CUDA Driver API 的 cuda 运行时环境☆16Jul 30, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- #[derive(Future, Stream, Sink, AsyncRead, AsyncWrite, AsyncSeek, AsyncBufRead)] for enums.☆18Updated this week
- #[derive(Iterator, DoubleEndedIterator, ExactSizeIterator, FusedIterator, Extend)] for enums.☆15Aug 8, 2026Updated last week
- Handy tools & graphics API abstraction for blazing fast prototyping☆10Jan 17, 2024Updated 2 years ago
- This repository open-sources our GEC system submitted by THU KELab (sz) in the CCL2023-CLTC Track 1: Multidimensional Chinese Learner Tex…☆15Nov 25, 2023Updated 2 years ago
- CCL2024 Chinese Essay Rhetoric Recognition and Understanding☆17Oct 1, 2024Updated last year
- A Feishu/Lark AI agent bot☆15Feb 27, 2026Updated 5 months ago
- ☆10Apr 29, 2023Updated 3 years ago
- 方便扩展的Cuda算子理解和优化框架,仅用在学习使用☆18Jun 13, 2024Updated 2 years ago
- ncnn export & infer mobileclip☆21Aug 18, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆12Aug 10, 2021Updated 5 years ago
- 🦀 A Rust implementation of a RoBERTa classification model for the SNLI dataset☆13Sep 13, 2021Updated 4 years ago
- pytorch Efficient GlobalPointer☆57Apr 12, 2022Updated 4 years ago
- Real-time AI video segmentation of USB camera and streaming over HTTP☆13Apr 23, 2025Updated last year
- Exploration of semantic chunking and chunk classification☆19Sep 16, 2024Updated last year
- An async-session implementation for MongoDB☆18Aug 21, 2022Updated 3 years ago
- ☆13Jul 28, 2026Updated 3 weeks ago
- 2023全球智能汽车AI挑战赛——赛道一:AI大模型检索问答, 75+ baseline☆58Dec 7, 2023Updated 2 years ago
- Concurrency and parallelism in async Rust☆13Oct 26, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code repo for "CritiPrefill: A Segment-wise Criticality-based Approach for Prefilling Acceleration in LLMs".☆17Sep 15, 2024Updated last year
- Pure sine wave inverter using sPWM signal generated from arduino☆12Jan 22, 2024Updated 2 years ago
- Rust implementation of the 7 GUI tasks by Eugen Kiss using Druid☆13Dec 27, 2020Updated 5 years ago
- A filter proxy for StatsD☆24Nov 29, 2022Updated 3 years ago
- Port of LibLine to Zig☆26May 20, 2026Updated 2 months ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆15Jun 27, 2026Updated last month
- Use common pre-trained ML models in Deno!☆19Nov 21, 2021Updated 4 years ago