☆72Feb 27, 2023Updated 3 years ago
Alternatives and similar repositories for huggingface-tokenizer-in-cxx
Users that are interested in huggingface-tokenizer-in-cxx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Universal cross-platform tokenizers binding to HF and sentencepiece☆497May 20, 2026Updated 2 months ago
- C++ implementation of tokenizers, including tiktoken.☆25Dec 7, 2023Updated 2 years ago
- transformer tokenizers (e.g. BERT tokenizer) in C++ (WIP)☆18Apr 7, 2022Updated 4 years ago
- Try to export the ONNX QDQ model that conforms to the AXERA NPU quantization specification. Currently, only w8a8 is supported.☆11Sep 10, 2024Updated last year
- Port of Funasr's Paraformer model in C/C++☆43Jun 19, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- HuggingFace Transformers WordPiece Tokenizer in C++☆22Mar 14, 2025Updated last year
- Minimal example of using a traced huggingface transformers model with libtorch☆36Sep 17, 2020Updated 5 years ago
- Converting Chinese sentences into pinyin sequences, implemented in C++, very fast and easy to deploy.☆23Jan 5, 2026Updated 6 months ago
- A SQLite extension for working with float and binary vectors. Work in progress!☆24Feb 10, 2023Updated 3 years ago
- BERT Tokenizer in C++☆79Jan 14, 2021Updated 5 years ago
- implement bert in pure c++☆37Apr 29, 2020Updated 6 years ago
- OneFlow Serving☆20Apr 10, 2025Updated last year
- GPU accelerated client-side embeddings for vector search, RAG etc.☆65Dec 4, 2023Updated 2 years ago
- Run Chinese MobileBert model on SNPE.☆15May 19, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Recording models☆12Sep 19, 2023Updated 2 years ago
- Fast and customizable text tokenization library with BPE and SentencePiece support☆334Jan 10, 2026Updated 6 months ago
- 不依赖 Go-Spring 框架的 Web 模块☆14Aug 8, 2020Updated 5 years ago
- ☆24Apr 2, 2026Updated 3 months ago
- Another reverse proxy that provides authentication with OpenID Connect☆10Jul 10, 2023Updated 3 years ago
- qwen2 and llama3 cpp implementation☆50Jun 7, 2024Updated 2 years ago
- ☆150Jan 9, 2025Updated last year
- Datastore for Tensors based on Xarray and Zarr☆18Sep 23, 2025Updated 9 months ago
- C++ SDK for Milvus☆53Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for our paper "Evaluating SIMD Compiler-Intrinsics for Database Systems"☆16Jul 5, 2023Updated 3 years ago
- ☆23Aug 14, 2024Updated last year
- NVIDIA TensorRT Hackathon 2023复赛选题:通义千问Qwen-7B用TensorRT-LLM模型搭建及优化☆43Oct 20, 2023Updated 2 years ago
- 基于 Sherpa-ONNX 实现在线下载模型的端侧实时语音识别应用(Implement speech recognition based on Sherpa-ONNX by downloading the model online.)☆29Feb 27, 2025Updated last year
- TTG: Template Task Graph C++ API☆26May 9, 2026Updated 2 months ago
- ☆13Nov 27, 2025Updated 7 months ago
- Parallel Wavelet Tree and Wavelet Matrix Construction☆26Jun 27, 2023Updated 3 years ago
- Sequence algorithms for use in Flashlight.☆14Jan 12, 2026Updated 6 months ago
- Inference RWKV v5, v6 and v7 with Qualcomm AI Engine Direct SDK☆94Jun 8, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆34Jul 23, 2024Updated last year
- Graph model execution API for Candle☆18Jul 27, 2025Updated 11 months ago
- c++实现的clip推理,模型有一点点改动,但是不大,改动和导出模型的代码可以在readme里找到,模型文件都在Releases里,包括AX650的模型。新增支持ChineseCLIP☆31Jun 19, 2025Updated last year
- Supporting code for "LLMs for your iPhone: Whole-Tensor 4 Bit Quantization"☆11Mar 31, 2024Updated 2 years ago
- Simple inference for Vits2 TTS Using ONNXRUNTIME and espeak-ng on C++☆19Apr 17, 2024Updated 2 years ago
- Tunnel is a Pipeline Execution Engine based on C++20 coroutine☆29Aug 17, 2023Updated 2 years ago
- A tool convert TensorRT engine/plan to a fake onnx☆41Nov 22, 2022Updated 3 years ago