☆72Feb 27, 2023Updated 3 years ago
Alternatives and similar repositories for huggingface-tokenizer-in-cxx
Users that are interested in huggingface-tokenizer-in-cxx are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Universal cross-platform tokenizers binding to HF and sentencepiece☆512May 20, 2026Updated 4 months ago
- C++ implementation of tokenizers, including tiktoken.☆25Dec 7, 2023Updated 2 years ago
- transformer tokenizers (e.g. BERT tokenizer) in C++ (WIP)☆18Apr 7, 2022Updated 4 years ago
- Try to export the ONNX QDQ model that conforms to the AXERA NPU quantization specification. Currently, only w8a8 is supported.☆11Sep 10, 2024Updated 2 years ago
- ONNX-compatible DocShadow: High-Resolution Document Shadow Removal. Supports TensorRT 🚀☆27Sep 13, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- HuggingFace Transformers WordPiece Tokenizer in C++☆22Mar 14, 2025Updated last year
- Minimal example of using a traced huggingface transformers model with libtorch☆36Sep 17, 2020Updated 6 years ago
- Source code of our implementation of the concurrent RMA☆12May 23, 2019Updated 7 years ago
- A four-dimensional Analysis of Partitioned Approximate Filters☆11Aug 6, 2025Updated last year
- Tutorials of Extending and importing TVM with CMAKE Include dependency.☆16Oct 11, 2024Updated last year
- A SQLite extension for working with float and binary vectors. Work in progress!☆24Feb 10, 2023Updated 3 years ago
- Port of Funasr's Paraformer model in C/C++☆43Jun 19, 2024Updated 2 years ago
- BERT Tokenizer in C++☆79Jan 14, 2021Updated 5 years ago
- C inference engine for running GLiClass (Generalist and Lightweight Classification) models☆17May 21, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Grizzly: Efficient Stream Processing Through Adaptive Query Compilation☆17Jun 13, 2020Updated 6 years ago
- implement bert in pure c++☆37Apr 29, 2020Updated 6 years ago
- OneFlow Serving☆20Apr 10, 2025Updated last year
- ☆45Sep 5, 2026Updated 2 weeks ago
- GPU accelerated client-side embeddings for vector search, RAG etc.☆64Dec 4, 2023Updated 2 years ago
- Vector functions and indexing for SQLite☆10Mar 26, 2023Updated 3 years ago
- A simple REPL for Lean 4, returning information about errors and sorries.☆12Jun 19, 2023Updated 3 years ago
- Fast and customizable text tokenization library with BPE and SentencePiece support☆339Jan 10, 2026Updated 8 months ago
- Uses the excellent silero VAD with onnxruntime C api for fast detection of audio segments with speech☆17Sep 20, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆26Apr 2, 2026Updated 5 months ago
- qwen2 and llama3 cpp implementation☆50Jun 7, 2024Updated 2 years ago
- C++ SDK for Milvus☆53Updated this week
- ☆18Dec 7, 2023Updated 2 years ago
- ☆16Mar 16, 2021Updated 5 years ago
- Code for our paper "Evaluating SIMD Compiler-Intrinsics for Database Systems"☆16Jul 5, 2023Updated 3 years ago
- ☆23Aug 14, 2024Updated 2 years ago
- Triton backend for https://github.com/OpenNMT/CTranslate2☆35Jul 7, 2023Updated 3 years ago
- Fast Cardinality Estimation of Multi-Join Queries Using Sketches☆16Feb 29, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for EMNLP 2023 paper: DALE: Generative Data Augmentation for Low-Resource Legal NLP☆11Oct 27, 2023Updated 2 years ago
- ☆34Apr 29, 2019Updated 7 years ago
- NVIDIA TensorRT Hackathon 2023复赛选题:通义千问Qwen-7B用TensorRT-LLM模型搭建及优化☆43Oct 20, 2023Updated 2 years ago
- ☆13Nov 27, 2025Updated 9 months ago
- Inference RWKV v5, v6 and v7 with Qualcomm AI Engine Direct SDK☆99Jul 27, 2026Updated last month
- ☆34Jul 23, 2024Updated 2 years ago
- Repository for "Propagating Knowledge Updates to LMs Through Distillation" (NeurIPS 2023).☆27Aug 25, 2024Updated 2 years ago