guillaume-be / rust-tokenizersLinks
Rust-tokenizer offers high-performance tokenizers for modern language models, including WordPiece, Byte-Pair Encoding (BPE) and Unigram (SentencePiece) models
☆326Updated 2 years ago
Alternatives and similar repositories for rust-tokenizers
Users that are interested in rust-tokenizers are comparing it to the libraries listed below
Sorting:
- Rust language bindings for Faiss☆234Updated last week
- Rust port of sentence-transformers (https://github.com/UKPLab/sentence-transformers)☆121Updated last year
- Rust wrapper for Microsoft's ONNX Runtime (version 1.8)☆307Updated last year
- Rust client for the huggingface hub aiming for minimal subset of features over `huggingface-hub` python package☆232Updated this week
- finalfusion embeddings in Rust☆103Updated last year
- Inference Llama 2 in one file of pure Rust 🦀☆234Updated 2 years ago
- HNSW ANN from the paper "Efficient and robust approximate nearest neighbor search using Hierarchical Navigable Small World graphs"☆245Updated last month
- Locality Sensitive Hashing in Rust with Python bindings☆118Updated 2 years ago
- Fast approximate nearest neighbor searching in Rust, based on HNSW index☆334Updated 3 weeks ago
- Rust implementation of the HNSW algorithm (Malkov-Yashunin)☆211Updated 3 weeks ago
- Example of tch-rs on M1☆55Updated last year
- fastText Rust binding☆62Updated last year
- pure rust implemention of word2vec☆85Updated 2 years ago
- Tutorial for Porting PyTorch Transformer Models to Candle (Rust)☆319Updated last year
- Low rank adaptation (LoRA) for Candle.☆162Updated 5 months ago
- Neural syntax annotator, supporting sequence labeling, lemmatization, and dependency parsing.☆78Updated last year
- pgvector support for Rust☆179Updated 3 weeks ago
- Llama2 LLM ported to Rust burn☆279Updated last year
- Rust library for generating vector embeddings, reranking. Re-write of qdrant/fastembed.☆617Updated 2 weeks ago
- Rust client for Qdrant vector search engine☆337Updated 3 weeks ago
- A rust implementation of some popular snowball stemming algorithms☆129Updated last year
- A Rust implementation of OpenAI's Whisper model using the burn framework☆324Updated last year
- Models and examples built with Burn☆286Updated 2 weeks ago
- Andrej Karpathy's Let's build GPT: from scratch video & notebook implemented in Rust + candle☆76Updated last year
- A machine learning library for Rust.☆329Updated last year
- An Approximate Nearest Neighbors library in Rust, based on random projections and LMDB and optimized for memory usage☆286Updated this week
- LLM Orchestrator built in Rust☆284Updated last year
- ☆33Updated 10 months ago
- The most accurate natural language detection library for Rust, suitable for short text and mixed-language text☆996Updated last week
- Ready-made tokenizer library for working with GPT and tiktoken☆338Updated this week