guillaume-be / rust-tokenizers
Rust-tokenizer offers high-performance tokenizers for modern language models, including WordPiece, Byte-Pair Encoding (BPE) and Unigram (SentencePiece) models
☆296Updated last year
Related projects ⓘ
Alternatives and complementary repositories for rust-tokenizers
- Rust wrapper for Microsoft's ONNX Runtime (version 1.8)☆283Updated 8 months ago
- Rust implementation of the HNSW algorithm (Malkov-Yashunin)☆158Updated 4 months ago
- Rust port of sentence-transformers (https://github.com/UKPLab/sentence-transformers)☆106Updated 2 months ago
- Rust client for the huggingface hub aiming for minimal subset of features over `huggingface-hub` python package☆153Updated 2 months ago
- finalfusion embeddings in Rust☆93Updated last year
- Rust language bindings for Faiss☆199Updated 2 months ago
- Inference Llama 2 in one file of pure Rust 🦀☆229Updated last year
- A machine learning library for Rust.☆314Updated 3 months ago
- Neural syntax annotator, supporting sequence labeling, lemmatization, and dependency parsing.☆71Updated last year
- Fast approximate nearest neighbor searching in Rust, based on HNSW index☆313Updated this week
- Llama2 LLM ported to Rust burn☆274Updated 7 months ago
- A Rust implementation of OpenAI's Whisper model using the burn framework☆270Updated 6 months ago
- pure rust implemention of word2vec☆80Updated last year
- Rust client for Qdrant vector search engine☆232Updated last month
- 🏆 A ranked list of awesome machine learning Rust libraries.☆307Updated this week
- fastText Rust binding☆57Updated 10 months ago
- Models and examples built with Burn☆185Updated this week
- Low rank adaptation (LoRA) for Candle.☆127Updated 3 months ago
- Tutorial for Porting PyTorch Transformer Models to Candle (Rust)☆252Updated 3 months ago
- LLaMa 7b with CUDA acceleration implemented in rust. Minimal GPU memory needed!☆101Updated last year
- A rust implementation of some popular snowball stemming algorithms☆114Updated 6 months ago
- pgvector support for Rust☆121Updated last week
- High-level, optionally asynchronous Rust bindings to llama.cpp☆179Updated 5 months ago
- Example of tch-rs on M1☆51Updated 8 months ago
- Library for generating vector embeddings, reranking in Rust☆285Updated this week
- HNSW ANN from the paper "Efficient and robust approximate nearest neighbor search using Hierarchical Navigable Small World graphs"☆226Updated 11 months ago
- Tensors and differentiable operations (like TensorFlow) in Rust☆485Updated last year
- Ready-made tokenizer library for working with GPT and tiktoken☆258Updated last week
- ONNX neural network inference engine☆124Updated this week
- A comprehensive library for machine learning and numerical computing. The library provides a set of tools for linear algebra, numerical c…☆707Updated 3 months ago