JS tokenizer for LLaMA 1 and 2
☆362Jun 27, 2024Updated 2 years ago
Alternatives and similar repositories for llama-tokenizer-js
Users that are interested in llama-tokenizer-js are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tensor library for machine learning☆273Apr 23, 2023Updated 3 years ago
- A toolbox for working with WebRTC, Audio and AI☆702Jul 29, 2023Updated 3 years ago
- A simple no-install web UI for Ollama and OAI-Compatible APIs!☆31Jan 30, 2025Updated last year
- Simple UI for LLM Model Finetuning☆2,054Dec 21, 2023Updated 2 years ago
- A fast inference library for running LLMs locally on modern consumer-class GPUs☆4,603Mar 4, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A simple "Be My Eyes" web app with a llama.cpp/llava backend☆496Nov 28, 2023Updated 2 years ago
- Falcon LLM ggml framework with CPU and GPU support☆250Jul 2, 2026Updated last month
- ☆16Dec 16, 2024Updated last year
- A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.☆2,937Sep 30, 2023Updated 2 years ago
- First instruction-tuning dataset distilled from Claude2 (52k Alpaca prompts)!☆13Oct 22, 2023Updated 2 years ago
- Finetune llama2-70b and codellama on MacBook Air without quantization☆449Mar 28, 2024Updated 2 years ago
- ☆17Dec 18, 2023Updated 2 years ago
- Bringing stable diffusion models to web browsers. Everything runs inside the browser with no server support.☆3,720Mar 12, 2024Updated 2 years ago
- High-performance In-browser LLM Inference Engine☆18,541Aug 4, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 🌸 Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading☆10,485Sep 7, 2024Updated last year
- OpenLLaMA, a permissively licensed open source reproduction of Meta AI’s LLaMA 7B trained on the RedPajama dataset☆7,532Jul 16, 2023Updated 3 years ago
- Generate High Quality textual or multi-modal datasets with Agents☆18Jun 7, 2023Updated 3 years ago
- 📊 x 📊 =❓An online gallery to explore the design space of comparative layouts☆18Nov 4, 2020Updated 5 years ago
- AI-managed code blocks in Python ⏪⏩☆465Oct 5, 2023Updated 2 years ago
- Llama 2 Everywhere (L2E)☆1,524Aug 27, 2025Updated 11 months ago
- code for training and using chess embeddings models☆14Jun 9, 2024Updated 2 years ago
- Visual Studio Code extension for WizardCoder☆148Aug 1, 2023Updated 3 years ago
- Quantized inference code for LLaMA models☆1,037Mar 17, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Prolog implemented in Python☆12Sep 6, 2024Updated last year
- A guidance language for controlling large language models.☆21,709May 21, 2026Updated 2 months ago
- Proof of concept for a generative AI application framework powered by WebAssembly and Extism☆14Aug 10, 2023Updated 3 years ago
- GPT in TensorFlow.js☆33Oct 16, 2023Updated 2 years ago
- LLMs build upon Evol Insturct: WizardLM, WizardCoder, WizardMath☆9,483Jun 7, 2025Updated last year
- Basaran is an open-source alternative to the OpenAI text completion API. It provides a compatible streaming API for your Hugging Face Tra…☆1,283Jan 24, 2024Updated 2 years ago
- A Bulletproof Way to Generate Structured JSON from Language Models☆4,935Feb 24, 2024Updated 2 years ago
- Browser-compatible JS library for running language models☆233Aug 25, 2022Updated 3 years ago
- Highly concurrent and fast content processing for Mighty Inference Server☆10Feb 6, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- simple ansible playbook to take clean ubuntu 18.04 to CUDA 10, PyTorch 1.0, fastai, miniconda heaven☆11Dec 16, 2018Updated 7 years ago
- State-of-the-art Machine Learning for the web. Run 🤗 Transformers directly in your browser, with no need for a server!☆16,241Jul 31, 2026Updated last week
- LLMFlows - Simple, Explicit and Transparent LLM Apps☆707Feb 20, 2025Updated last year
- Convenient wrapper for fine-tuning and inference of Large Language Models (LLMs) with several quantization techniques (GTPQ, bitsandbytes…☆143Oct 17, 2023Updated 2 years ago
- Text-To-Speech, RAG, and LLMs. All local!☆1,910Dec 9, 2024Updated last year
- Fork of Facebooks LLaMa model to run on CPU☆766Mar 6, 2023Updated 3 years ago
- Lightweight inference library for ONNX files, written in C++. It can run Stable Diffusion XL 1.0 on a RPI Zero 2 (or in 298MB of RAM) but…☆2,086Jun 18, 2026Updated last month