LiteRT, successor to TensorFlow Lite. is Google's On-device framework for high-performance ML & GenAI deployment on edge platforms, via efficient conversion, runtime, and optimization
☆3,312Aug 17, 2026Updated this week
Alternatives and similar repositories for LiteRT
Users that are interested in LiteRT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Support PyTorch model conversion with LiteRT.☆1,080Updated this week
- LiteRT-LM is Google's production-ready, high-performance, open-source inference framework for deploying Large Language Models on edge dev…☆6,224Updated this week
- LiteRT and LiteRT-LM sample apps, model recipes, agent skills and utilities.☆398Updated this week
- AI Edge Quantizer: flexible post training quantization for LiteRT models.☆190Updated this week
- On-device AI across mobile, embedded and edge for PyTorch☆4,923Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A gallery that showcases on-device ML/GenAI use cases and allows people to try and use models locally.☆24,471Updated this week
- A convenient CLI to streamline LiteRT related development workflows, including converting, quantizing, compiling, managing, running, benc…☆39Updated this week
- Official inference framework for 1-bit LLMs☆40,097Jul 27, 2026Updated 3 weeks ago
- ☆16,137Updated this week
- ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator☆21,396Updated this week
- Cross-platform, customizable ML solutions for live and streaming media.☆36,638Updated this week
- OpenRAG is a comprehensive, single package Retrieval-Augmented Generation platform built on Langflow, Docling, and Opensearch.☆4,431Updated this week
- Quantization, kernels, runtime and inference engine for mobiles, wearables, smart home and robots.☆5,802Updated this week
- ☆25Jan 22, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Qualcomm® AI Hub Models is our collection of state-of-the-art machine learning models optimized for performance (latency, memory etc.) an…☆1,187Updated this week
- High-efficiency floating-point neural network inference operators for mobile, server, and Web☆2,424Updated this week
- LLM inference in C/C++☆124,338Updated this week
- Tensor library for machine learning☆15,185Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆89,264Updated this week
- Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.☆73,218Updated this week
- Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces☆10,876Updated this week
- Infrastructure to enable deployment of ML models to low-power resource-constrained embedded targets (including microcontrollers and digit…☆3,048Updated this week
- LMCache: Supercharge Your LLM with the Fastest KV Cache Layer☆11,185Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- JavaScript in-page GUI agent. Control web interfaces with natural language.☆28,681Updated this week
- Hexagon-MLIR is a compiler toolchain for compiling and executing AI kernels and models on Qualcomm Hexagon Neural Processing Units (NPUs)…☆189Jul 2, 2026Updated last month
- A lightweight, lightning-fast, in-process vector database☆15,453Updated this week
- Open-Source Frontier Voice AI☆52,810Jul 24, 2026Updated 3 weeks ago
- MLX: An array framework for Apple silicon☆28,007Updated this week
- The Qualcomm® AI Hub apps are a collection of state-of-the-art machine learning models optimized for performance (latency, memory etc.) a…☆442Updated this week
- Fast Multimodal LLM on Mobile Devices☆1,587Updated this week
- SGLang is a high-performance serving framework for large language models and multimodal models.☆31,961Updated this week
- OpenVINO™ is an open source toolkit for optimizing and deploying AI inference☆10,666Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A modern model graph visualizer and debugger☆1,535Updated this week
- Run frontier AI locally.☆46,862Jun 23, 2026Updated last month
- Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.☆13,689Jul 24, 2026Updated 3 weeks ago
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆14,216Updated this week
- Port of OpenAI's Whisper model in C/C++☆52,957Updated this week
- This repository is a read-only mirror of https://gitlab.arm.com/kleidi/kleidiai☆180Updated this week
- Lightpanda: the headless browser designed for AI and automation☆33,993Updated this week