LiteRT, successor to TensorFlow Lite. is Google's On-device framework for high-performance ML & GenAI deployment on edge platforms, via efficient conversion, runtime, and optimization
☆3,221Jul 28, 2026Updated this week
Alternatives and similar repositories for LiteRT
Users that are interested in LiteRT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Support PyTorch model conversion with LiteRT.☆1,075Updated this week
- LiteRT-LM is Google's production-ready, high-performance, open-source inference framework for deploying Large Language Models on edge dev…☆6,034Updated this week
- ☆382Updated this week
- AI Edge Quantizer: flexible post training quantization for LiteRT models.☆185Updated this week
- On-device AI across mobile, embedded and edge for PyTorch☆4,838Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A gallery that showcases on-device ML/GenAI use cases and allows people to try and use models locally.☆24,301Updated this week
- A convenient CLI to streamline LiteRT related development workflows, including converting, quantizing, compiling, managing, running, benc…☆35Jun 18, 2026Updated last month
- Official inference framework for 1-bit LLMs☆39,785Updated this week
- ☆15,926Updated this week
- ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator☆21,210Updated this week
- Cross-platform, customizable ML solutions for live and streaming media.☆36,365Jul 17, 2026Updated last week
- OpenRAG is a comprehensive, single package Retrieval-Augmented Generation platform built on Langflow, Docling, and Opensearch.☆4,380Updated this week
- Quantization, kernels, runtime and inference engine for mobiles, wearables, smart home and robots.☆5,543Updated this week
- ☆24Jan 22, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Qualcomm® AI Hub Models is our collection of state-of-the-art machine learning models optimized for performance (latency, memory etc.) an…☆1,170Updated this week
- High-efficiency floating-point neural network inference operators for mobile, server, and Web☆2,409Updated this week
- LLM inference in C/C++☆121,878Updated this week
- Tensor library for machine learning☆15,071Jul 17, 2026Updated last week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆87,317Updated this week
- Unsloth is a local UI for training and running Gemma 4, Qwen3.6, DeepSeek, Kimi, GLM and other models.☆68,965Updated this week
- Infrastructure to enable deployment of ML models to low-power resource-constrained embedded targets (including microcontrollers and digit…☆3,020Updated this week
- Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces☆10,504Updated this week
- LMCache: Supercharge Your LLM with the Fastest KV Cache Layer☆10,922Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- JavaScript in-page GUI agent. Control web interfaces with natural language.☆28,055Updated this week
- Hexagon-MLIR is a compiler toolchain for compiling and executing AI kernels and models on Qualcomm Hexagon Neural Processing Units (NPUs)…☆180Jul 2, 2026Updated 3 weeks ago
- A lightweight, lightning-fast, in-process vector database☆15,298Updated this week
- Open-Source Frontier Voice AI☆50,834Updated this week
- MLX: An array framework for Apple silicon☆27,734Updated this week
- The Qualcomm® AI Hub apps are a collection of state-of-the-art machine learning models optimized for performance (latency, memory etc.) a…☆440Updated this week
- SGLang is a high-performance serving framework for large language models and multimodal models.☆30,854Updated this week
- OpenVINO™ is an open source toolkit for optimizing and deploying AI inference☆10,585Updated this week
- A modern model graph visualizer and debugger☆1,529Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Run frontier AI locally.☆46,507Jun 23, 2026Updated last month
- Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.☆13,532Updated this week
- Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime…☆13,837Updated this week
- Port of OpenAI's Whisper model in C/C++☆52,364Jul 11, 2026Updated 2 weeks ago
- This repository is a read-only mirror of https://gitlab.arm.com/kleidi/kleidiai☆174Updated this week
- Lightpanda: the headless browser designed for AI and automation☆32,802Updated this week
- The all-in-one, open-source backend platform for agentic coding. InsForge gives your coding agent database, auth, storage, compute, hosti…☆12,507Updated this week