Support PyTorch model conversion with LiteRT.
☆1,077Aug 7, 2026Updated this week
Alternatives and similar repositories for litert-torch
Users that are interested in litert-torch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AI Edge Quantizer: flexible post training quantization for LiteRT models.☆188Updated this week
- LiteRT, successor to TensorFlow Lite. is Google's On-device framework for high-performance ML & GenAI deployment on edge platforms, via e…☆3,279Updated this week
- A tool for converting ONNX files to LiteRT/TFLite/TensorFlow, PyTorch native code (nn.Module), TorchScript (.pt), state_dict (.pt), Expor…☆987Aug 1, 2026Updated last week
- On-device AI across mobile, embedded and edge for PyTorch☆4,882Updated this week
- LiteRT and LiteRT-LM sample apps, model recipes, agent skills and utilities.☆392Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Qualcomm® AI Hub Models is our collection of state-of-the-art machine learning models optimized for performance (latency, memory etc.) an…☆1,181Updated this week
- Pytorch to Keras/Tensorflow/TFLite conversion made intuitive☆344Mar 10, 2025Updated last year
- A modern model graph visualizer and debugger☆1,533Updated this week
- LiteRT-LM is Google's production-ready, high-performance, open-source inference framework for deploying Large Language Models on edge dev…☆6,148Updated this week
- High-efficiency floating-point neural network inference operators for mobile, server, and Web☆2,422Updated this week
- Generative AI extensions for onnxruntime☆1,097Updated this week
- A gallery that showcases on-device ML/GenAI use cases and allows people to try and use models locally.☆24,395Updated this week
- TinyNeuralNetwork is an efficient and easy-to-use deep learning model compression framework.☆880Mar 3, 2026Updated 5 months ago
- AIMET is a library that provides advanced quantization and compression techniques for trained neural network models.☆2,675Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Example apps and demos using PyTorch's ExecuTorch framework☆81Updated this week
- ☆202Jul 28, 2026Updated 2 weeks ago
- Fast Multimodal LLM on Mobile Devices☆1,584Updated this week
- ☆18Nov 30, 2023Updated 2 years ago
- The Qualcomm® AI Hub apps are a collection of state-of-the-art machine learning models optimized for performance (latency, memory etc.) a…☆444Updated this week
- Infrastructure to enable deployment of ML models to low-power resource-constrained embedded targets (including microcontrollers and digit…☆3,040Updated this week
- onnxruntime-qnn is the Qualcomm AI Runtime (QAIRT) execution provider for onnxruntime. It provides onnxruntime hardware acceleration and …☆42Updated this week
- Tensorflow Backend for ONNX☆1,326Mar 28, 2024Updated 2 years ago
- Hexagon-MLIR is a compiler toolchain for compiling and executing AI kernels and models on Qualcomm Hexagon Neural Processing Units (NPUs)…☆185Jul 2, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 🤗 Optimum ExecuTorch☆135Updated this week
- ☆2,796Updated this week
- Backward compatible ML compute opset inspired by HLO/MHLO☆683Updated this week
- Demonstration of running a native LLM on Android device.☆259Jul 31, 2026Updated last week
- ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator☆21,335Updated this week
- Cross-platform, customizable ML solutions for live and streaming media.☆36,559Updated this week
- Simplify your onnx model☆4,381Updated this week
- This repository contains the official implementation of the research papers, "MobileCLIP" CVPR 2024 and "MobileCLIP2" TMLR August 2025☆1,615Apr 15, 2026Updated 3 months ago
- Convert tflite to JSON and make it editable in the IDE. It also converts the edited JSON back to tflite binary.☆28Feb 21, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆346Feb 12, 2026Updated 5 months ago
- MobileLLM Optimizing Sub-billion Parameter Language Models for On-Device Use Cases. In ICML 2024.☆1,456Apr 30, 2026Updated 3 months ago
- PyTorch native quantization and sparsity for training and inference☆2,934Updated this week
- SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, …☆2,696Updated this week
- This repository is a read-only mirror of https://gitlab.arm.com/kleidi/kleidiai☆176Updated this week
- On-device Neural Engine☆577Jul 24, 2026Updated 2 weeks ago
- QAI AppBuilder is designed to help developers easily execute models on WoS and Linux platforms. It encapsulates the Qualcomm® AI Runtime …☆194Updated this week