Demonstration of running a native LLM on Android device.
☆260Aug 14, 2026Updated 2 weeks ago
Alternatives and similar repositories for Native-LLM-for-Android
Users that are interested in Native-LLM-for-Android are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Export YOLO Series to ONNX format.☆72Aug 23, 2026Updated last week
- Utilizes ONNX Runtime to transcribe audio into text.☆87Updated this week
- Utilizes ONNX Runtime for TTS model.☆70Updated this week
- 基于MNN-llm的安卓手机部署大语言模型:Qwen1.5-0.5B-Chat☆98Apr 8, 2024Updated 2 years ago
- Utilizes ONNX Runtime for speech activity detection.☆49Aug 17, 2026Updated 2 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Utilizes ONNX Runtime for audio denoising.☆137Aug 15, 2026Updated 2 weeks ago
- Android本地运行mnn-llm语言模型简单示例☆13Oct 2, 2025Updated 10 months ago
- Demonstrate Yolov9 model with Qualcomm Hexagon NPU and DirectML☆12Nov 27, 2024Updated last year
- ☆18Dec 7, 2023Updated 2 years ago
- Export the STFT or ISTFT process in ONNX format.☆45Jun 6, 2026Updated 2 months ago
- Transcribe subtitles and translate them offline with ease.☆47Jun 15, 2026Updated 2 months ago
- Fast Multimodal LLM on Mobile Devices☆1,600Aug 19, 2026Updated last week
- llm-export can export llm model to onnx.☆356May 8, 2026Updated 3 months ago
- Large Language Model Onnx Inference Framework☆35Nov 25, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- QAI AppBuilder is designed to help developers easily execute models on WoS and Linux platforms. It encapsulates the Qualcomm® AI Runtime …☆200Updated this week
- llm deploy project based onnx.☆49Oct 9, 2024Updated last year
- IRIS is an android app for interfacing with GGUF / llama.cpp models locally.☆290Feb 1, 2025Updated last year
- YOLOv12 TensorRT 端到端模型加速推理和INT8量化实现☆14Mar 5, 2025Updated last year
- Run pytorch models on GPU Android with Vulkan backend☆10Aug 15, 2023Updated 3 years ago
- MeloTTS demo on Axera☆14Jul 1, 2026Updated last month
- 安卓手机部署DeepSeek-R1 蒸馏的1.5B模型☆25Feb 4, 2025Updated last year
- Run local LLMs like Gemma, Qwen, and LLaMA on Android for offline, private, real-time chat and question answering with LiteRT and ONNX Ru…☆400Aug 13, 2026Updated 2 weeks ago
- llm deploy project based mnn. This project has merged into MNN.☆1,617Jan 20, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Running the F5-TTS by ONNX Runtime☆208Jul 29, 2026Updated last month
- Whisper in TensorRT-LLM☆16Sep 21, 2023Updated 2 years ago
- A prototype mobile text editor with Language Server Protocol support☆16Mar 19, 2021Updated 5 years ago
- mnn asr demo.☆27Mar 24, 2025Updated last year
- Inference RWKV v5, v6 and v7 with Qualcomm AI Engine Direct SDK☆96Jul 27, 2026Updated last month
- Create ONNX☆22Apr 13, 2026Updated 4 months ago
- ffmpeg+cuvid+tensorrt+multicamera☆12Dec 31, 2024Updated last year
- Memory Cleaner, Phone Booster and Optimizer.☆11Nov 20, 2018Updated 7 years ago
- ☆34Jul 23, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- some hexagon intrinsic examples based on Qualcomm Hexagon☆18Mar 7, 2025Updated last year
- The cpp-based deployment of CosyVoice2☆20Sep 7, 2025Updated 11 months ago
- 参考自mlc-llm,个人尝试在android手机上部署大模型并运行☆92Aug 5, 2024Updated 2 years ago
- Running the F5-TTS by ONNX Runtime standalone with GUI☆27Dec 10, 2024Updated last year
- PersonaTalk Hack☆15Jan 10, 2025Updated last year
- A lightweight vLLM-style inference engine in Rust☆153Updated this week
- ☆24Jul 17, 2024Updated 2 years ago