AXERA-TECH / ax-llmView external linksLinks
Explore LLM model deployment based on AXera's AI chips
☆139Feb 6, 2026Updated last week
Alternatives and similar repositories for ax-llm
Users that are interested in ax-llm are comparing it to the libraries listed below
Sorting:
- Samples code for world class Artificial Intelligence SoCs for computer vision applications.☆283Jan 20, 2026Updated 3 weeks ago
- ☆28Jun 30, 2025Updated 7 months ago
- Demo for Qwen2.5-VL-3B-Instruct on Axera device.☆17Sep 3, 2025Updated 5 months ago
- the python api for axengine runtime☆25Dec 4, 2025Updated 2 months ago
- linux bsp app & sample for axpi (ax620a)☆36Jun 21, 2023Updated 2 years ago
- The docs repository of Pulsar2 which is AXera's SoC 2rd AI toolchain. Such as AX650A, AX650N☆17Updated this week
- Linux BSP APP & Samples for AXera Pi Zero(AX620Q)☆21Nov 1, 2024Updated last year
- Python scripts performing Open Vocabulary Object Detection using the YOLO-World model in ONNX. And Export the ONNX model for AXera's NPU☆12Aug 11, 2025Updated 6 months ago
- Converts CLIP models to ONNX☆11Jan 17, 2023Updated 3 years ago
- ☆23Jan 3, 2024Updated 2 years ago
- Multiple GEMM operators are constructed with cutlass to support LLM inference.☆20Aug 3, 2025Updated 6 months ago
- MegEngine到其他框架的转换器☆69Apr 27, 2023Updated 2 years ago
- Arduino library for M5Stack LLM Module☆33Oct 29, 2025Updated 3 months ago
- Try to export the ONNX QDQ model that conforms to the AXERA NPU quantization specification. Currently, only w8a8 is supported.☆11Sep 10, 2024Updated last year
- ArduPlane, ArduCopter, ArduRover source☆10Updated this week
- A SDK to using the Realtime API with Microcontrollers like the ESP32☆23Dec 24, 2024Updated last year
- c++实现的clip推理,模型有一点点改动,但是不大,改动和导出模型的代码可以在readme里找到,模型文件都在Releases里,包括AX650的模型。新增支持ChineseCLIP☆31Jun 19, 2025Updated 7 months ago
- ☆16Nov 6, 2025Updated 3 months ago
- ☆125Dec 15, 2023Updated 2 years ago
- llm-export can export llm model to onnx.☆343Oct 24, 2025Updated 3 months ago
- 3D rendering by M5Stack☆26Apr 8, 2019Updated 6 years ago
- Large Language Model Onnx Inference Framework☆35Nov 25, 2025Updated 2 months ago
- Benchmark tests supporting the TiledCUDA library.☆18Nov 19, 2024Updated last year
- M5Stack dropdown menu code sample☆15Feb 5, 2019Updated 7 years ago
- ggml学习笔记,ggml是一个机器学习的推理框架☆18Mar 24, 2024Updated last year
- ☆1,232Nov 24, 2025Updated 2 months ago
- Example of SenseCraft Model Assistant Model deployment related to ESP32☆32Apr 9, 2025Updated 10 months ago
- ☆24Jun 27, 2023Updated 2 years ago
- TiledLower is a Dataflow Analysis and Codegen Framework written in Rust.☆14Nov 23, 2024Updated last year
- A high-throughput and memory-efficient inference and serving engine for LLMs☆17Jun 3, 2024Updated last year
- YOLOv5在高通AI Engine Direct环境下进行QNN量化,CPU推理的项目☆16Sep 10, 2024Updated last year
- Axera download protocol (AXDL) implementation in Rust (Unofficial)☆16Apr 4, 2025Updated 10 months ago
- a single-header math library☆17Nov 7, 2025Updated 3 months ago
- linux bsp app & sample for axpi pro (ax650n)☆30Nov 12, 2024Updated last year
- ☆97Mar 26, 2025Updated 10 months ago
- Quantized Attention on GPU☆44Nov 22, 2024Updated last year
- MegCC是一个运行时超轻量,高效,移植简单的深度学习模型编译器☆488Oct 23, 2024Updated last year
- llm deploy project based onnx.☆49Oct 9, 2024Updated last year
- A Triton JIT runtime and ffi provider in C++☆31Jan 26, 2026Updated 2 weeks ago