Explore LLM model deployment based on AXera's AI chips, It provides OpenAI‑compatible APIs and supports AX620E, AX650 and AX637 series chips.
☆166Sep 7, 2026Updated this week
Alternatives and similar repositories for ax-llm
Users that are interested in ax-llm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Samples code for world class Artificial Intelligence SoCs for computer vision applications.☆300Aug 10, 2026Updated 3 weeks ago
- The docs repository of Pulsar2 which is AXera's SoC 2rd AI toolchain. Such as AX650A, AX650N☆19Updated this week
- OpenAI Whisper demo on Axera☆17Jan 15, 2026Updated 7 months ago
- the python api for axengine runtime☆27Mar 24, 2026Updated 5 months ago
- ☆29Jun 30, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Demo for Qwen2.5-VL-3B-Instruct on Axera device.☆16Sep 3, 2025Updated last year
- Python scripts performing Open Vocabulary Object Detection using the YOLO-World model in ONNX. And Export the ONNX model for AXera's NPU☆12Aug 11, 2025Updated last year
- ☆72May 25, 2026Updated 3 months ago
- Converts CLIP models to ONNX☆11Jan 17, 2023Updated 3 years ago
- Try to export the ONNX QDQ model that conforms to the AXERA NPU quantization specification. Currently, only w8a8 is supported.☆11Sep 10, 2024Updated last year
- linux bsp app & sample for axpi (ax620a)☆36Jun 21, 2023Updated 3 years ago
- ☆23Jan 3, 2024Updated 2 years ago
- About Samples code for Axera's PCIE Card for computer vision applications.☆20Aug 10, 2026Updated 3 weeks ago
- MeloTTS demo on Axera☆14Jul 1, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Linux BSP APP & Samples for AXera Pi Zero(AX620Q)☆26Nov 1, 2024Updated last year
- SAM and lama inpaint,包含QT的GUI交互界面,实现了交互式可实时显示结果的画点、画框进行SAM,然后通过进行Inpaint,具体操作看readme里的视频。☆54Jan 30, 2024Updated 2 years ago
- The Pipeline example based on AX650N/AX8850 shows the software development skills of Image Processing, NPU, Codec, and Display modules, …☆23Updated this week
- ☆126Dec 15, 2023Updated 2 years ago
- Whisper in TensorRT-LLM☆16Sep 21, 2023Updated 2 years ago
- M5Stack dropdown menu code sample☆15Feb 5, 2019Updated 7 years ago
- c++实现的clip推理,模型有一点点改动,但 是不大,改动和导出模型的代码可以在readme里找到,模型文件都在Releases里,包括AX650的模型。新增支持ChineseCLIP☆31Jun 19, 2025Updated last year
- ☆18Dec 7, 2023Updated 2 years ago
- ☆17Aug 15, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- molly, an LLM designed to understand multi-omics data.☆26Dec 23, 2025Updated 8 months ago
- A SDK to using the Realtime API with Microcontrollers like the ESP32☆24Dec 24, 2024Updated last year
- stable diffusion using mnn☆68Sep 28, 2023Updated 2 years ago
- Multiple GEMM operators are constructed with cutlass to support LLM inference.☆20Aug 3, 2025Updated last year
- ☆1,666Jun 17, 2026Updated 2 months ago
- MegEngine到其他框架的转换器☆71Apr 27, 2023Updated 3 years ago
- ncnn和pnnx格式编辑器☆148Jul 15, 2026Updated last month
- linux bsp app & sample for axpi pro (ax650n)☆32Nov 12, 2024Updated last year
- Sample projects for InferenceHelper, a Helper Class for Deep Learning Inference Frameworks: TensorFlow Lite, TensorRT, OpenCV, ncnn, MNN,…☆22Mar 27, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- llm-export can export llm model to onnx.☆356May 8, 2026Updated 3 months ago
- Inference RWKV v5, v6 and v7 with Qualcomm AI Engine Direct SDK☆98Jul 27, 2026Updated last month
- ArduPlane, ArduCopter, ArduRover source☆10Updated this week
- LLaMa/RWKV onnx models, quantization and testcase☆367Jul 6, 2023Updated 3 years ago
- ☆17Jan 1, 2024Updated 2 years ago
- 3D rendering by M5Stack☆26May 31, 2026Updated 3 months ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆17Jun 3, 2024Updated 2 years ago