☆21Jul 22, 2022Updated 4 years ago
Alternatives and similar repositories for FastCNN
Users that are interested in FastCNN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆29Jun 30, 2025Updated last year
- Yet another Polyhedra Compiler for DeepLearning☆19Apr 14, 2023Updated 3 years ago
- Demo for Qwen2.5-VL-3B-Instruct on Axera device.☆16Sep 3, 2025Updated 11 months ago
- Tencent Distribution of TVM☆16Apr 7, 2023Updated 3 years ago
- Open deep learning compiler stack for cpu, gpu and specialized accelerators☆20Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆20Sep 28, 2024Updated last year
- VNEC: A Vectorized Non-Empty Column Format for SpMV on cross-platform multicore CPUs☆10Feb 6, 2024Updated 2 years ago
- A memory-centric profiling tool suite for heterogeneous memory☆10Nov 13, 2024Updated last year
- This is a demo how to write a high performance convolution run on apple silicon☆56Feb 8, 2022Updated 4 years ago
- ncnn export & infer mobileclip☆21Aug 18, 2025Updated last year
- Python scripts performing Open Vocabulary Object Detection using the YOLO-World model in ONNX. And Export the ONNX model for AXera's NPU☆12Aug 11, 2025Updated last year
- ☆10Aug 4, 2020Updated 6 years ago
- DDK for Rockchip NPU☆69Dec 29, 2020Updated 5 years ago
- Try to export the ONNX QDQ model that conforms to the AXERA NPU quantization specification. Currently, only w8a8 is supported.☆11Sep 10, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Profiling Spark Applications for Performance Comparison and Diagnosis☆16Nov 11, 2018Updated 7 years ago
- Code for reproducing work of ICML 2019 paper: Memory-Optimal Direct Convolutions for Maximizing Classification Accuracy in Embedded Appli…☆12Jun 8, 2019Updated 7 years ago
- An auxiliary project analysis of the characteristics of KV in DiT Attention.☆34Nov 29, 2024Updated last year
- ☆18Mar 18, 2024Updated 2 years ago
- Implementation of FusedMM method for IPDPS 2021 paper titled "FusedMM: A Unified SDDMM-SpMM Kernel for Graph Embedding and Graph Neural N…☆31Aug 12, 2022Updated 4 years ago
- SpV8 is a SpMV kernel written in AVX-512. Artifact for our SpV8 paper @ DAC '21.☆29Mar 16, 2021Updated 5 years ago
- ☆23Feb 18, 2025Updated last year
- Memory Sampling Tool using Linux perf_events☆15Aug 5, 2022Updated 4 years ago
- Lightweight performance and debugging tools☆17Feb 21, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A direct convolution library targeting ARM multi-core CPUs.☆12Nov 27, 2024Updated last year
- ☆32Jul 2, 2025Updated last year
- ncnn is a high-performance neural network inference framework optimized for the mobile platform☆14May 20, 2022Updated 4 years ago
- The Splash-4 benchmark suite☆15Oct 26, 2023Updated 2 years ago
- ☆17Aug 23, 2021Updated 5 years ago
- A framework and CLI toolkit for orchestrating teams of loosely-coupled AI agents.☆18Aug 9, 2026Updated 3 weeks ago
- Example implementation of the Alexa Voice Service Integration for AWS IoT Core for Arm Cortex-M series processors.☆12Feb 24, 2021Updated 5 years ago
- Quantized Attention on GPU☆45Nov 22, 2024Updated last year
- This a bridge for converting torch,and other AI training framework to C++ speed up infer library,like NCNN and ect☆20Mar 24, 2019Updated 7 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 使用NCNN推理框架和ByteTrack目标跟踪框架,对网络、文件流URL进行实时性视频推理,而UI界面则由Qt框架实现☆24Oct 16, 2024Updated last year
- Code for "Fast Sparse ConvNets" CVPR2020 submissions☆12Nov 20, 2019Updated 6 years ago
- Tutorials of Extending and importing TVM with CMAKE Include dependency.☆16Oct 11, 2024Updated last year
- Frame-agnostic XAI Library for Computer Vision, for understanding why models behave that way.☆11Feb 19, 2023Updated 3 years ago
- ☆23Dec 8, 2022Updated 3 years ago
- About Samples code for Axera's PCIE Card for computer vision applications.☆20Aug 10, 2026Updated 2 weeks ago
- PCJ library☆22Jan 16, 2026Updated 7 months ago