a computing kernel implementation in ML inference framework aiming at theoretical limit
☆12Dec 18, 2019Updated 6 years ago
Alternatives and similar repositories for speedup-aarch64-cpu
Users that are interested in speedup-aarch64-cpu are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Feb 26, 2026Updated 6 months ago
- Run OpenCL program on MOBILE GPU (Qualcomm & ARM) !☆18Jun 27, 2018Updated 8 years ago
- 阴阳师御魂方案计算工具,基于动态规划和剪枝☆14Sep 3, 2018Updated 8 years ago
- An easy way to run, test, benchmark and tune OpenCL kernel files☆24Aug 25, 2023Updated 3 years ago
- C++ deepsort on tensorflow☆18Apr 4, 2020Updated 6 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A set of tools to work with cgroup tree and process classification/QoS according to it☆10Oct 1, 2019Updated 6 years ago
- Erasure code library for Erlang☆12Sep 5, 2024Updated 2 years ago
- LLVM passes with usage instructions☆18Apr 23, 2017Updated 9 years ago
- linux内核异步内存回收的另一个思路:基于冷热文件的冷热区域精准的回收冷文件页page(可做成内核ko)☆13Jun 14, 2024Updated 2 years ago
- A structure from motion implemention in C++ and accelerated using CUDA☆48Oct 12, 2019Updated 6 years ago
- Clone of https://code.google.com/p/google-coredumper/ with enhancements by Amadeus☆13Jul 2, 2024Updated 2 years ago
- Reed-Solomon Erasure Coding in Haskell☆23Jan 22, 2017Updated 9 years ago
- Last Writer Slicing: data provenance tracking for concurrent program debugging & analysis☆13Nov 14, 2014Updated 11 years ago
- An AI/ML solution that provides a probability that a hard drive will fail within some pre-defined time period.☆13Dec 9, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The repository targets the OpenCL gemm function performance optimization. It compares several libraries clBLAS, clBLAST, MIOpenGemm, Inte…☆17Mar 28, 2019Updated 7 years ago
- [IJCNN'19, IEEE JSTSP'19] Caffe code for our paper "Structured Pruning for Efficient ConvNets via Incremental Regularization"; [BMVC'18] …☆14Feb 14, 2020Updated 6 years ago
- OpenDNN: An Open-source, cuDNN-like Deep Learning Primitive Library☆29Dec 9, 2019Updated 6 years ago
- ☆14Jun 9, 2023Updated 3 years ago
- Documentation for the entire CGRAFlow☆19Sep 17, 2021Updated 4 years ago
- Pytorch implementation of DAC-Net ("Zhongying Deng, Kaiyang Zhou, Yongxin Yang, Tao Xiang. Domain Attention Consistency for Multi-Source …☆24Dec 13, 2021Updated 4 years ago
- disk prediction papers☆18Oct 24, 2020Updated 5 years ago
- dnotify,inotify, and fanotify example code from http://www.lanedo.com/filesystem-monitoring-linux-kernel/☆14Apr 28, 2017Updated 9 years ago
- This is a caffe implementation of ShuffleNet model☆15Mar 15, 2018Updated 8 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Multi-Source Domain Adaptation via Optimal Transport for Student-Teacher Learning - UAI 2021☆23Jan 26, 2025Updated last year
- dump zmq messages on a socket☆15Aug 27, 2023Updated 3 years ago
- Heterogeneous Active Messages C++ library☆21Nov 8, 2019Updated 6 years ago
- A prototype implementation of AllReduce collective communication routine.☆19Sep 27, 2018Updated 7 years ago
- strace-perfetto runs strace and converts the raw output to a Trace Event JSON file. The JSON file can then be analyzed using Google's Per…☆13Apr 27, 2022Updated 4 years ago
- A simple cycle-accurate DaDianNao simulator☆13Mar 27, 2019Updated 7 years ago
- 一个尝试固液耦合的沙盒玩具☆11Feb 17, 2025Updated last year
- ☆14Dec 8, 2022Updated 3 years ago
- C/C++ header dependency list generator. Output can be used to create a dependency graph.☆15Apr 13, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆16Aug 11, 2016Updated 10 years ago
- NVM user-space Primitives API library repository☆18Mar 12, 2014Updated 12 years ago
- Basic algorithms for vslam.☆54Nov 20, 2020Updated 5 years ago
- This is a pintool that can analyze target dynamically and output code blocks and "key frames".☆14Mar 26, 2015Updated 11 years ago
- log, 仅包含头文件,追踪崩溃和数据的日志库☆16Dec 25, 2018Updated 7 years ago
- Training code of 'Joint 3D Face Reconstruction and Dense Alignment with Position Map Regression Network'. https://arxiv.org/abs/1803.0783…☆10Aug 14, 2018Updated 8 years ago
- 个人笔记☆18Updated this week