Implementation of convolution layer in different flavors
☆68Oct 8, 2017Updated 9 years ago
Alternatives and similar repositories for convolution-flavors
Users that are interested in convolution-flavors are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- HLS implemented systolic array structure☆41Nov 13, 2017Updated 8 years ago
- ICML2017 MEC: Memory-efficient Convolution for Deep Neural Network C++实现(非官方)☆17Apr 9, 2019Updated 7 years ago
- SDA: Low-Bit Stable Diffusion Acceleration on Edge FPGAs☆19May 23, 2024Updated 2 years ago
- Systolic matrix multiplication kernel implemented on Xilinx PYNQ FPGA board☆18Jun 23, 2020Updated 6 years ago
- An implementation of dilated convolutional layer based on Darknet Architecture☆30Jun 1, 2019Updated 7 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Matrix-Vector Library Designed for Neural Network Construction. cuda (gpu) support, openmp (multithreaded cpu) support, partial supp…☆13Nov 8, 2020Updated 5 years ago
- This is a collection of works on neural networks and neural accelerators.☆42Mar 3, 2019Updated 7 years ago
- An FPGA Accelerator for Transformer Inference☆98Apr 29, 2022Updated 4 years ago
- ☆22Jun 22, 2016Updated 10 years ago
- [FPGA'21] Microbenchmarks for Demystifying the Memory System of Modern Datacenter FPGAs for Software Programmers☆31Dec 16, 2021Updated 4 years ago
- Custom extensions to the RISC-V isa simulator for the UCB-BAR ESP project☆17Nov 27, 2022Updated 3 years ago
- ☆24Dec 1, 2016Updated 9 years ago
- ☆62Mar 15, 2018Updated 8 years ago
- Faster Non-Integer Sample Rate Conversion☆38Jul 12, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Reconstruction and Compression of Color Images Using Principal Component Analysis (PCA) Algorithm☆35Jun 3, 2020Updated 6 years ago
- Pruning methods for pytorch with an optimizer-like interface☆15Apr 14, 2020Updated 6 years ago
- softfloat and softposit in Python☆15Aug 2, 2019Updated 7 years ago
- Operating system for ARM processors☆10Aug 1, 2018Updated 8 years ago
- A simple wrapper for python-bioformats to convert .vsi CellSense format to TIFF.☆12Apr 9, 2022Updated 4 years ago
- ☆13Nov 29, 2024Updated last year
- ☆67May 14, 2022Updated 4 years ago
- Clean C++ CMake Ninja VSCode template☆12Oct 23, 2025Updated 11 months ago
- "Forked" from Xilinx/Edge-AI-Platform-Tutorials☆18Dec 14, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MAC system with IEEE754 compatibility☆14Nov 22, 2023Updated 2 years ago
- A Numpy implementation of a Convolutional Neural Network: slow & fast (im2col/col2im).☆60Jul 6, 2023Updated 3 years ago
- Scalable systolic array-based matrix-matrix multiplication implemented in Vivado HLS for Xilinx FPGAs.☆390Jan 20, 2025Updated last year
- An adapted version of the original caffe deep learning library to support training, finetuning and testing of convolutional neural networ…☆20Jul 11, 2017Updated 9 years ago
- 基于FP16的二维脉动阵列电路设计☆13Feb 23, 2023Updated 3 years ago
- Mad Zombie Classic 4th☆14Oct 8, 2023Updated 3 years ago
- Implementation of Hyena Hierarchy in JAX☆10Apr 30, 2023Updated 3 years ago
- c++ version of ViT☆13Nov 13, 2022Updated 3 years ago
- An HLS based winograd systolic CNN accelerator☆54Jul 18, 2021Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- BLAS / LAPACK for JavaScript☆52Aug 19, 2017Updated 9 years ago
- ☆14Feb 7, 2020Updated 6 years ago
- a highly-efficient library for deep neural networks based on Sunway TaihuLight supercomputer.☆17Sep 3, 2018Updated 8 years ago
- Convolutional Neural Networks☆31Feb 8, 2018Updated 8 years ago
- A Three-Dimensional, Serial Fast Multipole Method Code Based on the Work of Walter Dehnen☆10Jun 12, 2018Updated 8 years ago
- SDK for creating waPC WebAssembly Guest Modules in Zig☆14Dec 27, 2021Updated 4 years ago
- Code for the work in: https://www.nature.com/articles/s41534-020-00305-x Basically a generative neural network to tackle the classical ca…☆15Apr 17, 2025Updated last year