Official implementation of "Searching for Winograd-aware Quantized Networks" (MLSys'20)
☆27Oct 3, 2023Updated 2 years ago
Alternatives and similar repositories for WinogradAwareNets
Users that are interested in WinogradAwareNets are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Efficient Sparse-Winograd Convolutional Neural Networks (ICLR 2018)☆191May 7, 2019Updated 7 years ago
- Winograd minimal convolution algorithm generator for convolutional neural networks.☆628Feb 9, 2026Updated 6 months ago
- Implementation of the Winograd algorithm.☆24Nov 6, 2018Updated 7 years ago
- An HLS based winograd systolic CNN accelerator☆54Jul 18, 2021Updated 5 years ago
- [CF ’20] Verified Instruction-Level Energy Consumption Measurement for NVIDIA GPUs☆15Dec 11, 2020Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A Winograd Minimal Filter Implementation in CUDA☆31Aug 25, 2021Updated 4 years ago
- Torch Frontend for IREE☆26Dec 21, 2023Updated 2 years ago
- HeteroHalide: From Image Processing DSL to Efficient FPGA Acceleration☆15Sep 14, 2020Updated 5 years ago
- ☆11Sep 3, 2022Updated 3 years ago
- Deep learning with a multiplication budget☆47Jul 15, 2018Updated 8 years ago
- My implementation of an FPGA Deep Neural Network Hardware Accelerator, moved from my bitbucket☆30Jul 31, 2019Updated 7 years ago
- Official PyTorch implementation of "EvoGrad: Efficient Gradient-Based Meta-Learning and Hyperparameter Optimization"☆23Oct 24, 2021Updated 4 years ago
- Winograd-based convolution implementation in OpenCL☆29Jan 22, 2017Updated 9 years ago
- The code for AIM2022 compressed image super-resolution☆11Nov 30, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This is a sample implementation of "Robust Graph Convolutional Networks Against Adversarial Attacks", KDD 2019.☆10Dec 8, 2020Updated 5 years ago
- Graph Transforms to Quantize and Retrain Deep Neural Nets in TensorFlow☆170Dec 9, 2019Updated 6 years ago
- first-order deep learning accelerator model☆22Nov 27, 2017Updated 8 years ago
- [IPSN 2024] Lifelong Intelligence Beyond the Edge using Hyperdimensional Computing☆14May 16, 2024Updated 2 years ago
- This repository is an excuse to learn about Convolutional Neural Networks by implementing one in FPGA. The main goal is to learn, and to …☆12Jul 12, 2020Updated 6 years ago
- [NeurIPS 2023] ShiftAddViT: Mixture of Multiplication Primitives Towards Efficient Vision Transformer☆31Dec 6, 2023Updated 2 years ago
- Face recognition with loss of softmax, sphereface, cosface, arcface in pytorch of python3☆10Apr 27, 2020Updated 6 years ago
- GPU implementation of Winograd convolution☆10Oct 23, 2017Updated 8 years ago
- ☆13Mar 30, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Explore the energy-efficient dataflow scheduling for neural networks.☆237Apr 13, 2026Updated 4 months ago
- AI Accelerators-SC23-tutorial Repository☆12Nov 12, 2023Updated 2 years ago
- ☆23Jun 12, 2021Updated 5 years ago
- Post-training sparsity-aware quantization☆34Feb 26, 2023Updated 3 years ago
- FPGA-based hardware acceleration for dropout-based Bayesian Neural Networks.☆28Aug 15, 2023Updated 3 years ago
- GEMM and Winograd based convolutions using CUTLASS☆28Jul 15, 2020Updated 6 years ago
- implement of DoReFaNet with tensorflow based on cifar10 dataset☆28Nov 8, 2017Updated 8 years ago
- Face recognition using Tensorflow☆11Nov 6, 2018Updated 7 years ago
- CPrune: Compiler-Informed Model Pruning for Efficient Target-Aware DNN Execution☆17Jun 25, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Benchmarking PyTorch 2.0 different models☆20Mar 19, 2023Updated 3 years ago
- Adaptive Deep Learning Model Selection On Embedded Systems☆11May 6, 2018Updated 8 years ago
- ☆10Apr 24, 2023Updated 3 years ago
- ☆28Oct 26, 2019Updated 6 years ago
- [ICLR 2021] HW-NAS-Bench: Hardware-Aware Neural Architecture Search Benchmark☆118Apr 18, 2023Updated 3 years ago
- ☆50Jun 27, 2019Updated 7 years ago
- Exploring Motion Ambiguity and Alignment for High-Quality Video Frame Interpolation (CVPR2023)☆14Jul 21, 2023Updated 3 years ago