Implementation of a Quantized Transformer Model
☆20Mar 20, 2019Updated 7 years ago
Alternatives and similar repositories for QuantizedTransformer
Users that are interested in QuantizedTransformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Open Source Neural Machine Translation in PyTorch☆17Apr 15, 2019Updated 7 years ago
- PyTorch code for full quantization of DNN using BCGD☆14Jul 24, 2019Updated 7 years ago
- AFP is a hardware-friendly quantization framework for DNNs, which is contributed by Fangxin Liu and Wenbo Zhao.☆13Nov 8, 2021Updated 4 years ago
- 🔮 LLM GPU Calculator☆21Aug 19, 2023Updated 3 years ago
- BSQ: Exploring Bit-Level Sparsity for Mixed-Precision Neural Network Quantization (ICLR 2021)☆41Jan 12, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Implementation of several knowledge distillation techniques on PyTorch☆15Feb 25, 2019Updated 7 years ago
- An 8bit automated quantization conversion tool for the pytorch (Post-training quantization based on KL divergence)☆32Nov 17, 2019Updated 6 years ago
- Proximal Mean-field for Neural Network Quantization☆21Apr 9, 2020Updated 6 years ago
- Peking University Embedded Microprocessor System Lesson’s all Homework☆10Dec 28, 2021Updated 4 years ago
- HLS project modeling various sparse accelerators.☆12Jan 11, 2022Updated 4 years ago
- ☆13Jan 14, 2025Updated last year
- Code to accompany: https://arxiv.org/abs/2001.08349☆11Jul 6, 2023Updated 3 years ago
- ☆15Feb 18, 2022Updated 4 years ago
- DL quantization for pytorch☆26Mar 30, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Model for processing text sequences with coreference annotations☆14Nov 29, 2018Updated 7 years ago
- Implementation of Sparse Shift Layer and Active Shift Layer (3D, 4D, 5D tensors) for PyTorch(CPU,GPU)☆34May 5, 2021Updated 5 years ago
- ☆37Sep 3, 2023Updated 3 years ago
- Code for ViTAS_Vision Transformer Architecture Search☆50Jul 22, 2021Updated 5 years ago
- Example of Empirical Mode Decomposition algorithm☆14Mar 25, 2021Updated 5 years ago
- [KDD'22] Learned Token Pruning for Transformers☆98Feb 27, 2023Updated 3 years ago
- Tensorflow implementation of the Gradient Reversal layer from https://arxiv.org/abs/1505.07818☆13Jun 19, 2018Updated 8 years ago
- Question generation from Reading Comprehension☆18Feb 28, 2022Updated 4 years ago
- [AAAI 2023 Oral] Peeling the Onion: Hierarchical Reduction of Data Redundancy for Efficient Vision Transformer Training☆14Apr 19, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repository represents training examples for the CVPR 2018 paper "SYQ:Learning Symmetric Quantization For Efficient Deep Neural Netwo…☆31Jul 25, 2019Updated 7 years ago
- ☆21Jul 20, 2022Updated 4 years ago
- Caffe implementation of Optimal-Ternary-Weights-Approximation in "Two-Step Quantization for Low-bit Neural Networks" (CVPR2018).☆15Sep 21, 2018Updated 7 years ago
- ☆12Nov 24, 2023Updated 2 years ago
- [ICLR 2025] Code for the PopulationTransformer☆18Updated this week
- Official PyTorch implementation of "Evolving Search Space for Neural Architecture Search"☆12Aug 18, 2021Updated 5 years ago
- Neural Network Quantization With Fractional Bit-widths☆11Feb 19, 2021Updated 5 years ago
- Codes for AAAI2019 paper: Deep Neural Network Quantization via Layer-Wise Optimization using Limited Training Data☆41Jan 22, 2019Updated 7 years ago
- ☆14Oct 26, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official Repo for "Multi-objective Differentiable Neural Architecture Search"☆13Jul 12, 2024Updated 2 years ago
- Static Block Floating Point Quantization for CNN☆38Jun 9, 2021Updated 5 years ago
- Squeeze and Excitation network implementation.☆19May 26, 2019Updated 7 years ago
- LSTM neural network (verilog)☆16Dec 5, 2018Updated 7 years ago
- Digital Design Lab Spring 2019 Final Project☆13Jun 17, 2019Updated 7 years ago
- MathPrompter Implementation: This repository hosts an implementation based on the 'MathPrompter: Mathematical Reasoning Using Large Langu…☆17Apr 12, 2025Updated last year
- Contains code for Binary, Ternary, N-bit Quantized and Hybrid CNNs for low precision experiments.☆27Oct 30, 2018Updated 7 years ago