A block oriented training approach for inference time optimization.
☆34Aug 19, 2024Updated 2 years ago
Alternatives and similar repositories for superblock
Users that are interested in superblock are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official PyTorch implementation of "Efficient Latency-Aware CNN Depth Compression via Two-Stage Dynamic Programming" (ICML'23)☆13Apr 13, 2026Updated 4 months ago
- ☆19Mar 29, 2026Updated 5 months ago
- Official PyTorch implementation of "LayerMerge: Neural Network Depth Compression through Layer Pruning and Merging" (ICML 2024)☆32Apr 13, 2026Updated 4 months ago
- Official PyTorch implementation of QwT—“Quantization without Tears” (CVPR 2025): fast, accurate, and hassle-free post-training network qu…☆32Sep 30, 2025Updated 11 months ago
- Flash Attention from scratch, tiled CUDA forward kernel, online softmax with running max and correction factor, recomputation trick in ba…☆18Mar 6, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆27Mar 14, 2024Updated 2 years ago
- ☆13Sep 25, 2023Updated 2 years ago
- LLM Inference with Microscaling Format☆35Nov 12, 2024Updated last year
- ☆26Sep 9, 2024Updated last year
- MaskedTensors for PyTorch☆39Jul 17, 2022Updated 4 years ago
- Code for reproducing work of ICML 2019 paper: Memory-Optimal Direct Convolutions for Maximizing Classification Accuracy in Embedded Appli…☆12Jun 8, 2019Updated 7 years ago
- ☆14Mar 8, 2025Updated last year
- Invariant Feature Regularization for Fair Face Recognition (ICCV'23)☆15Oct 23, 2023Updated 2 years ago
- ☆13May 30, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICML 2024] Sparse Model Inversion: Efficient Inversion of Vision Transformers with Less Hallucination☆14Apr 29, 2025Updated last year
- NITEC: Versatile Hand-Annotated Eye Contact Dataset for Ego-Vision Interaction (WACV24)☆18Jul 17, 2024Updated 2 years ago
- MelGAN and Tacotron 2 in PyTorch☆11Oct 22, 2019Updated 6 years ago
- My attempt to improve the speed of the newton schulz algorithm, starting from the dion implementation.☆42Apr 30, 2026Updated 4 months ago
- Pebble REBBLE watchface☆13Mar 3, 2025Updated last year
- A family of efficient edge language models in 100M~1B sizes.☆19Feb 14, 2025Updated last year
- A fast implementation of Neural Image Caption by Chainer☆16Aug 9, 2018Updated 8 years ago
- Generates random utf-8 strings for fuzz t�sting character encoding probl�ms☆11Aug 21, 2015Updated 11 years ago
- [NeurIPS'24]Efficient and accurate memory saving method towards W4A4 large multi-modal models.☆102Jan 3, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆16Dec 9, 2023Updated 2 years ago
- This repository provides a multi task benchmark for instance segmentation, depth estimation, and 3D object detection.☆14Jul 29, 2023Updated 3 years ago
- 삼각형의 실전! Triton☆16Feb 15, 2024Updated 2 years ago
- ROOT講習会☆10Jul 28, 2026Updated last month
- Implementation of various generative models☆14Oct 1, 2018Updated 7 years ago
- MaXM is a suite of test-only benchmarks for multilingual visual question answering in 7 languages: English (en), French (fr), Hindi (hi),…☆13Jan 16, 2024Updated 2 years ago
- Source code of our TNNLS paper "Boosting Convolutional Neural Networks with Middle Spectrum Grouped Convolution"☆12Apr 14, 2023Updated 3 years ago
- ☆13Jul 14, 2025Updated last year
- ☆15Apr 11, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Pdf Query chat-bot using Gemini AI and Llma Index☆10Dec 24, 2023Updated 2 years ago
- Segmenting a given document using recursive xy-cut algorithm.☆12Oct 9, 2018Updated 7 years ago
- Continuous regular group convolutions for Pytorch☆12Jun 9, 2024Updated 2 years ago
- ☆14Feb 9, 2026Updated 6 months ago
- Code for "Fast Sparse ConvNets" CVPR2020 submissions☆12Nov 20, 2019Updated 6 years ago
- ☆13Jun 16, 2024Updated 2 years ago
- JSを言語仕様から把握し、ライブラリに振り回されない漢を目指すリポジトリ☆19Aug 22, 2018Updated 8 years ago