Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models (ICLR 2024)
☆14May 31, 2025Updated last year
Alternatives and similar repositories for K-prune
Users that are interested in K-prune are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PET: Parameter-efficient Knowledge Distillation on Transformer (PLOS One)☆14Aug 22, 2025Updated last year
- Sturctured pruning algorithm for pruning Transformer☆30Dec 6, 2023Updated 2 years ago
- ☆32Dec 9, 2022Updated 3 years ago
- SensiMix: Sensitivity-Aware 8-bit Index & 1-bit Value Mixed Precision Quantization for BERT Compression (PLOS One)☆33Aug 22, 2025Updated last year
- Flexible Convolutional Neural Network☆23Nov 15, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SynQ: Accurate Zero-shot Quantization by Synthesis-aware Fine-tuning (ICLR 2025)☆26Feb 7, 2025Updated last year
- Vector multiplication on Low-rank Matrix Factorization☆46Nov 10, 2023Updated 2 years ago
- Falcon: Lightweight and Accurate Convolution Based on Depthwise Separable Convolution (KAIS)☆47Aug 22, 2025Updated last year
- Pea-KD: Parameter-efficient and accurate knowledge distillation on BERT (PLOS One)☆34Aug 22, 2025Updated last year
- Model-Agnostic Augmentation for Accurate Graph Classification (WWW 2022)☆20Aug 22, 2025Updated last year
- [ICML 2024] Sparse Model Inversion: Efficient Inversion of Vision Transformers with Less Hallucination☆15Apr 29, 2025Updated last year
- ☆33Jul 8, 2024Updated 2 years ago
- Code and data for the ACM CIKM 2024 paper "Adversarial Text Rewriting for Text-aware Recommender Systems"☆13Aug 1, 2024Updated 2 years ago
- Fast and Accurate Partial Fourier Transform for Time Series Data (KDD 2021)☆16Aug 19, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆10Oct 29, 2020Updated 5 years ago
- ☆11Feb 22, 2020Updated 6 years ago
- [CVPR '24] Official implementation of the paper "Multiflow: Shifting Towards Task-Agnostic Vision-Language Pruning".☆24Mar 7, 2025Updated last year
- [ICLR 2025] Adaptive prompt tailored pruning of T2I diffusion models.☆15Feb 1, 2025Updated last year
- WWW22 - MiDaS: Representative Hypergraph Sampling☆13May 7, 2024Updated 2 years ago
- ☆55Nov 22, 2022Updated 3 years ago
- TensorCodec: Compact Lossy Compression of Tensors without Strong Data Assumptions (ICDM 23)☆12Jun 27, 2024Updated 2 years ago
- Codes for Merging Large Language Models☆37Aug 7, 2024Updated 2 years ago
- Official implementation of the ICLR paper "Streamlining Redundant Layers to Compress Large Language Models"☆45May 1, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Unofficial pytorch implementation of DisenGCN☆14Jun 19, 2023Updated 3 years ago
- ☆15Oct 19, 2021Updated 4 years ago
- Source code for ICDM 2022 paper "Set2Box: Similarity Preserving Representation Learning for Sets."☆14Dec 25, 2022Updated 3 years ago
- Pytorch implementation of our paper accepted by ECCV 2022-- Fine-grained Data Distribution Alignment for Post-Training Quantization☆16Sep 13, 2022Updated 4 years ago
- DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling☆39Jul 12, 2024Updated 2 years ago
- ☆14Mar 20, 2024Updated 2 years ago
- torch_quantizer is a out-of-box quantization tool for PyTorch models on CUDA backend, specially optimized for Diffusion Models.☆25Mar 29, 2024Updated 2 years ago
- ☆12Mar 15, 2022Updated 4 years ago
- Representation Surgery for Multi-Task Model Merging. ICML, 2024.☆50Oct 10, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Published as a journal paper at DAMI 2023☆19Feb 11, 2025Updated last year
- An implementation on "Curved-Voxel Clustering for Accurate Segmentation of 3D LiDAR Point Clouds with Real-Time Performance" from IROS 20…☆242Jan 25, 2022Updated 4 years ago
- Implementation of Time-aware Random Walk Diffusion to Improve Dynamic Graph Learning (AAAI-23)☆16Dec 1, 2022Updated 3 years ago
- [NeurIPS 2024] GACL: Exemplar-Free Generalized Analytic Continual Learning☆19Nov 5, 2024Updated last year
- [ICLR 2024] Hebbian Learning based Orthogonal Projection for Continual Learning of Spiking Neural Networks☆45Feb 20, 2024Updated 2 years ago
- DenseAlert: Incremental Dense-SubTensor Detection in Tensor Streams (KDD'17)☆19Oct 30, 2024Updated last year
- ☆18Oct 27, 2022Updated 3 years ago