A curated list of awesome papers on dataset reduction, including dataset distillation (dataset condensation) and dataset pruning (coreset selection).
☆62Jan 14, 2025Updated last year
Alternatives and similar repositories for Awesome-Dataset-Reduction
Users that are interested in Awesome-Dataset-Reduction are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official PyTorch code for ICLR 2025 paper "Gnothi Seauton: Empowering Faithful Self-Interpretability in Black-Box Models"☆24Mar 4, 2025Updated last year
- ATOM: A Framework of Detecting Query-Based Model Extraction Attacks for Graph Neural Networks☆18Jul 4, 2025Updated last year
- [ICCV 2025] Official Repository for "Multi-Cache Enhanced Prototype Learning for Test-Time Generalization of Vision-Language Models"☆21Nov 10, 2025Updated 10 months ago
- Official PyTorch implementation of the paper "Dataset Distillation with Neural Characteristic Function: A Minmax Perspective" (NCFM) in C…☆412Jun 3, 2026Updated 3 months ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An Easy and Unified Interface for Robots (and Grippers, etc.)☆14Nov 7, 2024Updated last year
- [CVPR 2025] An Implementation of the paper "Pre-Instruction Data Selection for Visual Instruction Tuning"☆17Jun 9, 2025Updated last year
- Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning☆37Oct 26, 2025Updated 10 months ago
- ☆18Jun 14, 2024Updated 2 years ago
- [NeurIPS 2023] Code release for "Going Beyond Linear Mode Connectivity: The Layerwise Linear Feature Connectivity"☆19Oct 19, 2023Updated 2 years ago
- (CVPR 2025) Official implementation to DELT: A Simple Diversity-driven EarlyLate Training for Dataset Distillation which outperforms SOTA…☆28Aug 23, 2025Updated last year
- [ICLR 2025] "Rethinking LLM Unlearning Objectives: A Gradient Perspective and Go Beyond"☆16Feb 27, 2025Updated last year
- Prioritize Alignment in Dataset Distillation☆21Dec 3, 2024Updated last year
- [ICLR 2025 Spotlight] Code release for "Sharpness-Aware Minimization Efficiently Selects Flatter Minima Late In Training"☆20Feb 20, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICRA 2025] CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation☆35Jan 14, 2025Updated last year
- [NeurIPS‘2021] "MEST: Accurate and Fast Memory-Economic Sparse Training Framework on the Edge", Geng Yuan, Xiaolong Ma, Yanzhi Wang et al…☆18Mar 16, 2022Updated 4 years ago
- [NeurIPS 2025@FoRLM] R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search☆17Jan 24, 2026Updated 7 months ago
- ☆44Oct 13, 2023Updated 2 years ago
- (Pattern Recognition 2025) Towards Trustworthy Dataset Distillation☆14Dec 8, 2024Updated last year
- several examples of the learning of the java☆11Nov 22, 2023Updated 2 years ago
- Preview code of ECCV'24 paper "Distill Gold from Massive Ores" (BiLP)☆25Jul 6, 2024Updated 2 years ago
- Official codebase for our NeurIPS paper, Symmetry-Informed Governing Equation Discovery.☆11Nov 13, 2024Updated last year
- This repository is the official implementation of Dataset Condensation with Contrastive Signals (DCC), accepted at ICML 2022.☆22Jun 8, 2022Updated 4 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- torch.optim.lr_scheduler☆10Mar 17, 2020Updated 6 years ago
- [ICCV 2023] DataDAM: Efficient Dataset Distillation with Attention Matching☆34Jun 20, 2024Updated 2 years ago
- Not All Patches Are Equal: Hierarchical Dataset Condensation for Single Image Super-Resolution☆10May 7, 2024Updated 2 years ago
- [ICCV 2025] 2D version of Dense Policy (DSP)☆34Jan 14, 2026Updated 8 months ago
- Source code for "Learning Deep Priors for Image Dehazing", ICCV 2019☆10Sep 18, 2020Updated 6 years ago
- ☆12Jul 30, 2025Updated last year
- [EMNLP 2025 main 🔥] Code for "Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More"☆122Oct 12, 2025Updated 11 months ago
- [IROS 2024] 📈 RISE: 3D Perception Makes Real-World Robot Imitation Simple and Effective☆160Nov 29, 2025Updated 9 months ago
- [WACV 2025] Code for Enhancing Vision-Language Few-Shot Adaptation with Negative Learning☆13Feb 24, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Welcome to the 'In Context Learning Theory' Reading Group☆31Nov 8, 2024Updated last year
- This is the repository for paper `Learning Task-Aware Effective Brain Connectivity for fMRI Analysis with Graph Neural Networks'.☆14Nov 22, 2023Updated 2 years ago
- [AAAI-2025] The offical code for SiTo (Similarity-based Token Pruning for Stable Diffusion Models)☆47Jun 2, 2025Updated last year
- AAAI23-Directed Acyclic Graph Structure Learning from Dynamic Graphs☆12Nov 25, 2022Updated 3 years ago
- Code for CVPR24 Paper - Resource-Efficient Transformer Pruning for Finetuning of Large Models☆12Oct 31, 2025Updated 10 months ago
- Official implementation of our ICML 2024 paper "UP2ME: Univariate Pre-training to Multivariate Fine-tuning as a General-purpose Framework…☆36May 12, 2025Updated last year
- Selen's Personal Homepage☆14Updated this week