A curated list of awesome papers on dataset reduction, including dataset distillation (dataset condensation) and dataset pruning (coreset selection).
☆61Jan 14, 2025Updated last year
Alternatives and similar repositories for Awesome-Dataset-Reduction
Users that are interested in Awesome-Dataset-Reduction are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official PyTorch code for ICLR 2025 paper "Gnothi Seauton: Empowering Faithful Self-Interpretability in Black-Box Models"☆23Mar 4, 2025Updated last year
- ATOM: A Framework of Detecting Query-Based Model Extraction Attacks for Graph Neural Networks☆18Jul 4, 2025Updated last year
- [ICCV 2025] Official Repository for "Multi-Cache Enhanced Prototype Learning for Test-Time Generalization of Vision-Language Models"☆20Nov 10, 2025Updated 9 months ago
- (NeurIPS 2025 🔥) Official implementation for "Efficient Multi-modal Large Language Models via Progressive Consistency Distillation"☆49Feb 11, 2026Updated 5 months ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- You Only Condense Once: Two Rules for Pruning Condensed Datasets (NeurIPS 2023)☆17Jul 30, 2026Updated last week
- An Easy and Unified Interface for Robots (and Grippers, etc.)☆14Nov 7, 2024Updated last year
- Pytorch implementation of OCFGAN-GP (CVPR 2020, Oral).☆15Apr 3, 2020Updated 6 years ago
- Data distillation benchmark☆73Jun 13, 2025Updated last year
- Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning☆37Oct 26, 2025Updated 9 months ago
- ☆18Jun 14, 2024Updated 2 years ago
- [NeurIPS 2023] Code release for "Going Beyond Linear Mode Connectivity: The Layerwise Linear Feature Connectivity"☆19Oct 19, 2023Updated 2 years ago
- (CVPR 2025) Official implementation to DELT: A Simple Diversity-driven EarlyLate Training for Dataset Distillation which outperforms SOTA…☆28Aug 23, 2025Updated 11 months ago
- [ICLR 2025] "Rethinking LLM Unlearning Objectives: A Gradient Perspective and Go Beyond"☆16Feb 27, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [IROS 2025] SIME: Enhancing Policy Self-Improvement with Modal-level Exploration☆17Mar 2, 2026Updated 5 months ago
- ☆30Nov 5, 2024Updated last year
- [RA-L 2025 & ICRA 2026] Motion Before Action: Diffusing Object Motion as Manipulation Condition☆73Nov 4, 2025Updated 9 months ago
- [ICLR 2025 Spotlight] Code release for "Sharpness-Aware Minimization Efficiently Selects Flatter Minima Late In Training"☆19Feb 20, 2025Updated last year
- [ICRA 2025] CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation☆35Jan 14, 2025Updated last year
- Official Implementation of paper "Distilling Long-tailed Datasets" [CVPR 2025]☆25Aug 13, 2025Updated 11 months ago
- [NeurIPS 2025@FoRLM] R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search☆17Jan 24, 2026Updated 6 months ago
- ☆44Oct 13, 2023Updated 2 years ago
- NeurIPS 2020 Spotlight Paper☆13Dec 20, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- (Pattern Recognition 2025) Towards Trustworthy Dataset Distillation☆14Dec 8, 2024Updated last year
- several examples of the learning of the java☆11Nov 22, 2023Updated 2 years ago
- [CVPR2024 highlight] Generalized Large-Scale Data Condensation via Various Backbone and Statistical Matching (G-VBSM)☆27Oct 9, 2024Updated last year
- Preview code of ECCV'24 paper "Distill Gold from Massive Ores" (BiLP)☆25Jul 6, 2024Updated 2 years ago
- This repository is the official implementation of Dataset Condensation with Contrastive Signals (DCC), accepted at ICML 2022.☆22Jun 8, 2022Updated 4 years ago
- ☆28Jun 12, 2023Updated 3 years ago
- torch.optim.lr_scheduler☆10Mar 17, 2020Updated 6 years ago
- CopT: Contrastive On-Policy Thinking with Continuous Spaces for General and Agentic Reasoning☆18May 21, 2026Updated 2 months ago
- Data and Baselines for AStitchInLanguageModels dataset☆13Oct 31, 2022Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Implementation of Mutan+ArticleNet on OKVQA☆10Jan 11, 2021Updated 5 years ago
- [ICCV 2025] 2D version of Dense Policy (DSP)☆34Jan 14, 2026Updated 6 months ago
- 本项目主要是2025届浙江大学软件学院夏令营(AI营)的考核项目☆12Mar 3, 2025Updated last year
- A curated list of awesome papers on dataset distillation and related applications.☆1,970Jul 21, 2026Updated 2 weeks ago
- The code for our work☆18Apr 7, 2024Updated 2 years ago
- Source code for "Learning Deep Priors for Image Dehazing", ICCV 2019☆10Sep 18, 2020Updated 5 years ago
- Official code of "ALIM: Adjusting Label Importance Mechanism for Noisy Partial Label Learning"☆23Sep 25, 2023Updated 2 years ago