A curated list of awesome papers on dataset reduction, including dataset distillation (dataset condensation) and dataset pruning (coreset selection).
☆62Jan 14, 2025Updated last year
Alternatives and similar repositories for Awesome-Dataset-Reduction
Users that are interested in Awesome-Dataset-Reduction are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Exploiting Inter-sample and Inter-feature Relations in Dataset Distillation (CVPR24)☆10Jun 16, 2024Updated 2 years ago
- Official PyTorch implementation of the paper "Dataset Distillation with Neural Characteristic Function: A Minmax Perspective" (NCFM) in C…☆413Jun 3, 2026Updated 2 months ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- An Easy and Unified Interface for Robots (and Grippers, etc.)☆14Nov 7, 2024Updated last year
- Pytorch implementation of OCFGAN-GP (CVPR 2020, Oral).☆15Apr 3, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2025] An Implementation of the paper "Pre-Instruction Data Selection for Visual Instruction Tuning"☆17Jun 9, 2025Updated last year
- The code of the paper "Minimizing the Accumulated Trajectory Error to Improve Dataset Distillation" (CVPR2023)☆40Mar 25, 2023Updated 3 years ago
- [NeurIPS 2023] Code release for "Going Beyond Linear Mode Connectivity: The Layerwise Linear Feature Connectivity"☆19Oct 19, 2023Updated 2 years ago
- (CVPR 2025) Official implementation to DELT: A Simple Diversity-driven EarlyLate Training for Dataset Distillation which outperforms SOTA…☆28Aug 23, 2025Updated last year
- Prioritize Alignment in Dataset Distillation☆21Dec 3, 2024Updated last year
- [ICLR 2025] "Rethinking LLM Unlearning Objectives: A Gradient Perspective and Go Beyond"☆16Feb 27, 2025Updated last year
- The officalimplement of dLLM-Factory☆25Jul 12, 2025Updated last year
- ☆30Nov 5, 2024Updated last year
- Official PyTorch implementation of "Loss-Curvature Matching for Dataset Selection and Condensation" (AISTATS 2023)☆22Mar 14, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICRA 2025] CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation☆35Jan 14, 2025Updated last year
- [NeurIPS‘2021] "MEST: Accurate and Fast Memory-Economic Sparse Training Framework on the Edge", Geng Yuan, Xiaolong Ma, Yanzhi Wang et al…☆18Mar 16, 2022Updated 4 years ago
- [NeurIPS 2025@FoRLM] R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search☆17Jan 24, 2026Updated 7 months ago
- [ICRA 2026] 🌠 DSPv2: Improved Dense Policy for Effective and Generalizable Whole-body Mobile Manipulation☆33Jan 14, 2026Updated 7 months ago
- Semi-Supervised Learning for Visual Bird’s Eye View Semantic Segmentation☆15Feb 27, 2024Updated 2 years ago
- NeurIPS 2020 Spotlight Paper☆13Dec 20, 2021Updated 4 years ago
- [BMVC 2022] "SUB-Depth: Self-distillation and Uncertainty Boosting Self-supervised Monocular Depth Estimation"☆16Apr 19, 2022Updated 4 years ago
- (Pattern Recognition 2025) Towards Trustworthy Dataset Distillation☆14Dec 8, 2024Updated last year
- several examples of the learning of the java☆11Nov 22, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [RAL 2022] S2G2: Semi-Supervised Semantic Bird-Eye-View Grid-Map Generation Using a Monocular Camera for Autonomous Driving☆11Nov 23, 2022Updated 3 years ago
- ☆32May 17, 2026Updated 3 months ago
- Preview code of ECCV'24 paper "Distill Gold from Massive Ores" (BiLP)☆25Jul 6, 2024Updated 2 years ago
- Official codebase for our NeurIPS paper, Symmetry-Informed Governing Equation Discovery.☆11Nov 13, 2024Updated last year
- ☆28Jun 12, 2023Updated 3 years ago
- torch.optim.lr_scheduler☆10Mar 17, 2020Updated 6 years ago
- Not All Patches Are Equal: Hierarchical Dataset Condensation for Single Image Super-Resolution☆10May 7, 2024Updated 2 years ago
- [ICCV 2023] DataDAM: Efficient Dataset Distillation with Attention Matching☆34Jun 20, 2024Updated 2 years ago
- Data and Baselines for AStitchInLanguageModels dataset☆13Oct 31, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Implementation of Mutan+ArticleNet on OKVQA☆10Jan 11, 2021Updated 5 years ago
- 本项目主要是2025届浙江大学软件学院夏令营(AI营)的考核项目☆12Mar 3, 2025Updated last year
- A curated list of awesome papers on dataset distillation and related applications.☆1,970Updated this week
- The code for our work☆18Apr 7, 2024Updated 2 years ago
- Official code of "ALIM: Adjusting Label Importance Mechanism for Noisy Partial Label Learning"☆23Sep 25, 2023Updated 2 years ago
- ☆12Jul 30, 2025Updated last year
- [EMNLP 2025 main 🔥] Code for "Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More"☆121Oct 12, 2025Updated 10 months ago