IMProv: Inpainting-based Multimodal Prompting for Computer Vision Tasks
☆57Sep 26, 2024Updated last year
Alternatives and similar repositories for IMProv
Users that are interested in IMProv are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of the paper "EgoPet: Egomotion and Interaction Data from an Animal's Perspective".☆29Dec 15, 2025Updated 3 months ago
- ☆41Jul 19, 2024Updated last year
- Syphus: Automatic Instruction-Response Generation Pipeline☆14Dec 14, 2023Updated 2 years ago
- Official implementation and data release of the paper "Visual Prompting via Image Inpainting".☆317Aug 7, 2023Updated 2 years ago
- [CVPR23] "Understanding and Improving Visual Prompting: A Label-Mapping Perspective" by Aochuan Chen, Yuguang Yao, Pin-Yu Chen, Yihua Zha…☆52Sep 17, 2023Updated 2 years ago
- DexPoint: Generalizable Point Cloud Reinforcement Learning for Sim-to-Real Dexterous Manipulation, CoRL 2022☆100May 22, 2024Updated last year
- a starter-kit for jaynes, the cloud-agnostic launch library☆17Aug 20, 2024Updated last year
- TTRV: Test-Time Reinforcement Learning for Vision–Language Models (CVPR 2026)☆37Mar 8, 2026Updated 2 weeks ago
- 🕊️ HATO: Learning Visuotactile Skills with Two Multifingered Hands [ICRA 2025]☆168May 27, 2024Updated last year
- (ICCV 2023) MasQCLIP for Open-Vocabulary Universal Image Segmentation☆38Oct 18, 2023Updated 2 years ago
- AllSight, is an optical tactile sensor with a round 3D structure, potentially designed for robotic in-hand manipulation tasks☆17Nov 28, 2025Updated 3 months ago
- [NeurIPS2023] Official implementation of the paper "Large Language Models are Visual Reasoning Coordinators"☆105Nov 9, 2023Updated 2 years ago
- Isaac Gym Python Stubs for Code Completion☆126Jun 10, 2024Updated last year
- ☆28Jun 5, 2025Updated 9 months ago
- [CVPR 2022] Joint hand motion and interaction hotspots prediction from egocentric videos☆71Jan 29, 2024Updated 2 years ago
- Code of CropMix: Sampling a Rich Input Distribution via Multi-Scale Cropping☆17Oct 8, 2022Updated 3 years ago
- Chain-of-Thought Predictive Control☆57May 1, 2023Updated 2 years ago
- ☆48Apr 25, 2024Updated last year
- [IJCV 2024] MosaicFusion: Diffusion Models as Data Augmenters for Large Vocabulary Instance Segmentation☆128Oct 8, 2024Updated last year
- OVSegmentor, CVPR23☆61Apr 22, 2024Updated last year
- Code for "Hierarchical World Models as Visual Whole-Body Humanoid Controllers"☆204Sep 18, 2025Updated 6 months ago
- PyTorch implementation of "Sample- and Parameter-Efficient Auto-Regressive Image Models" from CVPR 2025☆14Nov 21, 2025Updated 4 months ago
- Referring Image Segmentation Benchmarking with Segment Anything Model (SAM)☆38Apr 7, 2023Updated 2 years ago
- Code for Abstract-to-Executable Trajectory Translation for One Shot Task Generalization (ICML 2023)☆23May 12, 2023Updated 2 years ago
- Community-built test set to benchmark QP solvers☆15May 7, 2025Updated 10 months ago
- Vision package for robot manipulation and learning research☆25Apr 21, 2024Updated last year
- [ECCV 2024] Learning Video Context as Interleaved Multimodal Sequences☆43Mar 11, 2025Updated last year
- Directed masked autoencoders☆14Mar 17, 2026Updated last week
- MoDem Accelerating Visual Model-Based Reinforcement Learning with Demonstrations☆87Dec 12, 2022Updated 3 years ago
- Twisting Lids Off with Two Hands [CoRL 2024]☆39Mar 16, 2025Updated last year
- [TMLR 2025] The official repository of the paper "Unsupervised Discovery of Object-Centric Neural Fields"☆18Feb 15, 2026Updated last month
- Contrastive Learning of Image Representations with Cross-Video Cycle-Consistency☆17Dec 2, 2021Updated 4 years ago
- Code Release of "3D Concept Grounding on Neural Fields (NeurIPS2022)"☆15Feb 13, 2023Updated 3 years ago
- MCC-HO☆56Dec 2, 2024Updated last year
- ☆18Feb 19, 2024Updated 2 years ago
- LLMBind: A Unified Modality-Task Integration Framework☆19Jun 16, 2024Updated last year
- ☆13Sep 4, 2023Updated 2 years ago
- The Structure and Interpretation of Deep Networks Handbook☆14Dec 14, 2024Updated last year
- Official repo for the TMLR paper "Discffusion: Discriminative Diffusion Models as Few-shot Vision and Language Learners"☆29Apr 27, 2024Updated last year