IMProv: Inpainting-based Multimodal Prompting for Computer Vision Tasks
☆57Sep 26, 2024Updated last year
Alternatives and similar repositories for IMProv
Users that are interested in IMProv are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆49Nov 28, 2024Updated last year
- [NeurIPS2023] Official implementation and model release of the paper "What Makes Good Examples for Visual In-Context Learning?"☆182Mar 4, 2024Updated 2 years ago
- ☆41Jul 19, 2024Updated last year
- Official implementation and data release of the paper "Visual Prompting via Image Inpainting".☆319Aug 7, 2023Updated 2 years ago
- DexPoint: Generalizable Point Cloud Reinforcement Learning for Sim-to-Real Dexterous Manipulation, CoRL 2022☆105May 22, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [CVPR23] "Understanding and Improving Visual Prompting: A Label-Mapping Perspective" by Aochuan Chen, Yuguang Yao, Pin-Yu Chen, Yihua Zha…☆52Sep 17, 2023Updated 2 years ago
- A minimal and stable PPO.☆148Feb 9, 2024Updated 2 years ago
- TTRV: Test-Time Reinforcement Learning for Vision–Language Models (CVPR 2026)☆45Mar 8, 2026Updated 3 months ago
- 🕊️ HATO: Learning Visuotactile Skills with Two Multifingered Hands [ICRA 2025]☆169May 27, 2024Updated 2 years ago
- (ICCV 2023) MasQCLIP for Open-Vocabulary Universal Image Segmentation☆37Oct 18, 2023Updated 2 years ago
- Measuring the Signal to Noise Ratio in Language Model Evaluation☆30Aug 19, 2025Updated 10 months ago
- AllSight, is an optical tactile sensor with a round 3D structure, potentially designed for robotic in-hand manipulation tasks☆20Nov 28, 2025Updated 7 months ago
- [NeurIPS2023] Official implementation of the paper "Large Language Models are Visual Reasoning Coordinators"☆106Nov 9, 2023Updated 2 years ago
- Isaac Gym Python Stubs for Code Completion☆126Jun 10, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [CVPR 2022] Joint hand motion and interaction hotspots prediction from egocentric videos☆71Jan 29, 2024Updated 2 years ago
- Chain-of-Thought Predictive Control☆56May 1, 2023Updated 3 years ago
- ☆50Apr 25, 2024Updated 2 years ago
- ☆29Jun 5, 2025Updated last year
- Code of CropMix: Sampling a Rich Input Distribution via Multi-Scale Cropping☆17Oct 8, 2022Updated 3 years ago
- A curated list of papers & resources linked to concept learning☆13Aug 9, 2023Updated 2 years ago
- [IJCV 2024] MosaicFusion: Diffusion Models as Data Augmenters for Large Vocabulary Instance Segmentation☆129Oct 8, 2024Updated last year
- OVSegmentor, CVPR23☆62Apr 22, 2024Updated 2 years ago
- Code for "Hierarchical World Models as Visual Whole-Body Humanoid Controllers"☆211Sep 18, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- PyTorch implementation of "Sample- and Parameter-Efficient Auto-Regressive Image Models" from CVPR 2025☆14Nov 21, 2025Updated 7 months ago
- Referring Image Segmentation Benchmarking with Segment Anything Model (SAM)☆39Apr 7, 2023Updated 3 years ago
- Code for Abstract-to-Executable Trajectory Translation for One Shot Task Generalization (ICML 2023)☆23May 12, 2023Updated 3 years ago
- Community-built test set to benchmark QP solvers☆15May 7, 2025Updated last year
- [ECCV 2024] Learning Video Context as Interleaved Multimodal Sequences☆45Mar 11, 2025Updated last year
- Directed masked autoencoders☆15Mar 25, 2026Updated 3 months ago
- MoDem Accelerating Visual Model-Based Reinforcement Learning with Demonstrations☆87Dec 12, 2022Updated 3 years ago
- Twisting Lids Off with Two Hands [CoRL 2024]☆41Mar 16, 2025Updated last year
- [TMLR 2025] The official repository of the paper "Unsupervised Discovery of Object-Centric Neural Fields"☆18Feb 15, 2026Updated 4 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Contrastive Learning of Image Representations with Cross-Video Cycle-Consistency☆17Dec 2, 2021Updated 4 years ago
- Code Release of "3D Concept Grounding on Neural Fields (NeurIPS2022)"☆15Feb 13, 2023Updated 3 years ago
- ☆18Feb 19, 2024Updated 2 years ago
- ☆13Sep 4, 2023Updated 2 years ago
- LLMBind: A Unified Modality-Task Integration Framework☆19Jun 16, 2024Updated 2 years ago
- The Structure and Interpretation of Deep Networks Handbook☆14Dec 14, 2024Updated last year
- ☆33Jan 17, 2026Updated 5 months ago