[ICLR2025] Code Release of Refining CLlP's Spatial Awareness: A Visual-centric Perspective
☆21Apr 11, 2025Updated last year
Alternatives and similar repositories for CLIPRefiner
Users that are interested in CLIPRefiner are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Implement of the work "Coherent and Multi-modality Image Inpainting via Latent Space Optimization"☆54Apr 10, 2025Updated last year
- Code for Point-Level Regin Contrast (https//arxiv.org/abs/2202.04639)☆35Dec 23, 2022Updated 3 years ago
- Code for **Spatiotemporal Self-supervised Learning for Point Clouds in the Wild** (STSSL) CVPR2023☆49Mar 4, 2024Updated 2 years ago
- ☆20Jun 4, 2025Updated last year
- ☆12May 26, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A curated publication list on visual dialog☆14May 8, 2023Updated 3 years ago
- Language planning labeling for VLA-OS☆16Jun 25, 2025Updated last year
- A minimalist (educational) implementation of Latent Diffusion Models (LDM) with PyTorch distributed training.☆13Dec 22, 2024Updated last year
- ☆15Feb 23, 2023Updated 3 years ago
- Supervised Training of Conditional Monge Maps☆19Oct 30, 2023Updated 2 years ago
- differentiable top-k operator☆23Dec 30, 2024Updated last year
- [ICCV 2025] Unbiased Region-Language Alignment for Open-Vocabulary Dense Prediction☆53Sep 22, 2025Updated 10 months ago
- Official code for "Vision Transformers with Self-Distilled Registers" (NeurIPS 2025 Spotlight)☆35Dec 6, 2025Updated 7 months ago
- ☆100Aug 28, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official implementation of "Streaming Communication in Multi-Agent Reasoning"☆34Jun 6, 2026Updated last month
- Let's learn how to do 3D stuff with Kivy!☆19Jan 19, 2017Updated 9 years ago
- ☆31Jun 30, 2026Updated 3 weeks ago
- ☆30Jul 17, 2026Updated last week
- [CVPR 2022 Oral] Towards Open Set Temporal Action Localization☆55Sep 4, 2023Updated 2 years ago
- PyTorch implementation of paper "ARTrack" and "ARTrackV2"☆317Oct 20, 2025Updated 9 months ago
- PyTorch implementation of "UNIT: Unifying Image and Text Recognition in One Vision Encoder", NeurlPS 2024.☆34Sep 26, 2024Updated last year
- ☆53Jan 3, 2023Updated 3 years ago
- [CVPR 2023] Official repository for paper "Stare at What You See: Masked Image Modeling without Reconstruction"☆71Jul 2, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆39Apr 16, 2025Updated last year
- The offical repository of "So-Fake: Benchmarking and Explaining Social Media Image Forgery Detection"☆34Oct 29, 2025Updated 8 months ago
- Official Implementation for *PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward Modeling*☆42Dec 13, 2025Updated 7 months ago
- ZJUT的保研分享库☆36Mar 12, 2025Updated last year
- AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding☆50Jun 5, 2026Updated last month
- repository containing analysis scripts and auxiliary files☆40Apr 9, 2020Updated 6 years ago
- Implementation of the proposed LVMAE, from the paper, Extending Video Masked Autoencoders to 128 frames, in Pytorch☆55Nov 25, 2024Updated last year
- Mini-Kinetics-200 data splits used in paper "Rethinking Spatiotemporal Feature Learning For Video Understanding"☆80Dec 24, 2017Updated 8 years ago
- This is the official code repo for GLOVER and GLOVER++.☆58Aug 6, 2025Updated 11 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ICRA 2026] Official implemetation of the paper "InSpire: Vision-Language-Action Models with Intrinsic Spatial Reasoning"☆51Feb 2, 2026Updated 5 months ago
- Official codes for paper: Localizing Anomalies from Weakly-Labeled Videos☆85Sep 23, 2022Updated 3 years ago
- Official codebase for Margin-aware Preference Optimization for Aligning Diffusion Models without Reference (MaPO).☆83Jun 11, 2024Updated 2 years ago
- ☆58Jul 8, 2025Updated last year
- Hierarchical Universal Language Conditioned Policies☆78Mar 19, 2024Updated 2 years ago
- ☆55Dec 10, 2025Updated 7 months ago
- Displays plots on a graph. DEPRECATED for https://github.com/kivy-garden/graph☆55Jun 3, 2018Updated 8 years ago