[CVPR 2024] Code for "Improved Visual Grounding through Self-Consistent Explanations".
☆28Mar 1, 2024Updated 2 years ago
Alternatives and similar repositories for SelfEQ
Users that are interested in SelfEQ are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆37Mar 24, 2026Updated 6 months ago
- The official Pytorch Implementation for ElasticDiffusion: Training-free Arbitrary Size Image Generation through Global-Local Content Sepa…☆160Dec 24, 2024Updated last year
- [ICME 2024 Oral] DARA: Domain- and Relation-aware Adapters Make Parameter-efficient Tuning for Visual Grounding☆22Feb 26, 2025Updated last year
- [AAAI 2024]Weakly Supervised Multimodal Affordance Grounding for Egocentric Images☆13Nov 10, 2024Updated last year
- Code of "What Images are More Memorable to Machines?"☆15Feb 13, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Codes for the AAAI 2023 paper (Oral) "Efficient Mirror Detection via Multi-level Heterogeneous Learning" https://arxiv.org/pdf/2211.1564…☆15Jan 18, 2023Updated 3 years ago
- [EMNLP 2017] Obj2Text: Generating Visually Descriptive Language from Object Layouts☆10Jul 26, 2017Updated 9 years ago
- AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation☆16Aug 3, 2025Updated last year
- Implementation for the project: Variational Image Captioning Using Deterministic Attention☆13Dec 14, 2018Updated 7 years ago
- ☆18Jul 16, 2024Updated 2 years ago
- [ICCV 2023] Distilling Coarse-to-fine Semantic Matching Knowledge for Weakly Supervised 3D Visual Grounding☆14Oct 2, 2024Updated last year
- [CVPR 2024] Tune-An-Ellipse: CLIP Has Potential to Find What You Want☆14Jan 5, 2025Updated last year
- Exploiting Inter-sample and Inter-feature Relations in Dataset Distillation (CVPR24)☆10Jun 16, 2024Updated 2 years ago
- M2-Reasoning: Empowering MLLMs with Unified General and Spatial Reasoning☆46Jul 17, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆40Jun 28, 2023Updated 3 years ago
- This is the repo for the work "Where and What: Driver Attention-based Object Detection".☆10May 10, 2022Updated 4 years ago
- Official code for CVPR 2024 paper, "SC-Tune: Unleashing Self-Consistent Referential Comprehension in Large Vision Language Models"☆16Apr 22, 2024Updated 2 years ago
- ☆48Apr 16, 2026Updated 5 months ago
- Official Implementation of "Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Func…☆31Dec 2, 2024Updated last year
- [ECCV'24] Official Implementation of Autoregressive Visual Entity Recognizer.☆14Mar 2, 2024Updated 2 years ago
- Git for "Stepwise Self-Consistent Mathematical Reasoning with Large Language Models"☆12Nov 26, 2024Updated last year
- CoNeTTE: An efficient Audio Captioning system leveraging multiple datasets with Task Embedding☆23Dec 17, 2025Updated 9 months ago
- A reading list of papers about Visual Grounding.☆31Aug 24, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 华中科技大学计算机视觉课程实验☆17Sep 29, 2024Updated last year
- [AAAI 2024] Mono3DVG: 3D Visual Grounding in Monocular Images, AAAI, 2024☆73Apr 9, 2024Updated 2 years ago
- ☆21Apr 2, 2024Updated 2 years ago
- ☆14Jul 13, 2021Updated 5 years ago
- [CVPR 2023] Code for "Improving Visual Grounding by Encouraging Consistent Gradient-based Explanations"☆19Oct 10, 2023Updated 2 years ago
- Official implement of "Point Long-Term Locality-Aware Transformer for Point Cloud Video Understanding"☆28Mar 24, 2026Updated 6 months ago
- [CVPR'24] Code for Emergent Open-Vocabulary Semantic Segmentation from Off-the-shelf Vision-Language Models☆18Jul 22, 2024Updated 2 years ago
- Uncertainty-Aware Curriculum Learning for Neural Machine Translation (ACL 2020)☆11Jun 12, 2020Updated 6 years ago
- ☆24Jul 8, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for paper: Reinforced Vision Perception with Tools☆73Oct 3, 2025Updated 11 months ago
- SINDER: Repairing the Singular Defects of DINOv2 (ECCV 2024 Oral)☆41Oct 17, 2025Updated 11 months ago
- [CVPR 2024] Do you remember? Dense Video Captioning with Cross-Modal Memory Retrieval☆67Jun 19, 2024Updated 2 years ago
- Source Code of the ROAD benchmark for feature attribution methods (ICML22)☆26Jun 26, 2023Updated 3 years ago
- Code for the "Long Context Needs Some R&R" paper.☆12Mar 11, 2024Updated 2 years ago
- Human Pose Classification☆17Feb 19, 2023Updated 3 years ago
- A personal reimplementation with TensorFlow of NIPS2018 paper: Joint Autoregressive and Hierarchical Priors for Learned Image Compression☆15Jan 17, 2023Updated 3 years ago