How Good is Google Bard's Visual Understanding? An Empirical Study on Open Challenges
☆30Sep 11, 2026Updated 2 weeks ago
Alternatives and similar repositories for GoogleBard-VisUnderstand
Users that are interested in GoogleBard-VisUnderstand are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ARB: A Comprehensive Arabic Multimodal Reasoning Benchmark☆17May 25, 2025Updated last year
- [MICCAI 2024 🔥] HLSS, the first study to explore hierarchical information inherent in histopathology images and their language descripti…☆28Aug 5, 2024Updated 2 years ago
- PyTorch code for "ADEM-VL: Adaptive and Embedded Fusion for Efficient Vision-Language Tuning"☆21Oct 28, 2024Updated last year
- Official repository for "Boosting Adversarial Transferability using Dynamic Cues " (ICLR 2023)☆20Aug 24, 2023Updated 3 years ago
- A Benchmark and Agentic Framework for Omni-Modal Reasoning and Tool Use in Long Videos☆26Jun 20, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆20Dec 29, 2020Updated 5 years ago
- Concealed Object Detection, IEEE TPAMI 2021 (CVPR2020 extended version).☆16Feb 21, 2022Updated 4 years ago
- Perceptual Grouping in Contrastive Vision-Language Models (ICCV'23)☆37Jan 1, 2024Updated 2 years ago
- [ACM MM-2024] RefMask3D: Language-Guided Transformer for 3D Referring Segmentation☆65Jul 29, 2024Updated 2 years ago
- This is a python library. Install with "python3 -m pip install rp" then run with "python3 -m rp" or just "rp". Requires python≥3.5☆13Jul 13, 2026Updated 2 months ago
- ☆29Jan 23, 2024Updated 2 years ago
- MIMIC: Masked Image Modeling with Image Correspondences☆16Jun 14, 2024Updated 2 years ago
- ☆14Jun 25, 2022Updated 4 years ago
- Original code base for On Pretraining Data Diversity for Self-Supervised Learning☆14Dec 30, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Python package to download and use the SSB datasets☆11Aug 3, 2023Updated 3 years ago
- repo for paper titled: Towards Realistic Zero-Shot Classification via Self Structural Semantic Alignment (AAAI'24 Oral)☆25May 16, 2024Updated 2 years ago
- ☆14Aug 14, 2026Updated last month
- [ECCV 2024] Official Implementation of CoPT: Unsupervised Domain Adaptive Segmentation using Domain-Agnostic Text Embeddings☆10Feb 24, 2025Updated last year
- Official repository for "Stylized Adversarial Training" (TPAMI 2022)☆11Dec 30, 2022Updated 3 years ago
- ☆20Aug 21, 2026Updated last month
- TerraFM is a scalable foundation model for unified multisensor Earth observation, trained on 18.7M Sentinel-1/2 tiles and achieving state…☆51Jun 9, 2025Updated last year
- Official repo for “Unlocking Attributes' Contribution to Successful Camouflage: A Combined Textual and Visual Analysis Strategy”☆14Nov 26, 2024Updated last year
- A One-key fast evaluation on saliency object detection with Muti-thread and GPU implementation including MAE, Max F-measure, S-measure, E…☆11Apr 20, 2019Updated 7 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆44Jul 18, 2022Updated 4 years ago
- 🏠🔍 Auto check for new apartments in Hamburg from various real estate provides☆16Apr 15, 2026Updated 5 months ago
- Official repository for "On Improving Adversarial Transferability of Vision Transformers" (ICLR 2022--Spotlight)☆73Nov 19, 2022Updated 3 years ago
- [MICCAI 2025] Hierarchical Self-Supervised Adversarial Training for Robust Vision Models in Histopathology☆12Jun 17, 2025Updated last year
- ☆27May 6, 2024Updated 2 years ago
- Lidar Panoptic Segmentation without Bells and Whistles (IROS 2023)☆25Oct 21, 2023Updated 2 years ago
- Official implementation for MGN☆20Dec 22, 2022Updated 3 years ago
- A new multi-task learning framework using Vision Transformers☆11Jun 19, 2024Updated 2 years ago
- Submission to the inverse scaling prize☆23Jul 23, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [NeurIPS 2024] RaVL: Discovering and Mitigating Spurious Correlations in Fine-Tuned Vision-Language Models☆32Nov 12, 2024Updated last year
- ☆25Sep 19, 2023Updated 3 years ago
- ☆11Oct 29, 2024Updated last year
- 【ICCV 2023】Towards Instance-adaptive Inference for Federated Learning☆12Mar 31, 2025Updated last year
- ☆27Aug 28, 2023Updated 3 years ago
- Answering Ambiguous Questions via Iterative Prompting☆13May 25, 2024Updated 2 years ago
- Official repository for "A Self-supervised Approach for Adversarial Robustness" (CVPR 2020--Oral)☆101Apr 30, 2021Updated 5 years ago