ARB: A Comprehensive Arabic Multimodal Reasoning Benchmark
β17May 25, 2025Updated last year
Alternatives and similar repositories for ARB
Users that are interested in ARB are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2025 π₯] Time Travel is a Comprehensive Benchmark to Evaluate LMMs on Historical and Cultural Artifactsβ19May 22, 2025Updated last year
- VideoMathQA is a benchmark designed to evaluate mathematical reasoning in real-world educational videosβ25Sep 5, 2026Updated 3 weeks ago
- AIN - The First Arabic Inclusive Large Multimodal Model. It is a versatile bilingual LMM excelling in visual and contextual understandingβ¦β54Mar 13, 2025Updated last year
- [MICCAI 2025] Hierarchical Self-Supervised Adversarial Training for Robust Vision Models in Histopathologyβ12Jun 17, 2025Updated last year
- Learnable Weight Initialization for Volumetric Medical Image Segmentation [Elsevier AIM2024]β22Oct 27, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Reasoning DriveLMMβ16Mar 15, 2025Updated last year
- [NAACL 2025 π₯] CAMEL-Bench is an Arabic benchmark for evaluating multimodal models across eight domains with 29,000 questions.β38Apr 17, 2025Updated last year
- [MICCAI 2023][Early Accept] Official code repository of paper titled "Cross-modulated Few-shot Image Generation for Colorectal Tissue Claβ¦β46Sep 28, 2023Updated 2 years ago
- [ICCVW 2025 (Oral)] Robust-LLaVA: On the Effectiveness of Large-Scale Robust Image Encoders for Multi-modal Large Language Modelsβ29Sep 9, 2026Updated 2 weeks ago
- [MICCAI 2023] Official code repository of paper titled "Frequency Domain Adversarial Training for Robust Volumetric Medical Segmentation"β¦β52Nov 14, 2023Updated 2 years ago
- Composed Video Retrievalβ62May 2, 2024Updated 2 years ago
- A Large Multimodal Model for Remote Sensing Change Description (IGARSS 2025)β22Dec 17, 2025Updated 9 months ago
- [EMNLP'23] ClimateGPT: a specialized LLM for conversations related to Climate Change and Sustainability topics in both English and Arabiβ¦β80Sep 24, 2024Updated 2 years ago
- (ICCV 2023) Generative Multiplane Neural Radiance for 3D Aware Image Generation.β18Sep 28, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR 2025 π₯] ALM-Bench is a multilingual multi-modal diverse cultural benchmark for 100 languages across 19 categories. It assesses theβ¦β47Sep 5, 2026Updated 3 weeks ago
- Video-CoM: Interactive Video Reasoning via Chain of Manipulationsβ23Sep 5, 2026Updated 3 weeks ago
- A new multi-task learning framework using Vision Transformersβ11Jun 19, 2024Updated 2 years ago
- A Benchmark and Agentic Framework for Omni-Modal Reasoning and Tool Use in Long Videosβ26Jun 20, 2026Updated 3 months ago
- How Good is Google Bard's Visual Understanding? An Empirical Study on Open Challengesβ30Sep 11, 2026Updated 2 weeks ago
- Language Grounded Single Source Domain Generalization in Medical Image Segmentation [ISBI2024]β34Oct 27, 2024Updated last year
- [CVPR 2025 π₯]A Large Multimodal Model for Pixel-Level Visual Grounding in Videosβ106Sep 5, 2026Updated 3 weeks ago
- [BMVC 2024] On Evaluating Adversarial Robustness of Volumetric Medical Segmentation Modelsβ15Nov 1, 2024Updated last year
- Reinforcement Training of Robotβ11Dec 1, 2019Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [CVPRW-25 MMFM] Official repository of paper titled "How Good is my Video LMM? Complex Video Reasoning and Robustness Evaluation Suite foβ¦β50Aug 23, 2024Updated 2 years ago
- [MICCAI 2024 π₯] HLSS, the first study to explore hierarchical information inherent in histopathology images and their language descriptiβ¦β28Aug 5, 2024Updated 2 years ago
- [MICCAI 2024] Official code repository of paper titled "BAPLe: Backdoor Attacks on Medical Foundation Models using Prompt Learning" accepβ¦β57Oct 22, 2024Updated last year
- Kaggle Competition Dstl Satellite Imagery Feature Detectionβ10Apr 1, 2017Updated 9 years ago
- Self Evolving Large Multimodal Models with Continuous Rewardsβ27Sep 5, 2026Updated 3 weeks ago
- [CVPR -2025] GroupMamba: Parameter-Efficient and Accurate Group Visual State Space Modelβ143Mar 22, 2025Updated last year
- Official repository for "Boosting Adversarial Transferability using Dynamic Cues " (ICLR 2023)β20Aug 24, 2023Updated 3 years ago
- [NeurIPS2023] 3D-OWIS is capable of detecting unknown instances in inference, and progressively learning novel classes in the process of β¦β68Dec 3, 2023Updated 2 years ago
- (BMVC 2022--Oral) Official repository for "Adversarial Pixel Restoration as a Pretext Task for Transferable Perturbations" β¦β35Jan 8, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ACL 2025 π₯] A Comprehensive Multi-Domain Benchmark for Arabic OCR and Document Understandingβ79Aug 10, 2026Updated last month
- OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoningβ29May 24, 2025Updated last year
- [CVPRW 2025] Official repository of paper titled "Towards Evaluating the Robustness of Visual State Space Models"β25Jun 8, 2025Updated last year
- [ICML2026] Official Implementations "FasterVAR: Plug-and-Play Acceleration for Visual Autoregressive Models"β30Aug 8, 2026Updated last month
- [WACV 2025] Efficient Video Object Segmentation via Modulated Cross-Attention Memoryβ61Feb 28, 2025Updated last year
- Abstract. Person search is a challenging problem with various real- world applications, that aims at joint person detection and re-identiβ¦β13Feb 28, 2024Updated 2 years ago
- Official repository for "Stylized Adversarial Training" (TPAMI 2022)β11Dec 30, 2022Updated 3 years ago