☆51May 24, 2023Updated 3 years ago
Alternatives and similar repositories for LLaVA
Users that are interested in LLaVA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [RSS 23] Dynamic-Resolution Model Learning for Object Pile Manipulation☆37Jan 29, 2024Updated 2 years ago
- Master of Research project investigating deep learning models for the estimation of fruit quality attributes using Near-infrared (NIR) S…☆19Nov 5, 2024Updated last year
- [ECCV24] VISA: Reasoning Video Object Segmentation via Large Language Model☆22Jul 20, 2024Updated 2 years ago
- [EMNLP 2025] Official code for the paper "SafeKey: Amplifying Aha-Moment Insights for Safety Reasoning"☆16May 12, 2026Updated 4 months ago
- Pluggin and utils for viewing voxelgrids in RViz☆11Jun 14, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- C++ and Python utilities. ARC -> ARM☆13Sep 3, 2025Updated last year
- ☆24Feb 5, 2024Updated 2 years ago
- Object Detection in images using Selective Search and EdgeBoxes algorithm☆33Oct 4, 2019Updated 6 years ago
- Beyond Known Clusters: Probe New Prototypes for Efficient Generalized Class Discovery☆15Apr 28, 2024Updated 2 years ago
- ☆15Aug 3, 2021Updated 5 years ago
- A Simple Framwork for CV Pre-training Model (SOCO, VirTex, BEiT)☆15Oct 18, 2021Updated 4 years ago
- [CVPR 2025] COSMOS: Cross-Modality Self-Distillation for Vision Language Pre-training☆42Mar 27, 2025Updated last year
- ☆10May 26, 2022Updated 4 years ago
- ☆11Nov 5, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [TPAMI 2023] Object Affinity Learning: Towards Annotation-free Instance Segmentation☆14Sep 14, 2023Updated 3 years ago
- ☆11May 24, 2024Updated 2 years ago
- [CVPR 24] This is official implication for our paper: ''CroSel: Cross Selection of Confident Pseudo Labels for Partial-Label Learning''.☆15Apr 27, 2025Updated last year
- Code for Enhancing Self-supervised Video Representation Learning via Multi-level Feature Optimization.☆10Sep 28, 2021Updated 4 years ago
- ☆21Jul 5, 2024Updated 2 years ago
- [AAAI 2026] A²LC: Active and Automated Label Correction for Semantic Segmentation☆16Jun 25, 2026Updated 2 months ago
- This is the official code for NeurIPS 2023 paper "Learning Unseen Modality Interaction"☆18Jan 22, 2024Updated 2 years ago
- multi-bit language model watermarking (NAACL 24)☆22Sep 20, 2024Updated 2 years ago
- Pytorch implementation of ICML-2024 "Navigating Complexity: Toward Lossless Graph Condensation via Expanding Window Matching"☆25Jun 23, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Framework for Symbolic MUsic Graph Explanations☆11Jul 30, 2025Updated last year
- [CVPR 2021] Semi-Supervised Indoor Layout Estimation from 360-Degree Panorama☆12Jun 1, 2021Updated 5 years ago
- ☆18Jul 6, 2026Updated 2 months ago
- Historical Report Guided Bi-modal Concurrent Learning for Pathology Report Generation☆15Nov 24, 2025Updated 9 months ago
- Repository for "Training Audio Captioning Models without Audio"☆10Sep 26, 2023Updated 2 years ago
- ☆14Oct 23, 2018Updated 7 years ago
- ☆10Sep 25, 2024Updated last year
- ☆39Nov 25, 2025Updated 9 months ago
- Official repository of the "Active Learning for Semantic Segmentation with Multi-class Label Query (NeurIPS'23)"☆16Jan 16, 2024Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ICCV 2025] Official implementation of "InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models"☆56Feb 10, 2025Updated last year
- ☆11Feb 15, 2019Updated 7 years ago
- "Visual Prompt Selection for In-Context Learning Segmentation Framework"☆14Dec 13, 2024Updated last year
- Reproduced the DFT method without using Verl. https://arxiv.org/abs/2508.05629☆24Oct 14, 2025Updated 11 months ago
- 🎭 Official code and dataset for our CCGPK@COLING 2022 paper - "PersonaChatGen: Generating Personalized Dialogue using GPT-3"☆13Mar 26, 2024Updated 2 years ago
- CaMML:Context-Aware MultiModal Learner for Large Models (ACL 2024 SAC Award)☆15May 21, 2025Updated last year
- PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails☆22Jul 8, 2026Updated 2 months ago