Reproduction of LLaVA-v1.5 based on Llama-3-8b LLM backbone.
☆64Oct 25, 2024Updated last year
Alternatives and similar repositories for LLaVA-Llama-3
Users that are interested in LLaVA-Llama-3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCVW 25] LLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning☆160Aug 8, 2025Updated last year
- code for Learning the Unlearned: Mitigating Feature Suppression in Contrastive Learning☆20Jul 16, 2024Updated 2 years ago
- Fast LLM Training CodeBase With dynamic strategy choosing [Deepspeed+Megatron+FlashAttention+CudaFusionKernel+Compiler];☆42Jan 4, 2024Updated 2 years ago
- [ICLR2025] LLaVA-HR: High-Resolution Large Language-Vision Assistant☆249Aug 14, 2024Updated 2 years ago
- ☆17Oct 21, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- This is the official code for the paper "Reconstruct before Query: Continual Missing Modality Learning with Decomposed Prompt Collaborati…☆12Aug 13, 2024Updated 2 years ago
- ☆23Aug 27, 2025Updated 11 months ago
- 基于PaddlePaddle以及wechaty框架 建立的宇宙漫游指南机器人☆17Aug 3, 2021Updated 5 years ago
- ☆10Nov 29, 2022Updated 3 years ago
- ☆20Oct 19, 2023Updated 2 years ago
- ☆10Sep 25, 2019Updated 6 years ago
- Official code of *Virgo: A Preliminary Exploration on Reproducing o1-like MLLM*☆20May 27, 2025Updated last year
- ☆24May 23, 2025Updated last year
- The efficient tuning method for VLMs☆83Mar 10, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆15Oct 27, 2023Updated 2 years ago
- A simple Web AI model deployment tool using JavaScript based on OpenCV.js and ONNXRuntime☆58Jul 8, 2024Updated 2 years ago
- [ECCV 2024] Official PyTorch implementation of DreamLIP: Language-Image Pre-training with Long Captions☆138May 8, 2025Updated last year
- ☆20Apr 8, 2025Updated last year
- 天池 NVIDIA TensorRT Hackathon 2023 —— 生成式AI模型优化赛 初赛第三名方案☆50Aug 16, 2023Updated 3 years ago
- Official repo for ICML 2025 paper "RollingQ: Reviving the Cooperation Dynamics in Multimodal Transformer"☆17Jun 21, 2025Updated last year
- ☆35Jan 21, 2025Updated last year
- Setup scripts for the WebArena benchmark☆22Jun 19, 2025Updated last year
- [WIP@Oct 13] 质衡-基准测试 (Q-Bench in Chinese),包含中文版【底层视觉问答】和【底层视觉描述】数据集,以及中文提示下的图片质量评价。 We will release Q-Bench in more languages in the futu…☆24Jan 7, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- On-Device Domain Generalization☆47Nov 9, 2022Updated 3 years ago
- 🧩 Official code repository for “Puzzled by Puzzles: When Vision-Language Models Can’t Take a Hint.”☆15Sep 22, 2025Updated 10 months ago
- ☆26Feb 2, 2025Updated last year
- This repository contains the code for the paper “Neuro-Symbolic Query Compiler”, accepted to the Findings of ACL 2025.☆18Oct 20, 2025Updated 9 months ago
- ☆13Oct 23, 2024Updated last year
- [ICLR 2024 Spotlight] DreamLLM: Synergistic Multimodal Comprehension and Creation☆462Dec 2, 2024Updated last year
- Official code for the CVPR2019 workshop paper -- Scan-flood-Fill: an Efficient Automatic Precise Region Filling Algorithm for Complicated…☆17Jun 17, 2019Updated 7 years ago
- Dynamic, high-resolution poverty measurement in data-scarce environments☆11Dec 8, 2024Updated last year
- ☆4,714Jun 15, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- a TensorFlow implementation of the paper "Feature Super-Resolution Based Facial Expression Recognition for Multi-scale Low-Resolution Ima…☆13Nov 30, 2021Updated 4 years ago
- [CVPR 2024 🔥] Grounding Large Multimodal Model (GLaMM), the first-of-its-kind model capable of generating natural language responses tha…☆967Aug 5, 2025Updated last year
- Masking tokens to modify the predictions of a pretrained sentence classifier☆16Feb 4, 2020Updated 6 years ago
- Official Repo for FoodieQA paper (EMNLP 2024)☆20Jun 26, 2025Updated last year
- Official repository for Robust Multimodal Large Language Models Against Modality Conflict☆22Jul 9, 2025Updated last year
- ☆10Sep 21, 2024Updated last year
- ☆10Jul 11, 2022Updated 4 years ago