Reproduction of LLaVA-v1.5 based on Llama-3-8b LLM backbone.
☆64Oct 25, 2024Updated last year
Alternatives and similar repositories for LLaVA-Llama-3
Users that are interested in LLaVA-Llama-3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCVW 2025] LLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning☆160Aug 8, 2025Updated last year
- code for Learning the Unlearned: Mitigating Feature Suppression in Contrastive Learning☆20Jul 16, 2024Updated 2 years ago
- 🔥🔥 LLaVA++: Extending LLaVA with Phi-3 and LLaMA-3 (LLaVA LLaMA-3, LLaVA Phi-3)☆842Sep 5, 2026Updated 3 weeks ago
- ☆12Dec 20, 2024Updated last year
- Fast LLM Training CodeBase With dynamic strategy choosing [Deepspeed+Megatron+FlashAttention+CudaFusionKernel+Compiler];☆42Jan 4, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR2025] LLaVA-HR: High-Resolution Large Language-Vision Assistant☆249Aug 14, 2024Updated 2 years ago
- MemSearcher is a search agent that keeps a compact, iteratively-updated memory instead of the full interaction history, trained end-to-en…☆29Jun 29, 2026Updated 2 months ago
- [NeurIPS 2024] MoVA: Adapting Mixture of Vision Experts to Multimodal Context☆174Sep 25, 2024Updated 2 years ago
- ☆17Oct 21, 2024Updated last year
- This is the official code for the paper "Reconstruct before Query: Continual Missing Modality Learning with Decomposed Prompt Collaborati…☆12Aug 13, 2024Updated 2 years ago
- ☆20Oct 19, 2023Updated 2 years ago
- Official code of *Virgo: A Preliminary Exploration on Reproducing o1-like MLLM*☆20May 27, 2025Updated last year
- ☆24May 23, 2025Updated last year
- 飞桨模型加密库☆10Nov 13, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 基于LLaVA1.6微调的Xray识别的多模态大模型☆10Oct 22, 2024Updated last year
- A toolkit for automated alignment research.☆16Jul 3, 2026Updated 2 months ago
- The efficient tuning method for VLMs☆84Mar 10, 2024Updated 2 years ago
- ☆15Oct 27, 2023Updated 2 years ago
- [ECCV 2024] Official PyTorch implementation of DreamLIP: Language-Image Pre-training with Long Captions☆139May 8, 2025Updated last year
- ☆20Apr 8, 2025Updated last year
- a set of tools for computer vision processing☆18Jul 9, 2016Updated 10 years ago
- [NeurIPS 2023 Datasets and Benchmarks Track] LAMM: Multi-Modal Large Language Models and Applications as AI Agents☆318Apr 16, 2024Updated 2 years ago
- 天池 NVIDIA TensorRT Hackathon 2023 —— 生成式AI模型优化赛 初赛第三名方案☆50Aug 16, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- How to export PyTorch models with unsupported layers to ONNX and then to Intel OpenVINO☆28Feb 20, 2025Updated last year
- ☆35Jan 21, 2025Updated last year
- tensorflow Implementation of https://github.com/facebookresearch/MIXER☆11Mar 8, 2017Updated 9 years ago
- [WIP@Oct 13] 质衡-基准测试 (Q-Bench in Chinese),包含中文版【底层视觉问答】和【底层视觉描述】数据集,以及中文提示下的图片质量评价。 We will release Q-Bench in more languages in the futu…☆24Jan 7, 2024Updated 2 years ago
- On-Device Domain Generalization☆48Nov 9, 2022Updated 3 years ago
- ☆12Jun 2, 2024Updated 2 years ago
- 🧩 Official code repository for “Puzzled by Puzzles: When Vision-Language Models Can’t Take a Hint.”☆15Sep 22, 2025Updated last year
- ☆26Feb 2, 2025Updated last year
- This repository contains the code for the paper “Neuro-Symbolic Query Compiler”, accepted to the Findings of ACL 2025.☆19Oct 20, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is AlpaGasus2-QLoRA based on LLaMA2 with AlpaGasus mechanism using QLoRA!☆15Nov 22, 2023Updated 2 years ago
- [ICLR 2024 Spotlight] DreamLLM: Synergistic Multimodal Comprehension and Creation☆462Dec 2, 2024Updated last year
- Dynamic, high-resolution poverty measurement in data-scarce environments☆11Dec 8, 2024Updated last year
- ☆4,722Jun 15, 2026Updated 3 months ago
- [CVPR 2024 🔥] Grounding Large Multimodal Model (GLaMM), the first-of-its-kind model capable of generating natural language responses tha…☆970Sep 5, 2026Updated 3 weeks ago
- It analyze facial micro-expressions, provides emotional states indicative of PTSD. The tool supports real-time analysis of live video str…☆13May 7, 2024Updated 2 years ago
- Official repository for Robust Multimodal Large Language Models Against Modality Conflict☆22Jul 9, 2025Updated last year