Fine-tune Qwen2.5-VL-7B on custom visual QA tasks using LoRA + Accelerate, supporting single/multi-GPU training on COCO 2014 dataset.
☆30Apr 28, 2025Updated last year
Alternatives and similar repositories for Qwen2.5-VL-Finetune
Users that are interested in Qwen2.5-VL-Finetune are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [IJBHI 2024] This is the official implementation of CAMANet: Class Activation Map Guided Attention Network for Radiology Report Generati…☆11May 14, 2025Updated last year
- Experimental adapter for fine-tuning Qwen3-VL as a Vision-Language-Action (VLA) model☆15Dec 21, 2025Updated 9 months ago
- Implementation of "DeepWriter: A Multi-Stream Deep CNN for Text-independent Writer Identification"☆16Feb 3, 2020Updated 6 years ago
- The official code and model for ACL 2023 paper 'mCLIP: Multilingual CLIP via Cross-lingual Transfer'☆10Jan 23, 2024Updated 2 years ago
- ☆10Apr 7, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Demo for Qwen2.5-VL-3B-Instruct on Axera device.☆16Sep 3, 2025Updated last year
- ☆14Oct 14, 2019Updated 6 years ago
- ☆11May 16, 2025Updated last year
- [ACL Main 2025] I0T: Embedding Standardization Method Towards Zero Modality Gap☆12Jun 18, 2025Updated last year
- ☆20Jun 13, 2025Updated last year
- This is a project based on an accepted paper "Weighted Poisson-disk Resampling on Large-Scale Point Clouds"☆16Dec 19, 2024Updated last year
- ☆16May 31, 2023Updated 3 years ago
- [AAAI 2024] DenoSent: A Denoising Objective for Self-Supervised Sentence Representation Learning☆15Apr 29, 2024Updated 2 years ago
- Models for verification and identification of document writers☆18May 26, 2020Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆12May 3, 2024Updated 2 years ago
- ☆15Dec 31, 2024Updated last year
- The official PyTorch code for AAAI'23 Paper "Sparse Coding in a Dual Memory System for Lifelong Learning"☆12Feb 15, 2023Updated 3 years ago
- Codeformer Tensorrt Face Restoration☆14Apr 15, 2024Updated 2 years ago
- Cross-modal Active Complementary Learning with Self-refining Correspondence (NeurIPS 2023, Pytorch Code)☆15Jun 6, 2024Updated 2 years ago
- Codebase for VidHal: Benchmarking Hallucinations in Vision LLMs☆14Apr 23, 2026Updated 5 months ago
- Repositiory of paper "Continual Learning for LiDAR Semantic Segmentation: Class-Incremental and Coarse-to-Fine strategies on Sparse Data"☆14Oct 26, 2024Updated last year
- Coherent Point Drift Networks: Unsupervised Learning of Non-Rigid Point Set Registration (CPD-Net). Lingjing Wang, Xiang Li, Jianchun Ch…☆34Jul 19, 2019Updated 7 years ago
- ☆16Apr 21, 2016Updated 10 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Enhancing Complex Question Answering over Knowledge Graphs through Evidence Pattern Retrieval, WWW 2024☆15Oct 22, 2024Updated last year
- Holistic Coverage and Faithfulness Evaluation of Large Vision-Language Models (ACL-Findings 2024)☆16Apr 23, 2024Updated 2 years ago
- Official implementation for the paper "Transferring Visual Knowledge with Pre-trained Models for Multimodal Machine Translation", publish…☆20Jun 3, 2024Updated 2 years ago
- ☆16Sep 8, 2021Updated 5 years ago
- Reinforcing LLM Reasoning through Self-Training and Value-Guided Decoding☆19May 6, 2026Updated 4 months ago
- 基于LLaVA1.6微调的Xray识别的多模态大模型☆10Oct 22, 2024Updated last year
- ☆20Jun 16, 2021Updated 5 years ago
- Calculating Disparity Maps using openCVs implemented algorithms.☆11May 5, 2017Updated 9 years ago
- Adding Scene-Centric Forecasting Control to Occupancy World Model☆48Jul 1, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 中南大学Java+网络课设,使用JavaFx设计的邮件客户端☆10Oct 4, 2021Updated 4 years ago
- ☆16Jan 12, 2026Updated 8 months ago
- VAP-Diffusion: Enriching Descriptions with MLLMs for Enhanced Medical Image Generation☆12Apr 11, 2026Updated 5 months ago
- [ICCV 2025] MMReason, MLLMs, step by step, reasoning benchmark, AGI☆15Apr 25, 2026Updated 5 months ago
- ☆11Jul 25, 2018Updated 8 years ago
- This repository contains the source code related to the paper Compressed Volumetric Heatmaps for Multi-Person 3D Pose Estimation☆11Jun 23, 2020Updated 6 years ago
- ☆13Dec 10, 2023Updated 2 years ago