☆34Mar 2, 2025Updated last year
Alternatives and similar repositories for Qwen2.5-VL-Fine-Tuning
Users that are interested in Qwen2.5-VL-Fine-Tuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Real-time motion planner and autonomous vehicle simulator in the browser, built with WebGL and Three.js.☆13Jun 25, 2026Updated last month
- Code for 'Geospatial Entity Resolution' paper (WWW 2022)☆20Apr 27, 2023Updated 3 years ago
- [CVPR-2023 Workshop@NFVLR] Official PyTorch implementation of Learning CLIP Guided Visual-Text Fusion Transformer for Video-based Pedestr…☆30Mar 20, 2025Updated last year
- 2025.01:从零到一实现了一个多模态大模型,并命名为Reyes(睿视),R:睿,eyes:眼。Reyes的参数量为8B,视觉编码器使用的是InternViT-300M-448px-V2_5,语言模型侧使用的是Qwen2.5-7B-Instruct,Reyes也通过一个两…☆34Feb 10, 2026Updated 6 months ago
- ☆12Feb 13, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [Neural Networks 2025]Text-guided Image Restoration and Semantic Enhancement for Text-to-Image Person Retrieval☆12Dec 24, 2024Updated last year
- SparseGNV: Generating Novel Views of Indoor Scenes with Sparse Input Views☆19Feb 27, 2024Updated 2 years ago
- ☆16Nov 13, 2022Updated 3 years ago
- Official Pytorch implementation of 'Facing the Elephant in the Room: Visual Prompt Tuning or Full Finetuning'? (ICLR2024)☆13Mar 8, 2024Updated 2 years ago
- 基于多模态检索的互联网图文匹配☆15Mar 17, 2024Updated 2 years ago
- ☆13Apr 5, 2023Updated 3 years ago
- SISR: RDN & RCAN☆11Dec 23, 2024Updated last year
- 使用FastAPI+vLLM部署Qwen2.5☆25Sep 29, 2024Updated last year
- ☆11Oct 31, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR 2026] CityLens: Evaluating Large Vision-Language Models for Urban Socioeconomic Sensing☆17Jan 31, 2026Updated 6 months ago
- ☆10May 17, 2024Updated 2 years ago
- [NeurIPS 2025] The official repository of "Inst-IT: Boosting Multimodal Instance Understanding via Explicit Visual Prompt Instruction Tun…☆40Feb 20, 2025Updated last year
- An automated imageJ macro which can be used to detect and measure the size of particles in images, maps of images or videos☆12Mar 25, 2024Updated 2 years ago
- FNIN: A Fourier Neural Operator-based Numerical Integration Network for Surface-form-gradients☆13Jan 22, 2025Updated last year
- [AAAI 2026] Towards high-level person/vehicle re-id using an event camera☆18Jul 10, 2026Updated last month
- ☆14Aug 31, 2023Updated 2 years ago
- 这是一个不基于任何框架实现的从0到1的VLM finetune(包括Pre-train和SFT)☆39Aug 22, 2025Updated 11 months ago
- ☆18May 18, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- pymesh for windows☆19Aug 18, 2016Updated 9 years ago
- ☆15Mar 25, 2016Updated 10 years ago
- 1Z实验室教程源代码仓库☆12Nov 10, 2018Updated 7 years ago
- Code Releasement for 'Generative Object Insertion in Gaussian Splatting with a Multi-View Diffusion Model'☆15Apr 26, 2025Updated last year
- Monocular Distance Measurement☆26Apr 16, 2026Updated 3 months ago
- [CVPR‘26 Oral] VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation☆23May 23, 2026Updated 2 months ago
- T-Rex2: Towards Generic Object Detection via Text-Visual Prompt Synergy☆28May 13, 2025Updated last year
- ☆12Jun 10, 2025Updated last year
- Unofficial implementation of STAN paper published at ISBI 2020 by researchers from University of Idaho using Tensorflow Keras 2.0.☆12Jun 8, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Latex template for poster☆12Sep 6, 2023Updated 2 years ago
- A survey on MM-LLMs for long video understanding: From Seconds to Hours: Reviewing MultiModal Large Language Models on Comprehensive Long…☆25Sep 12, 2025Updated 11 months ago
- 无人飞行器智能感知技术竞赛-精准定位☆26Dec 8, 2023Updated 2 years ago
- C++ OpenGL implementation for the viewer of SNeRG (or Baking NeRF)☆15Mar 21, 2023Updated 3 years ago
- [ICCVw TRICKY 2023] Ref-DVGO: Reflection-Aware Direct Voxel Grid Optimization for an Improved Quality-Efficiency Trade-Off in Reflective …☆16Apr 16, 2026Updated 3 months ago
- ☆23Feb 6, 2024Updated 2 years ago
- SRD: A Tree Structure Based Decoder for Online Handwritten Mathematical Expression Recognition☆21Jul 20, 2020Updated 6 years ago