一个低成本、易于上手的多模态大模型学习项目。基于Qwen3-0.6B和CLIP构建,使用LLaVA架构和LoRA微调,在消费级16G显卡上数小时即可完成训练
☆51Sep 15, 2025Updated 10 months ago
Alternatives and similar repositories for TinyLLaVA-Qwen3
Users that are interested in TinyLLaVA-Qwen3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 一个包含了多种主流大模型微调方案的实战代码库,基于Qwen3系列模型☆134Aug 10, 2025Updated 11 months ago
- Qwen3 Fine-tuning: Medical R1 Style Chat☆332May 31, 2025Updated last year
- hrtem lattice fringe image analysis algorithms - developed for carbon nanostructure analysis☆10Jan 9, 2026Updated 6 months ago
- [ICASSP 2026] The official pytorch implementation of ACVIS☆15Jan 19, 2026Updated 6 months ago
- A real-time inferencing of multistreaming YOWOv3(Spatio Temporal Action Detection task) using (UCF101-24) dataset. The repo is extension …☆26May 15, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 推荐系统-腾讯广告算法大赛☆15Sep 12, 2025Updated 10 months ago
- 基于qwen3的医疗大模型研发全流程 0.分词训练 1.增量预训练 2.微调 3.强化 4.量化 5.蒸馏 6.评估 7.lora模型合并 8.服务 9.部署☆47Jan 3, 2026Updated 6 months ago
- 用Paddle复现论文ChineseBERT: Chinese Pretraining Enhanced by Glyph and Pinyin Information(ACL2021)☆10Nov 15, 2021Updated 4 years ago
- [CVPR'26, Findings] AuralSAM2: Enabling SAM2 Hear Through Pyramid Audio-Visual Feature Prompting☆15May 18, 2026Updated 2 months ago
- ☆14Oct 30, 2024Updated last year
- ☆14Nov 4, 2021Updated 4 years ago
- Local DeepSearch (Advantage: Low Threshold): an implementation of Agentic RAG based on DeepSeek-R1 API and Tavily API☆17Jun 21, 2025Updated last year
- [CVPR 2023] Learning Steerable Function for Efficient Image Resampling☆16Jun 1, 2023Updated 3 years ago
- 基于MCP协议和LangChain框架实现的企业级AI多Agent多模态系统,包含RAG技术增强的知识检索能力。☆43Mar 10, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of PFLD(Paper: "A Practical Facial Landmark Detector") by pytorch.☆15Feb 16, 2021Updated 5 years ago
- A travel agent based on Qwen2.5, fine-tuned by SFT + DPO/PPO/GRPO using traveling question-answer dataset, a mindmap can be output using …☆81Jul 6, 2026Updated 3 weeks ago
- Advancing Medical Foundation Models with Unified Medical Image Grounding for Clinical Reasoning☆17Sep 25, 2025Updated 10 months ago
- ☆15Aug 31, 2025Updated 10 months ago
- STFNet: Self-supervised Transformer for Infrared and Visible Image Fusion☆14Mar 25, 2024Updated 2 years ago
- SmartCLIP: A training method to improve CLIP with both short and long texts☆43Jun 18, 2025Updated last year
- ☆18Dec 12, 2023Updated 2 years ago
- [2026 AAAI] Think Before You Segment: An Object-aware Reasoning Agent for Referring Audio-Visual Segmentation☆20Nov 8, 2025Updated 8 months ago
- Python基于OpenCV的人脸表情识别系统[源码&部署教程]☆19Nov 17, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 本项目是一个基于LangChain构建的多Agent系统,结合Streamlit实现的Web界面,能够根据用户输入进行网络搜索并提供旅游相关的聊天服务。此外,该系统还具备基于本地知识库的推销功能,为用户提供个性化的旅游产品推荐。☆15Apr 20, 2025Updated last year
- This is the official implementation of YOLA, NeurIPS2024☆42Mar 8, 2025Updated last year
- "DenseFusion: 6D Object Pose Estimation by Iterative Dense Fusion" code repository☆12Aug 19, 2020Updated 5 years ago
- 自己调的网络☆11Sep 5, 2019Updated 6 years ago
- 个人工作站项目☆15Mar 17, 2026Updated 4 months ago
- [NeurIPS 25] VR-Drive: Viewpoint-Robust End-to-End Driving with Feed-Forward 3D Gaussian Splatting☆28Jan 4, 2026Updated 6 months ago
- A Codex skill for jointly analyzing research papers and code, with a static web reader for follow-up questions.☆21May 11, 2026Updated 2 months ago
- (2021' TIM) This is the official implementation for the paper titled "UNIFusion: A Lightweight Unified Image Fusion Network".☆11Apr 12, 2023Updated 3 years ago
- Prompt-first starter kit for Codex and Claude Code workspaces, with bootstrap, upkeep, and archive workflows.☆20Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- CurricuVLM: Towards Safe Autonomous Driving via Personalized Safety-Critical Curriculum Learning with Vision-Language Models☆25May 1, 2026Updated 2 months ago
- Havard Medical Image Fusion Datasets CT-MRI PET-MRI SPECT-MRI☆10Oct 27, 2025Updated 9 months ago
- Multi-Modal-AI-Orchestrator (Reset version),AI Full-modal Full-agent:Text → Image → Music → Lights → Video, Includes "Scenario Director,…☆103Nov 5, 2025Updated 8 months ago
- [ICRA 2024] WLST: Weak Labels Guided Self-training for Weakly-supervised Domain Adaptation on 3D Object Detection☆12Feb 6, 2024Updated 2 years ago
- [AAAI2024] An official pytorch implement of the paper: Vision-Language Pre-training with Object Contrastive Learning for 3D Scene Underst…☆13Dec 8, 2024Updated last year
- [ICCV 2025] Revisiting Efficient Semantic Segmentation: Learning Offsets for Better Spatial and Class Feature Alignment☆59Oct 14, 2025Updated 9 months ago
- ☆15Dec 6, 2024Updated last year