☆41Nov 16, 2025Updated 9 months ago
Alternatives and similar repositories for vlm_reproduce
Users that are interested in vlm_reproduce are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pretrain、Posttrain、RAG、Agent等大模型相关的基础项目合集☆39Dec 7, 2025Updated 9 months ago
- A comparison of deepseek grpo and qwen gspo on Qwen2.5-1.5B-Instruct fine tunning.☆170Mar 28, 2026Updated 5 months ago
- 基于轻量级 Qwen2.5-0.5B 和 SigLIP 的视觉语言多模态模型实现,包含训练和 SFT 代码。分享训练和 SFT 相关代码,记录一下探索和学习的过程。欢迎一起交流讨论~☆22Aug 31, 2025Updated last year
- code for paper “PET Image Denoising with Score-Based Diffusion Probabilistic Models”☆11Dec 23, 2023Updated 2 years ago
- 新手友好的基于 Qwen2 的生成式推荐系统,通过大模型理解用户偏好生成候选物品,融合 TF-IDF 关键词召回、热门物品召回构建多源策略,平衡相关性与多样性。采用 LightGBM 排序模型精准打分,内置召回率、NDCG 等评估指标量化效果。通过 Flask 封装 API…☆55Dec 6, 2025Updated 9 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- HydraRNA is a full-length RNA language model.☆18Dec 15, 2025Updated 8 months ago
- Personal Project: MPP-Qwen14B & MPP-Qwen-Next(Multimodal Pipeline Parallel based on Qwen-LM). Support [video/image/multi-image] {sft/conv…☆686Mar 10, 2025Updated last year
- Denoising Diffusion Probabilistic Model for retinal image generation☆36Nov 29, 2024Updated last year
- Local-first interview recording review reports with a Codex skill and CLI.☆81May 16, 2026Updated 3 months ago
- [ICML 2022 Spotlight] Finding the Task-Optimal Low-Bit Sub-Distribution in Deep Neural Networks☆11May 21, 2023Updated 3 years ago
- ☆29Feb 27, 2026Updated 6 months ago
- Full-dose Whole-body PET Synthesis from Low-dose PET using Consistency Model☆14Apr 17, 2024Updated 2 years ago
- (Pattern Recognition 2025) Towards Trustworthy Dataset Distillation☆14Dec 8, 2024Updated last year
- 超简单使用监督微调SFT和强化学习RL去训练领域Agent☆35Oct 20, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆16Jun 18, 2026Updated 2 months ago
- This repository is used for learning how to perform 3d reconstruction of the heart from various imaging modalities (e.g. CT, MRI, US), bu…☆13Jun 8, 2023Updated 3 years ago
- ☆10Jan 10, 2022Updated 4 years ago
- ☆13Dec 17, 2024Updated last year
- DocMind - 企业级 RAG 智能问答系统☆18Dec 29, 2025Updated 8 months ago
- Open Source Road Datasets☆19Aug 30, 2024Updated 2 years ago
- ☆12Nov 17, 2023Updated 2 years ago
- 基于Qwen2+SFT+DPO的医疗问答系统,项目中使用了自定义的 SFTTrainer/DPOTrainer/TRPOTrainer用于训练,其次,项目还调用各种知识库工具(neo4j, milvus, LDA, 等)进行自动化训练数据生成。另外,使用 vllm 用于推理…☆88Apr 29, 2026Updated 4 months ago
- Benchmarking pipeline for single-cell perturbation prediction tools☆36Jul 11, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Pytorch implementation of our paper accepted by ECCV 2022-- Fine-grained Data Distribution Alignment for Post-Training Quantization☆16Sep 13, 2022Updated 3 years ago
- Implementation of Differentiable ODE Solvers using Jittor☆13May 18, 2025Updated last year
- ☆16Nov 25, 2022Updated 3 years ago
- Self Evolving Large Multimodal Models with Continuous Rewards☆27Updated this week
- [Nature Communications 2023] "Wearable in-sensor reservoir computing using optoelectronic polymers with through-space charge-transport ch…☆17Jan 11, 2023Updated 3 years ago
- Steganography with help of LLM☆13Apr 28, 2025Updated last year
- Codex skill for reconstructing diagram images into editable Draw.io files☆28Aug 28, 2026Updated last week
- ☆22Oct 6, 2023Updated 2 years ago
- 🎓Automatically Update CV Papers Daily using Github Actions (Update Every 12th hours)☆12May 17, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Multi-Modal-AI-Orchestrator (Reset version),AI Full-modal Full-agent:Text → Image → Music → Lights → Video, Includes "Scenario Director,…☆103Nov 5, 2025Updated 10 months ago
- 🎯 Build a winning recommendation system with this effective generative framework, advancing to the finals of the 2025 Tencent Advertisin…☆27Updated this week
- Official implementation of Bifrost-1: Bridging Multimodal LLMs and Diffusion Models with Patch-level CLIP Latents (NeurIPS 2025)☆47Nov 24, 2025Updated 9 months ago
- 毕业设计:《基于CLIP模型的视频文本检索设计与实现》☆17Jul 21, 2024Updated 2 years ago
- 使用Flutter开发的一款App,主要功能是日程管理,根据优先级进行划分,设置时间提醒☆11Oct 29, 2018Updated 7 years ago
- [CVPR2025] Hybrid-Level Instruction Injection for Video Token Compression in Multi-modal Large Language Models☆21Apr 30, 2025Updated last year
- Image Denoising Using Anisotropic Diffusion☆12Jun 15, 2016Updated 10 years ago