Qwen3 Fine-tuning: Medical R1 Style Chat
☆334May 31, 2025Updated last year
Alternatives and similar repositories for Qwen3-Medical-SFT
Users that are interested in Qwen3-Medical-SFT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 大语言模型微调,Qwen2VL、Qwen2、GLM4指令微调☆658May 26, 2025Updated last year
- 一个包含了多种主流大模型微调方案的实战代码库,基于Qwen3系列模型☆141Aug 10, 2025Updated last year
- ChatGLM4微调简介☆27Apr 8, 2025Updated last year
- Qwen2.5 0.5B GRPO☆87Feb 16, 2025Updated last year
- 基于 Qwen/Qwen3-0.6B 的医疗问答微调与推理示例项目,记录健康数据、生成建议与报告☆16Sep 30, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A comparison of deepseek grpo and qwen gspo on Qwen2.5-1.5B-Instruct fine tunning.☆171Mar 28, 2026Updated 6 months ago
- 这是一个从头训练大语言模型的项目,包括预训练、微调和直接偏好优化,模型拥有1B参数,支持中英文。☆876Feb 18, 2025Updated last year
- 复现大模型相关算法及一些学习记录☆3,547Jul 2, 2026Updated 3 months ago
- 集成Qwen与DeepSeek等先进大语言模型,支持纯LLM+分类层模式及LLM+LoRA+分类层模式,使用transformers模块化设计和训练便于根据需要调整或替换组件。☆22Sep 1, 2025Updated last year
- 将SmolVLM2的视觉头与Qwen3-0.6B模型进行了拼接微调☆612Sep 8, 2025Updated last year
- MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training Pipeline. 训练医疗大模型,实现了包括增量预训练(PT)、有监督微调(SFT)、RLHF、DPO、ORPO、GRPO。☆5,857Sep 15, 2026Updated 3 weeks ago
- 《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程☆32,425Sep 12, 2026Updated 3 weeks ago
- ☆14Aug 19, 2025Updated last year
- Reliable Source Approximation: Source-Free Domain Adaptation for Vestibular Schwannoma MRI Segmentation☆11Dec 28, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 项目的issue会存放我的所有blog☆20Sep 12, 2025Updated last year
- ☆10Apr 30, 2025Updated last year
- A travel agent based on Qwen2.5, fine-tuned by SFT + DPO/PPO/GRPO using traveling question-answer dataset, a mindmap can be output using …☆83Jul 6, 2026Updated 3 months ago
- 基于deepseek、qwen3大模型,lora sft 医疗行业数据☆15Apr 10, 2026Updated 6 months ago
- Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.8, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL…☆15,807Updated this week
- ☆30May 17, 2026Updated 4 months ago
- Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)☆75,401Updated this week
- 👀 Train a 65M-parameter VLM from scratch in just 2h!☆8,734Sep 22, 2026Updated 2 weeks ago
- 基于大模型ChatGLM,微调方式为LORA,集SFT、RM、PPO算法为一体项目☆14Jun 20, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Accelerating GOT-OCRv2 with VLLM☆10Nov 15, 2024Updated last year
- A federated image segmentation method based on style transfer☆16Sep 28, 2024Updated 2 years ago
- 中文领域多轮心理健康思维链对话数据集SoulChat-R1☆21Nov 12, 2025Updated 10 months ago
- 使用多轮对话数据集对deepseek进行lora微调教程☆61Dec 26, 2024Updated last year
- 《大模型白盒子构建指南》:一个全手搓的Tiny-Universe☆5,098Feb 12, 2026Updated 7 months ago
- 本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)☆25,150Jul 19, 2026Updated 2 months ago
- 🧠 Train a 64M-parameter LLM from scratch in just 2h!☆63,508Sep 22, 2026Updated 2 weeks ago
- 天池Better Synth多模态大模型数据合成挑战赛-打赢baseline就算成功方案☆30Oct 30, 2025Updated 11 months ago
- ⚡️SwanLab - an open-source, modern-design AI training tracking and visualization tool. Supports Cloud / Self-hosted use. Integrated with …☆4,272Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 基于自由度(熵)、凝固度 新词发现算法实现☆12Oct 7, 2018Updated 8 years ago
- 2024百度商业AI技术创新大赛赛道一:基于大模型的广告检索全国一等奖获奖方案☆19Feb 23, 2025Updated last year
- Qwen1.5-SFT(阿里, Ali), Qwen_Qwen1.5-2B-Chat/Qwen_Qwen1.5-7B-Chat微调(transformers)/LORA(peft)/推理☆72May 17, 2024Updated 2 years ago
- 基于DeepSpeed的大模型微调教程,详细介绍如何使用DeepSpeed进行微调和分布式训练文本总结大模型☆19May 6, 2026Updated 5 months ago
- ☆34Jul 8, 2025Updated last year
- ☆39Sep 17, 2025Updated last year
- 基于Qwen2+SFT+DPO的医疗问答系统,项目中使用了自定义的 SFTTrainer/DPOTrainer/TRPOTrainer用于训练,其次,项目还调用各种知识库工具(neo4j, milvus, LDA, 等)进行自动化训练数据生成。另外,使用 vllm 用于推理…☆90Apr 29, 2026Updated 5 months ago