对llama3进行全参微调、lora微调以及qlora微调。
☆222Oct 4, 2024Updated last year
Alternatives and similar repositories for Llama3.1-Finetuning
Users that are interested in Llama3.1-Finetuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LlaMA3-SFT, Meta-Llama-3-8B/Meta-Llama-3-8B-Instruct微调(transformers)/LORA(peft)/推理, 支持中文(chinese, zh)☆34May 17, 2024Updated 2 years ago
- Qwen1.5-SFT(阿里, Ali), Qwen_Qwen1.5-2B-Chat/Qwen_Qwen1.5-7B-Chat微调(transformers)/LORA(peft)/推理☆73May 17, 2024Updated 2 years ago
- Qwen1.5大模型微调、基于PEFT框架LoRA微调,在数据集HC3-Chinese上实现文本分类。☆12Jun 29, 2024Updated 2 years ago
- 大语言模型微调,Qwen2VL、Qwen2、GLM4指令微调☆656May 26, 2025Updated last year
- 基于Unsloth框架下,使用llama3大模型为基底的模型微调☆18Mar 3, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Gemma-SFT, gemma-2b/gemma-7b微调(finetune,transformers)/LORA(peft)/推理(inference)☆32May 17, 2024Updated 2 years ago
- ☆15Apr 27, 2024Updated 2 years ago
- 基于ChatGLM-6B、ChatGLM2-6B、ChatGLM3-6B模型,进行下游具体任务微调,涉及Freeze、Lora、P-tuning、全参微调等☆2,770Dec 12, 2023Updated 2 years ago
- ☆37Feb 16, 2025Updated last year
- Official code for the paper Improving Language Plasticity via Pretraining with Active Forgetting, NeurIPS 2023☆21Mar 12, 2026Updated 5 months ago
- ☆21Dec 7, 2025Updated 8 months ago
- 基于大模型ChatGLM,微调方式为LORA,集SFT、RM、PPO算法为一体项目☆14Jun 20, 2023Updated 3 years ago
- 以InternLM2-chat-7为基座模型,以常用中药等为数据集,微调的大模型。中医聊天小助手。☆18Feb 29, 2024Updated 2 years ago
- Alpaca Chinese Dataset -- 中文指令微调数据集☆219Oct 6, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 行业调研助手是一个基于AI的自动化行业研究工具,旨在帮助研究人员、分析师和决策者快速获取、分析和整合行业信息,生成结构化的研究报告。本项目利用先进的自然语言处理技术,实现了从信息搜集、内容提取到报告生成的全流程自动化。☆25Mar 26, 2025Updated last year
- Code of Paper: Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Models☆16Mar 17, 2026Updated 5 months ago
- [SIGIR '25] This is the code repo for our SIGIR '25 paper: Enhancing the Patent Matching Capability of Large Language Models via Memory G…☆19Apr 22, 2025Updated last year
- 中文文本摘要生成模型☆21Jul 29, 2022Updated 4 years ago
- 陶弘景中医药大模型,包括命名实体识别,关系抽取,知识图谱构建,大模型增量微调,RAG☆19Jul 28, 2025Updated last year
- [ACL'25] Code for "Aligning Large Language Models to Follow Instructions and Hallucinate Less via Effective Data Filtering"☆21Jul 23, 2025Updated last year
- 简单易懂的LLaMA微调指南。☆411Jul 5, 2023Updated 3 years ago
- The official implementation of our work Hawkeye: Discovering and Grounding Implicit Anomalous Sentiment in Recon-videos via Scene-enhanc…☆13Oct 14, 2024Updated last year
- 大模型微调工具集合☆26Mar 15, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 基于qlora对baichuan-7B大模型进行指令微调。☆22Jun 22, 2023Updated 3 years ago
- Multi-Task instruction-tuned LLaMA☆14May 5, 2023Updated 3 years ago
- 通义千问的DPO训练☆67Sep 21, 2024Updated last year
- Retrieval Augmented Generation (RAG) implementation through libraries like Tavily, LangChain, ChatGLM3☆21Jul 23, 2025Updated last year
- [EMNLP2024] Aligning Large Language Models on Information Extraction☆57Nov 4, 2024Updated last year
- 使用UniLM实现中文文本摘要☆43Mar 25, 2020Updated 6 years ago
- 结合知识图谱做的有关诗词的问答demo☆11Mar 11, 2020Updated 6 years ago
- [EMNLP 2024 Findings] ICON: Improving Inter-Report Consistency in Radiology Report Generation via Lesion-aware Mixup Augmentation☆19Dec 11, 2024Updated last year
- To convert CosyVoice model to ONNX☆16Dec 22, 2025Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The official code and dataset for EMNLP 2022 paper "COPEN: Probing Conceptual Knowledge in Pre-trained Language Models".☆21Mar 9, 2023Updated 3 years ago
- Llama3-Tutorial(XTuner、LMDeploy、OpenCompass)☆506May 10, 2024Updated 2 years ago
- 在中文开源大模型的基础上进行定制化的微调,拥有自己专属的语言模型。☆53May 20, 2023Updated 3 years ago
- Merging Generated and Retrieved Knowledge for Open-Domain QA (EMNLP 2023)☆21Oct 8, 2023Updated 2 years ago
- 基于torch的,支持logits、crf、span、global_pointer四种方法的命名实体识别任务框架☆24Feb 29, 2024Updated 2 years ago
- WiSR: Wireless Domain Generalization Based on Style Randomization☆13Nov 17, 2023Updated 2 years ago
- Community-aware Graph Transformer (CGT) is a novel Graph Transformer model that utilizes community structures to address node degree bias…☆15Aug 27, 2025Updated 11 months ago