deepspeed+trainer简单高效实现多卡微调大模型
☆132May 27, 2023Updated 3 years ago
Alternatives and similar repositories for ChatGLM_mutli_gpu_tuning
Users that are interested in ChatGLM_mutli_gpu_tuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code to reproduce THUIR‘s submissions for COLIEE 2023 Task1 and Task2☆28May 12, 2023Updated 3 years ago
- The official repo for our SIGIR'23 Full paper: Structure-aware Pre-trained Language Model for Legal Case Retrieval☆98May 9, 2023Updated 3 years ago
- The official repo for our SIGIR'23 Full paper: Constructing Tree-based Index for Efficient and Effective Dense Retrieval☆28Jun 7, 2023Updated 3 years ago
- LLM with LuXun (鲁迅) style☆91May 15, 2023Updated 3 years ago
- Large Language Models as Evaluators for Recommendation Explanations (RecSys 2024 Reproducibility)☆21Aug 13, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for KERM: Incorporating Explicit Knowledge in Pre-trained Language Models for Passage Re-ranking, accepted at SIGIR 2022.☆19Oct 31, 2022Updated 3 years ago
- LexiLaw - 中文法律大模型☆1,042Mar 12, 2026Updated 5 months ago
- chatglm多gpu用deepspeed和☆409Jul 8, 2024Updated 2 years ago
- The official implementation for Collaborative Word-based Pre-trained Item Representation for Transferable Recommendation.☆25Jan 30, 2024Updated 2 years ago
- chatglm3-6b, 微调/LORA/推理/单机多卡/deepspeed/支持多轮对话☆16Nov 30, 2023Updated 2 years ago
- An evaluation framework to test AI in a trial-and-error process. It is a simplified Natural Selection test.☆22Mar 11, 2025Updated last year
- T2Ranking: A large-scale Chinese benchmark for passage ranking.☆162Jul 3, 2023Updated 3 years ago
- A Large-Scale Chinese Legal Case Retrieval Dataset☆105Dec 29, 2024Updated last year
- BLOOM 模型的指令微调☆24Jun 15, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- LeXFiles and LegalLAMA: Facilitating English Multinational Legal Language Model Development☆23Jul 24, 2023Updated 3 years ago
- 一套代码指令微调大模型☆39Aug 1, 2023Updated 3 years ago
- Fine-tuning ChatGLM-6B with PEFT | 基于 PEFT 的高效 ChatGLM 微调☆3,713Oct 12, 2023Updated 2 years ago
- 中文对话数据清洗☆32Nov 8, 2022Updated 3 years ago
- ☆43Dec 15, 2023Updated 2 years ago
- ☆29Jul 25, 2025Updated last year
- This is our implementation of IntEL-Intent-aware Ranking Ensemble for Personalized Recommendation (SIGIR2023)☆24Nov 17, 2023Updated 2 years ago
- Code for the paper "A Comprehensive Evaluation of Large Language Models on Legal Judgment Prediction"☆13Oct 20, 2023Updated 2 years ago
- Code for MBGE-recognition: Emotion recognition based on multi-view body gestures, accepted at ICIP 2019.☆12Apr 6, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 基于ChatGLM-6B + LoRA的Fintune方案☆3,740Nov 25, 2023Updated 2 years ago
- 基于ChatGLM-6B、ChatGLM2-6B、ChatGLM3-6B模型,进行下游具体任务微调,涉及Freeze、Lora、P-tuning、全参微调等☆2,771Dec 12, 2023Updated 2 years ago
- Open ChatGLM Eyes to See the World☆13Mar 30, 2023Updated 3 years ago
- 中文法律LLaMA (LLaMA for Chinese legel domain)☆1,000Aug 28, 2024Updated 2 years ago
- 模型压缩的小白入门教程☆22Jul 7, 2024Updated 2 years ago
- An optimized deep prompt tuning strategy comparable to fine-tuning across scales and tasks☆2,078Nov 16, 2023Updated 2 years ago
- ☆22Apr 22, 2025Updated last year
- 基于 LoRA 和 P-Tuning v2 的 ChatGLM-6B 高效参数微调☆55May 17, 2023Updated 3 years ago
- 一个即插即用的llm增强器,项目初衷是通过工程化架构让小参数模型获得能媲美大模型的能力。☆19May 23, 2026Updated 3 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Repo. for RLCF.☆15Apr 1, 2024Updated 2 years ago
- ☆84Sep 9, 2023Updated 2 years ago
- Technical Report: Is ChatGPT a Good NLG Evaluator? A Preliminary Study☆43Mar 8, 2023Updated 3 years ago
- 基于ChatGLM3基座模型和LLAMA-Factory框架进行微调的一个中医问答机器人☆115Jan 3, 2024Updated 2 years ago
- 中文nlp解决方案(大模型、数据、模型、训练、推理)☆3,836Aug 5, 2025Updated last year
- "桃李“: 国际中文教育大模型☆195Nov 13, 2023Updated 2 years ago
- code for paper 《RankingGPT: Empowering Large Language Models in Text Ranking with Progressive Enhancement》☆34Jan 9, 2024Updated 2 years ago