chatglm3base模型的有监督微调SFT
☆80Nov 5, 2023Updated 2 years ago
Alternatives and similar repositories for chatglm3-base-tuning
Users that are interested in chatglm3-base-tuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A High Performance Web Server and CDN Cache Server Based on Nginx and LuaJIT☆25Sep 29, 2026Updated last week
- ☆23Sep 9, 2023Updated 3 years ago
- 基于ChatGLM-6B、ChatGLM2-6B、ChatGLM3-6B模型,进行下游具体任务微调,涉及Freeze、Lora、P-tuning、全参微调等☆2,773Dec 12, 2023Updated 2 years ago
- 可以成功Lora微调的Qwen-VL模型☆16Oct 27, 2023Updated 2 years ago
- 基于chatglm3-6b模型的lora方法的微调☆81Apr 18, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 膨胀卷积,Multi Scale Context Aggregation by Dilated Convolutions的实现☆12Dec 24, 2017Updated 8 years ago
- 阿里通义千问(Qwen-7B-Chat/Qwen-7B), 微调/LORA/推理☆143May 17, 2024Updated 2 years ago
- InstantID with multiple controlnets and SDXL-Lightning☆20May 5, 2024Updated 2 years ago
- 用于微调LLM的中文指令数据集☆28Apr 12, 2023Updated 3 years ago
- 基于 LoRA 和 P-Tuning v2 的 ChatGLM-6B 高效参数微调☆55May 17, 2023Updated 3 years ago
- ☆22May 7, 2025Updated last year
- ChatGLM3 series: Open Bilingual Chat LLMs | 开源双语对话语言模型☆13,641Jan 13, 2025Updated last year
- A freeswitch esl server for make a callcenter core,ex:ACD,IVR and so on......☆12Sep 26, 2016Updated 10 years ago
- 基于Qwen2.5模型、使用DISC-Law-SFT-Pair数据集微调的法律大模型☆11Dec 29, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 基于ChatGLM2-6B进行微调,包括全参数、参数有效性、量化感知训练等,可实现指令微调、多轮对话微调等。☆24Jul 29, 2023Updated 3 years ago
- 本项目对Deepseek-R1-Distill-Qwen-7B进行心理咨询CoT数据的LoRA微调,以进一步提升Deepseek-R1-Distill-Qwen-7B在心理咨询领域的慢思考能力。☆12Mar 11, 2025Updated last year
- ☆11Oct 24, 2024Updated last year
- 大语言模型微调的项目,包含了使用QLora微调ChatGLM和LLama☆29Jun 26, 2023Updated 3 years ago
- 学习ChatGLM3模型和LangChain框架的架构与核心功能,并基于LangChain+ChatGLM3实现本地知识库问答。☆36Jan 30, 2024Updated 2 years ago
- 谷歌插件:标签整理器☆11Apr 2, 2024Updated 2 years ago
- 基于prompt learning的NLP算法,涉及文本分类、信息抽取等☆15Jan 12, 2023Updated 3 years ago
- 在ChatGLM大模型上利用LoRA方法进行小参数学习,训练语料库选择中文的[alpaca-zh](https://huggingface.co/datasets/shibing624/alpaca-zh)☆26Apr 13, 2023Updated 3 years ago
- KuaiSearch PERKS☆12Nov 16, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- Research project for task-oriented dialogue system with jointly training multi-intent classification and slot filling☆10Sep 11, 2023Updated 3 years ago
- 受到self-instruct启发,除了通用LLM还能做垂直领域的小LLM实现定制效果,通过GPT获得question和answer来作为训练数据☆18May 12, 2023Updated 3 years ago
- 离线部署大模型,构建一个可以上传本地知识库进行RAG问答且可以自行调用工具的Agent。☆44Apr 23, 2024Updated 2 years ago
- “达观杯”长文本智能处理挑战赛。达观数据提供了一批长文本数据和分类信息,希望选手动用自己的智慧,结合当下最先进的NLP和人工智能技术,深入分析文本内在结构和语义信息,构建文本分类模型,实现精准分类。☆10Jul 20, 2018Updated 8 years ago
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- 2021科大讯飞试题标签预测挑战赛亚军方案☆13Dec 4, 2021Updated 4 years ago
- Reinforcement learning based pricing for demand response☆19Jun 17, 2023Updated 3 years ago
- 基于ChatGLM-6B + LoRA的Fintune方案☆3,738Nov 25, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- SGF parsing library for Rust.☆19Jun 19, 2026Updated 3 months ago
- Fine-tuning ChatGLM-6B with PEFT | 基于 PEFT 的高效 ChatGLM 微调☆3,711Oct 12, 2023Updated 2 years ago
- 用于生成文本纠错模型(如Gector)需要的大量数据。☆15Jan 5, 2023Updated 3 years ago
- A repo for update and debug Mixtral-7x8B、MOE、ChatGLM3、LLaMa2、 BaChuan、Qwen an other LLM models include new models mixtral, mixtral 8x7b, …☆47Oct 8, 2025Updated last year
- Alpaca Chinese Dataset -- 中文指令微调数据集☆219Oct 6, 2024Updated 2 years ago
- A Android client of Stable Diffusion.☆13Mar 29, 2024Updated 2 years ago
- AUTOMATIC111/stable-difusion-webui的Golang API服务端☆13Jul 10, 2023Updated 3 years ago