用RLHF可选LoRA对LLaMA和MOSS进行训练|Training LLaMA or MOSS with RLHF [LoRA]
☆21May 16, 2023Updated 3 years ago
Alternatives and similar repositories for LLaMA-MOSS-RLHF-LoRA
Users that are interested in LLaMA-MOSS-RLHF-LoRA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Mixture of Expert (MoE) techniques for enhancing LLM performance through expert-driven prompt mapping and adapter combinations.☆11Feb 11, 2024Updated 2 years ago
- ☆18Apr 3, 2023Updated 3 years ago
- This repository provides the code for applying Contrastive Learning Penalty Loss (CLPL) and Mixture of Experts (MoE) to the BGE-M3 text e…☆11Dec 27, 2024Updated last year
- The model for edge classification by transforming edges to nodes.☆15Dec 22, 2020Updated 5 years ago
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ChatGLM-Peft-Tuning☆13Mar 19, 2023Updated 3 years ago
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- ☆16Nov 10, 2023Updated 2 years ago
- 对ChatGLM直接使用RLHF提升或降低目标输出概率|Modify ChatGLM output with only RLHF☆196May 23, 2023Updated 3 years ago
- 论文一体化写作神器(Python)☆17Apr 11, 2020Updated 6 years ago
- Demo for the subjective interface☆14Mar 4, 2018Updated 8 years ago
- This is a deep-learning based model for Electronic Design Automation(EDA), predicting the Design Rule Check (DRC) violation location.☆13Jun 24, 2023Updated 3 years ago
- My notes on reinforcement learning papers