[NIPS2023] RRHF & Wombat
☆804Sep 22, 2023Updated 2 years ago
Alternatives and similar repositories for RRHF
Users that are interested in RRHF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)☆4,754Jan 8, 2024Updated 2 years ago
- Instruction Tuning with GPT-4☆4,332Jun 11, 2023Updated 3 years ago
- Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback☆1,613Nov 24, 2025Updated 9 months ago
- Open Academic Research on Improving LLaMA to SOTA LLM☆1,602Aug 30, 2023Updated 3 years ago
- Secrets of RLHF in Large Language Models Part I: PPO☆1,430Mar 3, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)☆8,273Oct 16, 2024Updated last year
- We unified the interfaces of instruction-tuning data (e.g., CoT data), multiple LLMs and parameter-efficient methods (e.g., lora, p-tunin…☆2,788Dec 12, 2023Updated 2 years ago
- Multi-agent Social Simulation + Efficient, Effective, and Stable alternative of RLHF. Code for the paper "Training Socially Aligned Langu…☆357Jun 18, 2023Updated 3 years ago
- Human preference data for "Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback"☆1,857Jun 17, 2025Updated last year
- Reference implementation for DPO (Direct Preference Optimization)☆2,907Aug 11, 2024Updated 2 years ago
- 800,000 step-level correctness labels on LLM solutions to MATH problems☆2,148Jun 1, 2023Updated 3 years ago
- A simulation framework for RLHF and alternatives. Develop your RLHF method without collecting human data.