This repository serves as a collection of research notes and resources on training large language models (LLMs) and Reinforcement Learning from Human Feedback (RLHF). It focuses on the latest research, methodologies, and techniques for fine-tuning language models.
☆142Jul 28, 2025Updated last year
Alternatives and similar repositories for reasoning_models_how_to
Users that are interested in reasoning_models_how_to are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Transform your PDFs into captivating audio podcasts with this PDF-to-Podcast pipeline! Combining advanced language models and high-qualit…☆17Nov 11, 2024Updated last year
- ☆13May 12, 2023Updated 3 years ago
- BanglaWriting: A multi-purpose offline Bangla handwriting dataset☆14Nov 18, 2020Updated 5 years ago
- harvard-cs-2881-classroom-hw0-c2881-hw0 created by GitHub Classroom