Train transformer language models with reinforcement learning.
☆19Feb 25, 2025Updated last year
Alternatives and similar repositories for trl
Users that are interested in trl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Evolutionary Search for expert-level performance on any task with environmental feedback☆14Oct 12, 2025Updated 9 months ago
- Oak National Academy's AI Auto Eval tools provide LLM as a judge evaluation on lesson plans and resources☆17Jun 11, 2026Updated last month
- various experiments for scaling inference time compute with small reasoning models☆17Jan 16, 2025Updated last year
- Open-source Human Feedback Library☆11Oct 25, 2023Updated 2 years ago
- Lego for GRPO☆30May 27, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 🔥 Animotion examples☆20Apr 3, 2025Updated last year
- English and Chinese LaTeX template for reports/projects/proposal at Beijing Institute of Technology☆10Nov 19, 2020Updated 5 years ago
- ☆19Dec 12, 2023Updated 2 years ago
- ☆12Apr 19, 2024Updated 2 years ago
- Based on the R1-Zero method, using rule-based rewards and GRPO on the Code Contests dataset.☆18Apr 22, 2025Updated last year
- An implementation of the Augmented Random Search algorithm☆14Jan 29, 2022Updated 4 years ago
- Clustered Compositional Embeddings☆13Oct 25, 2023Updated 2 years ago
- A peer-to-peer communication system. BIT 小学期软件开发实训。☆11Sep 7, 2018Updated 7 years ago
- An unofficial API for royalroad.com☆23Jul 27, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆33Feb 6, 2025Updated last year
- ScribePal is an Open Source intelligent browser extension that leverages AI to empower your web experience by providing contextual insigh…☆22Apr 6, 2026Updated 3 months ago
- An RL environment similar to Cognition's SWE-Grep☆16Mar 10, 2026Updated 4 months ago
- ☆116Jan 21, 2025Updated last year
- Meta-Reinforcement Learning with Self-Reflection☆33Mar 26, 2026Updated 3 months ago
- ☆13Apr 16, 2025Updated last year
- Exploring the classical regression capabilities of LLMs.☆18May 20, 2024Updated 2 years ago
- Exploring the Limitations of Large Language Models on Multi-Hop Queries☆33Mar 2, 2025Updated last year
- Fork of Flame repo for training of some new stuff in development