AdaRFT: Efficient Reinforcement Finetuning via Adaptive Curriculum Learning
☆56Jun 13, 2025Updated last year
Alternatives and similar repositories for verl
Users that are interested in verl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Implementation for the paper "Discovering Knowledge Deficiencies of Language Models on Massive Knowledge Base"☆27Sep 2, 2025Updated 10 months ago
- ☆16Jul 10, 2025Updated last year
- ☆18Feb 2, 2026Updated 5 months ago
- [COLM 2025] "C3PO: Critical-Layer, Core-Expert, Collaborative Pathway Optimization for Test-Time Expert Re-Mixing"☆21Apr 9, 2025Updated last year
- Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination.