LeslieTrue / SFTvsRL

Official implementation of paper: SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
255Updated last month

Alternatives and similar repositories for SFTvsRL:

Users that are interested in SFTvsRL are comparing it to the libraries listed below