Code for "Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate" [COLM 2025]
☆182Jul 8, 2025Updated last year
Alternatives and similar repositories for CritiqueFineTuning
Users that are interested in CritiqueFineTuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official repo for “Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem” [EMNLP25]☆33Sep 1, 2025Updated 11 months ago
- Code for "[COLM'25] RepoST: Scalable Repository-Level Coding Environment Construction with Sandbox Testing"☆24Mar 18, 2025Updated last year
- [ICML 2025] Teaching Language Models to Critique via Reinforcement Learning☆127May 6, 2025Updated last year
- General Reasoner: Advancing LLM Reasoning Across All Domains [NeurIPS25]☆229Nov 27, 2025Updated 8 months ago
- Simple RL training for reasoning☆3,872Dec 23, 2025Updated 7 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The official repo for "AceCoder: Acing Coder RL via Automated Test-Case Synthesis" [ACL25]☆100Apr 9, 2025Updated last year
- Code for Blog Post: Can Better Cold-Start Strategies Improve RL Training for LLMs?☆20Mar 9, 2025Updated last year
- A unified benchmark for math reasoning☆90Jan 25, 2023Updated 3 years ago
- Exploration of automated dataset selection approaches at large scales.☆55Mar 4, 2025Updated last year
- Code for "Unifying Molecular and Textual Representations via Multi-task Language Modelling" @ ICML 2023