Klear-Reasoner: Advancing Reasoning Capability via Gradient-Preserving Clipping Policy Optimization
☆82Dec 25, 2025Updated 8 months ago
Alternatives and similar repositories for KlearReasoner
Users that are interested in KlearReasoner are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CE-GPPO: Controlling Entropy via Gradient-Preserving Clipping Policy Optimization in Reinforcement Learning☆16Jan 23, 2026Updated 7 months ago
- ☆19Sep 7, 2025Updated last year
- RL with Experience Replay☆59Jul 27, 2025Updated last year
- Official implementation of the paper "Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following"☆40Jan 11, 2026Updated 8 months ago
- The official repository of paper "Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models''☆112Aug 15, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Large language models designed for formal theorem proving through tool-integrated reasoning.☆33Aug 13, 2025Updated last year
- An Ultra-Long Output Reinforcement Learning Approach☆23Jul 31, 2025Updated last year
- Revisiting Mid-training in the Era of Reinforcement Learning Scaling☆188Jul 23, 2025Updated last year
- ☆17Jun 10, 2025Updated last year
- ☆65Mar 30, 2026Updated 5 months ago
- Extrapolating RLVR to General Domains without Verifiers☆206Aug 12, 2025Updated last year
- CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning