Official repository for Beyond Binary Rewards: Training LMs to Reason about Their Uncertainty
☆68Aug 20, 2025Updated 11 months ago
Alternatives and similar repositories for RLCR
Users that are interested in RLCR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the repo for constructing a comprehensive and rigorous evaluation framework for LLM calibration.☆14Apr 9, 2024Updated 2 years ago
- ☆28Jul 18, 2025Updated last year
- Materials for the course Principles of AI: LLMs at UPenn (Stat 9911, Spring 2025). LLM architectures, training paradigms (pre- and post-t…☆48Jun 14, 2025Updated last year
- ☆11Apr 23, 2023Updated 3 years ago
- Optimizing Anytime Reasoning via Budget Relative Policy Optimization☆54Jul 15, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2025] What Makes a Reward Model a Good Teacher? An Optimization Perspective☆44Sep 18, 2025Updated 10 months ago
- Scaling Test-time Training for LLM Reasoning☆27Apr 14, 2026Updated 3 months ago
- Code for "Multi-Modal Neural Machine Translation with Deep Semantic Interactions" (Information Sciences)☆16May 21, 2021Updated 5 years ago
- A holistic benchmark for LLM abstention