☆21Dec 14, 2024Updated last year
Alternatives and similar repositories for Reward-Calibration
Users that are interested in Reward-Calibration are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12Apr 18, 2025Updated last year
- codes for Efficient Test-Time Scaling via Self-Calibration☆22Sep 13, 2025Updated last year
- ☆27May 14, 2026Updated 4 months ago
- ☆41Aug 25, 2026Updated 3 weeks ago
- AbstainQA, ACL 2024☆30Feb 4, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2024] This is the official implementation for the paper: "Beyond imitation: Leveraging fine-grained quality signals for alignment"☆11May 5, 2024Updated 2 years ago
- Extensive Self-Contrast Enables Feedback-Free Language Model Alignment☆20Apr 2, 2024Updated 2 years ago
- ☆21Apr 16, 2025Updated last year
- ☆43Feb 2, 2024Updated 2 years ago
- [ACL 2024] Benchmarking Knowledge Boundary for Large Language Models: A Different Perspective on Model Evaluation☆10May 26, 2024Updated 2 years ago
- Text-to-dysarthric speech (TTDS) synthesis. An implementation using the Grad-TTS model with the TORGO database.☆15Mar 15, 2025Updated last year
- [ICLR 2025] On Evluating the Durability of Safegurads for Open-Weight LLMs☆13Jun 20, 2025Updated last year
- [ICLR 2025] FLAT: LLM Unlearning via Loss Adjustment with Only Forget Data☆14Feb 26, 2025Updated last year
- This is the oficial repository for "Safer-Instruct: Aligning Language Models with Automated Preference Data"☆17Feb 22, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AdaptiveStep: Automatically Dividing Reasoning Step through Model Confidence☆11Mar 2, 2025Updated last year
- [ACL'24 Findings] Official code for "TLCR: Token-Level Continuous Reward for Fine-grained Reinforcement Learning from Human Feedback"☆12Dec 6, 2024Updated last year
- ☆15May 28, 2024Updated 2 years ago
- Code and Data for "Long-context LLMs Struggle with Long In-context Learning" [TMLR2025]☆113Feb 20, 2025Updated last year
- Sparse Backpropagation for Mixture-of-Expert Training☆30Jul 2, 2024Updated 2 years ago
- Code for paper Empowering Large Language Model Agents through Action Learning☆35Aug 8, 2024Updated 2 years ago
- [COLM'25] CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing☆19Jun 25, 2025Updated last year
- Repo for paper: Examining LLMs' Uncertainty Expression Towards Questions Outside Parametric Knowledge☆14Feb 20, 2024Updated 2 years ago
- Code for SafeMERGE (ICLR 2025).☆15Apr 1, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The offical repo for "Parallel-Probe: Towards Efficient Parallel Thinking via 2D Probing"☆21Feb 3, 2026Updated 7 months ago
- Code for the paper "Studying Large Language Model Behaviors Under Context-Memory Conflicts With Real Documentss"☆14Oct 8, 2024Updated last year
- Code for experiments on transformers using Markovian data.☆22Nov 22, 2024Updated last year
- personalized-llms with allen institute☆13Jun 22, 2023Updated 3 years ago
- [ACL' 25] The official code repository for PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models.☆94Feb 15, 2025Updated last year
- Working note for WSI analysis☆10Apr 3, 2023Updated 3 years ago
- 把成语转成 emoji 来猜谜的小玩具☆26Jan 2, 2024Updated 2 years ago
- An NLP research and data collection platform.☆17Jul 4, 2026Updated 2 months ago
- ☆26Feb 11, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Repo. for RLCF.☆15Apr 1, 2024Updated 2 years ago
- Entity-Based Knowledge Conflicts in Question Answering. Code repo for EMNLP2021 paper: https://aclanthology.org/2021.emnlp-main.565/☆77Sep 11, 2026Updated last week
- TallyQA: Answering Complex Counting Questions dataset☆31Feb 19, 2024Updated 2 years ago
- Code for paper "Factual Confidence of LLMs: on Reliability and Robustness of Current Estimators"☆17Dec 4, 2024Updated last year
- Analyzing LLM Alignment via Token distribution shift☆17Jan 26, 2024Updated 2 years ago
- Causal tracing for language models☆12Apr 2, 2024Updated 2 years ago
- MultiMath: Bridging Visual and Mathematical Reasoning for Large Language Models☆33Jan 22, 2025Updated last year