SCoRe: Training Language Models to Self-Correct via Reinforcement Learning
☆16May 14, 2026Updated 3 months ago
Alternatives and similar repositories for SCoRe
Users that are interested in SCoRe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Concise Reasoning via Reinforcement Learning☆13Apr 16, 2025Updated last year
- Implementation of Stein Variational Gradient Descent with TensorFlow 2.0☆12Sep 11, 2019Updated 6 years ago
- Official repository for AutoRule: Reasoning Chain-of-thought Extracted Rule-based Rewards Improve Preference Learning☆17Jul 24, 2025Updated last year
- ☆19Mar 25, 2025Updated last year
- ACL 2026☆27Nov 19, 2025Updated 9 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official source code of ICDM2023 paper "Hypergraph Contrastive Learning for Drug Trafficking Community Detection".☆11Nov 3, 2023Updated 2 years ago
- Fully open reproduction of DeepSeek-R1☆11Mar 24, 2025Updated last year
- ☆14Apr 18, 2020Updated 6 years ago
- Binding Affinity Prediction using Deep learning models☆12Jun 9, 2021Updated 5 years ago
- Unofficial Implementation of Selective Attention Transformer☆20Oct 31, 2024Updated last year
- ☆13Jul 14, 2024Updated 2 years ago
- ☆11Jun 16, 2024Updated 2 years ago
- ☆16Apr 16, 2024Updated 2 years ago
- Code for the "Long Context Needs Some R&R" paper.☆12Mar 11, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Can LLMs Convert Graphs to Text-Attributed Graphs? NAACL 25☆16Mar 7, 2025Updated last year
- ☆14Oct 12, 2024Updated last year
- ☆17Apr 11, 2025Updated last year
- Implementation of AdaCQR(COLING 2025)☆15Dec 30, 2024Updated last year
- Learning Protein-Ligand Properties with Atomic Environment Vectors☆10Apr 19, 2024Updated 2 years ago
- Reinforcement Leanring Algorithms Trained with Unity☆13Apr 26, 2019Updated 7 years ago
- lol助手秒选亚索☆12Jun 12, 2022Updated 4 years ago
- Chapter 15 AlphaZero in book Deep Reinforcement Learning: code example of AlphaZero solving Gomoku game.☆36Feb 18, 2020Updated 6 years ago
- ☆54Feb 12, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Water Network-Augmented Two-State model for Protein−Ligand Binding Affinity Prediction☆12Jun 10, 2023Updated 3 years ago
- ☆18Jun 30, 2025Updated last year
- This is a PyTorch implementation of a Transformer Decoder based model that plays chess.☆17Mar 15, 2024Updated 2 years ago
- Implementation of DiffDock: Diffusion Steps, Twists, and Turns for Molecular Docking☆13Jun 22, 2023Updated 3 years ago
- Official implementation of "Figure It Out: Improve the Frontier of Reasoning with Active Visual Thinking"☆17Jan 13, 2026Updated 7 months ago
- [ICML 2025] ReflectionBench: Evaluating Epistemic Agency in Large Language Models☆22Jun 24, 2025Updated last year
- A custom chatbot built with Next.js and OpenAI's GPT-3 language model API.☆13Aug 13, 2026Updated 2 weeks ago
- ☆12Jul 4, 2024Updated 2 years ago
- [ACL 2026] R-Search: Empowering LLM Reasoning with Search via Multi-Reward Reinforcement Learning☆36Aug 11, 2026Updated 2 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Q-Probe: A Lightweight Approach to Reward Maximization for Language Models☆41Jun 10, 2024Updated 2 years ago
- Constrained Decoding Project☆20Nov 10, 2023Updated 2 years ago
- Code and data for the EMNLP 2021 paper "Just Say No: Analyzing the Stance of Neural Dialogue Generation in Offensive Contexts". Coming so…☆17Jul 27, 2023Updated 3 years ago
- This repository contains a Jupyter Notebook that implements Gaussian Mixture Model (GMM) for semantic segmentation and background extract…☆15Jun 5, 2023Updated 3 years ago
- INSCIT: Information-Seeking Conversations with Mixed-Initiative Interactions☆16Jan 21, 2025Updated last year
- ☆16Aug 14, 2019Updated 7 years ago
- Connect6 AI based on reinforcement learning☆12Sep 13, 2019Updated 6 years ago