Code for the paper <SelfCheck: Using LLMs to Zero-Shot Check Their Own Step-by-Step Reasoning>
☆48Aug 1, 2023Updated 2 years ago
Alternatives and similar repositories for SelfCheck
Users that are interested in SelfCheck are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Release of the ConditionalQA dataset☆21Nov 2, 2021Updated 4 years ago
- Codes and Data for Scaling Relationship on Learning Mathematical Reasoning with Large Language Models☆270Sep 12, 2024Updated last year
- ☆11Nov 27, 2022Updated 3 years ago
- ☆26May 30, 2023Updated 2 years ago
- The is the official implementation of "Lyra: Orchestrating Dual Correction in Automated Theorem Proving"☆15Jul 2, 2024Updated last year
- The Lean Theorem Proving Environment☆15May 7, 2023Updated 2 years ago
- ☆30Dec 27, 2024Updated last year
- A dataset for natural language code search.☆14Feb 13, 2020Updated 6 years ago
- This is the artifact for paper “Are Machine Learning Cloud APIs Used Correctly? (#421)” in ICSE2021☆16Feb 27, 2021Updated 5 years ago
- [ NeurIPS 2023 ] Official Codebase for "Aligning Synthetic Medical Images with Clinical Knowledge using Human Feedback"☆20Oct 19, 2023Updated 2 years ago
- ☆14Oct 11, 2023Updated 2 years ago
- MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models☆454Feb 1, 2024Updated 2 years ago
- Guidelines for our secondary layer of annotation adding multi-sentence AMR links☆12Sep 6, 2017Updated 8 years ago
- ☆13Oct 4, 2022Updated 3 years ago
- About The corresponding code from our paper " REFINER: Reasoning Feedback on Intermediate Representations" (EACL 2024). Do not hesitate t…☆74Jan 27, 2026Updated last month
- Source code for ACL 2021 paper "Automatic ICD Coding via Interactive Shared Representation Networks with Self-distillation Mechanism"☆14Jun 1, 2021Updated 4 years ago
- ☆25Aug 23, 2024Updated last year
- Code for RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs. ACL 2023.☆64Nov 27, 2024Updated last year
- Grade-School Math with Irrelevant Context (GSM-IC) benchmark is an arithmetic reasoning dataset built upon GSM8K, by adding irrelevant se…☆65Feb 13, 2023Updated 3 years ago
- Supporting code for ReCEval paper☆31Sep 14, 2024Updated last year
- Automated Machine Learning (AutoML) for Kaggle Competition☆32Jul 6, 2023Updated 2 years ago
- Code and data accompanying the paper "TRUE: Re-evaluating Factual Consistency Evaluation".☆84Feb 20, 2026Updated last month
- [NeurIPS 2023] PyTorch code for Can Language Models Teach? Teacher Explanations Improve Student Performance via Theory of Mind☆66Dec 21, 2023Updated 2 years ago
- Repo to reproduce the First-Explore paper results☆39Dec 25, 2024Updated last year
- Code for Arxiv 2023: Improving Language Model Negociation with Self-Play and In-Context Learning from AI Feedback☆209May 24, 2023Updated 2 years ago
- The official repository of "ChatCoT: Tool-Augmented Chain-of-Thought Reasoning on Chat-based Large Language Models"☆46Jun 2, 2023Updated 2 years ago
- An AI-powered coding assistant plugin for the Eclipse IDE.☆13Oct 28, 2025Updated 4 months ago
- [ 𝐄𝐌𝐍𝐋𝐏 𝐅𝐢𝐧𝐝𝐢𝐧𝐠𝐬 𝟐𝟎𝟐𝟒 & 𝐀𝐂𝐋 𝟐𝟎𝟐𝟒 𝐍𝐋𝐑𝐒𝐄 𝐎𝐫𝐚𝐥] 𝘌𝘯𝘩𝘢𝘯𝘤𝘪𝘯𝘨 𝘔𝘢𝘵𝘩𝘦𝘮𝘢𝘵𝘪𝘤𝘢𝘭 𝘙𝘦𝘢𝘴𝘰𝘯𝘪𝘯…☆51May 4, 2024Updated last year
- ☆14Jul 24, 2024Updated last year
- {DeepL, Google, WMT-Best, davinci-003, turbo, gpt-4} × {En-De, En-Cs, En-Ru, En-Zh, De-Fr, En-Ja, Uk-En, Uk-Cs, En-Hr, En-Ha, En-Is}☆14Jun 18, 2023Updated 2 years ago
- VQA-Med 2021☆22Jul 11, 2022Updated 3 years ago
- 中国执业医师、药师、护士资格考试数据集和ChatGPT评估☆14Mar 13, 2026Updated last week
- ⚡Research papers about leveraging the capabilities of language models⚡☆53Jan 13, 2026Updated 2 months ago
- CVPR2024 highlight.☆13Oct 10, 2024Updated last year
- The data and implementation for the experiments in the paper "Flows: Building Blocks of Reasoning and Collaborating AI".☆31Feb 12, 2024Updated 2 years ago
- Adapt MLLMs to Domains via Post-Training (EMNLP 2025 Findings)☆13Nov 11, 2025Updated 4 months ago
- [APSIPA ASC 2023] The official code of paper, "FactLLaMA: Optimizing Instruction-Following Language Models with External Knowledge for Au…☆17Mar 7, 2024Updated 2 years ago
- A minimal language for Isabelle/HOL, designed for easing machine learning.☆25Jan 13, 2026Updated 2 months ago
- Loop Nest - Linear algebra compiler and code generator.☆20Oct 22, 2022Updated 3 years ago