Code for the paper <SelfCheck: Using LLMs to Zero-Shot Check Their Own Step-by-Step Reasoning>
☆46Aug 1, 2023Updated 3 years ago
Alternatives and similar repositories for SelfCheck
Users that are interested in SelfCheck are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The baseline method for CCIR 22 https://www.datafountain.cn/competitions/573☆13Aug 2, 2022Updated 4 years ago
- Release of the ConditionalQA dataset☆22Nov 2, 2021Updated 4 years ago
- Codes and Data for Scaling Relationship on Learning Mathematical Reasoning with Large Language Models☆269Sep 12, 2024Updated 2 years ago
- ☆26May 30, 2023Updated 3 years ago
- ☆30Dec 27, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A dataset for natural language code search.☆14Feb 13, 2020Updated 6 years ago
- The is the official implementation of "Lyra: Orchestrating Dual Correction in Automated Theorem Proving"☆15Jul 2, 2024Updated 2 years ago
- The Lean Theorem Proving Environment☆15May 7, 2023Updated 3 years ago
- ☆14Oct 11, 2023Updated 2 years ago
- The official code of TACL 2021, "Did Aristotle Use a Laptop? A Question Answering Benchmark with Implicit Reasoning Strategies".☆88Oct 31, 2022Updated 3 years ago
- PyFed generic framework of benchmark for Federated Learning☆11Oct 9, 2021Updated 4 years ago
- This is the repo for the paper Shepherd -- A Critic for Language Model Generation☆225Aug 10, 2023Updated 3 years ago
- Experiment for lsat☆51Jan 20, 2023Updated 3 years ago
- Guidelines for our secondary layer of annotation adding multi-sentence AMR links☆12Sep 6, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- About The corresponding code from our paper " REFINER: Reasoning Feedback on Intermediate Representations" (EACL 2024). Do not hesitate t…☆77Jan 27, 2026Updated 7 months ago
- Source code for ACL 2021 paper "Automatic ICD Coding via Interactive Shared Representation Networks with Self-distillation Mechanism"☆14Jun 1, 2021Updated 5 years ago
- Code for Arxiv 2023: Improving Language Model Negociation with Self-Play and In-Context Learning from AI Feedback☆210May 24, 2023Updated 3 years ago
- [NeurIPS 2023] PyTorch code for Can Language Models Teach? Teacher Explanations Improve Student Performance via Theory of Mind☆66Dec 21, 2023Updated 2 years ago
- Repo to reproduce the First-Explore paper results☆39May 6, 2026Updated 4 months ago
- Grade-School Math with Irrelevant Context (GSM-IC) benchmark is an arithmetic reasoning dataset built upon GSM8K, by adding irrelevant se…☆67Feb 13, 2023Updated 3 years ago
- Supporting code for ReCEval paper☆33Sep 14, 2024Updated 2 years ago
- ☆26Aug 23, 2024Updated 2 years ago
- Code for AAAI 2023 accepted paper titled "Knowledge-Bridged Causal Interaction Network for Causal Emotion Entailment"☆14May 6, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [𝐄𝐌𝐍𝐋𝐏 𝐅𝐢𝐧𝐝𝐢𝐧𝐠𝐬 𝟐𝟎𝟐𝟒 & 𝐀𝐂𝐋 𝟐𝟎𝟐𝟒 𝐍𝐋𝐑𝐒𝐄 𝐎𝐫𝐚𝐥] 𝘌𝘯𝘩𝘢𝘯𝘤𝘪𝘯𝘨 𝘔𝘢𝘵𝘩𝘦𝘮𝘢𝘵𝘪𝘤𝘢𝘭 𝘙𝘦𝘢𝘴𝘰𝘯𝘪𝘯…☆52May 4, 2024Updated 2 years ago
- Software Testing | Tongji Univ. SSE Course Project☆18Jul 2, 2020Updated 6 years ago
- {DeepL, Google, WMT-Best, davinci-003, turbo, gpt-4} × {En-De, En-Cs, En-Ru, En-Zh, De-Fr, En-Ja, Uk-En, Uk-Cs, En-Hr, En-Ha, En-Is}☆14Jun 18, 2023Updated 3 years ago
- ⚡Research papers about leveraging the capabilities of language models⚡☆53Apr 15, 2026Updated 5 months ago
- [APSIPA ASC 2023] The official code of paper, "FactLLaMA: Optimizing Instruction-Following Language Models with External Knowledge for Au…☆18Mar 7, 2024Updated 2 years ago
- We have released the code and demo program required for LLM with self-verification☆61Oct 18, 2023Updated 2 years ago
- ☆22Jan 5, 2024Updated 2 years ago
- A minimal language for Isabelle/HOL, designed for easing machine learning.☆31Updated this week
- ☆10Nov 18, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆14Nov 20, 2022Updated 3 years ago
- 测试 https://huggingface.co/OFA-Sys/gsm8k-rft-llama7b-u13b 的 GSM8K 分数☆15Aug 10, 2023Updated 3 years ago
- Data and Code for Program of Thoughts [TMLR 2023]☆317May 15, 2024Updated 2 years ago
- ☆16Mar 6, 2025Updated last year
- ☆73Apr 2, 2024Updated 2 years ago
- Code and data for the paper: IntentionQA: A Benchmark for Evaluating Purchase Intention Comprehension Abilities of Large Language Models …☆12Apr 27, 2024Updated 2 years ago
- Code & Data for our Paper "Alleviating Hallucinations of Large Language Models through Induced Hallucinations"☆71Feb 27, 2024Updated 2 years ago