Code and data accompanying our paper on arXiv "Faithful Chain-of-Thought Reasoning".
☆169May 7, 2024Updated 2 years ago
Alternatives and similar repositories for Faithful-COT
Users that are interested in Faithful-COT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆88Jun 1, 2023Updated 3 years ago
- ☆76Apr 27, 2024Updated 2 years ago
- [EMNLP 2023] The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning☆258Oct 31, 2023Updated 2 years ago
- ☆15Nov 22, 2023Updated 2 years ago
- ☆19Nov 7, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- NaturalProver: Grounded Mathematical Proof Generation with Language Models☆40Mar 24, 2023Updated 3 years ago
- paper list on reasoning in NLP☆197Apr 7, 2025Updated last year
- implementation of paper "Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners"☆20Aug 17, 2023Updated 2 years ago
- ☆157Mar 18, 2023Updated 3 years ago
- This repository contains a collection of papers and resources on Reasoning in Large Language Models.☆572Nov 13, 2023Updated 2 years ago
- PaL: Program-Aided Language Models (ICML 2023)☆525Jun 30, 2023Updated 3 years ago
- Your finetuned model's back to its original safety standards faster than you can say "SafetyLock"!☆11Oct 16, 2024Updated last year
- [ACL 2023] Reasoning with Language Model Prompting: A Survey☆1,008May 21, 2025Updated last year
- Probabilistic LLM evaluations. [CogSci2023; ACL2023]☆73Jul 27, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆16Mar 22, 2025Updated last year
- [ACL 2023 Findings] What In-Context Learning “Learns” In-Context: Disentangling Task Recognition and Task Learning☆21Jul 9, 2023Updated 3 years ago
- ☆103Dec 7, 2023Updated 2 years ago
- A central, open resource for data and tools related to chain-of-thought reasoning in large language models. Developed @ Samwald research …☆1,014Dec 16, 2024Updated last year
- [EMNLP 2024] Official implementation of "Hierarchical Deconstruction of LLM Reasoning: A Graph-Based Framework for Analyzing Knowledge Ut…☆23Dec 4, 2024Updated last year
- Grade-School Math with Irrelevant Context (GSM-IC) benchmark is an arithmetic reasoning dataset built upon GSM8K, by adding irrelevant se…☆67Feb 13, 2023Updated 3 years ago
- Codes and Data for Scaling Relationship on Learning Mathematical Reasoning with Large Language Models☆270Sep 12, 2024Updated last year
- GenRM-CoT: Data release for verification rationales☆68Oct 16, 2024Updated last year
- Large Language Models Are Reasoning Teachers (ACL 2023)☆345Mar 7, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Progressive Prompts: Continual Learning for Language Models☆96Apr 24, 2023Updated 3 years ago
- IntructIR, a novel benchmark specifically designed to evaluate the instruction following ability in information retrieval models. Our foc…☆32Jun 13, 2024Updated 2 years ago
- This is the official implementation of "Progressive-Hint Prompting Improves Reasoning in Large Language Models"☆208Oct 11, 2023Updated 2 years ago
- Code for paper "LEVER: Learning to Verifiy Language-to-Code Generation with Execution" (ICML'23)☆90Jul 5, 2023Updated 3 years ago
- Benchmarking large language models' complex reasoning ability with chain-of-thought prompting☆2,776Aug 4, 2024Updated last year
- A trend starts from "Chain of Thought Prompting Elicits Reasoning in Large Language Models".☆2,105Oct 5, 2023Updated 2 years ago
- the instructions and demonstrations for building a formal logical reasoning capable GLM☆54Sep 3, 2024Updated last year
- ☆57Oct 23, 2023Updated 2 years ago
- Repository for "Scaling Evaluation-time Compute with Reasoning Models as Process Evaluators"☆12Mar 25, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Data and Code for Program of Thoughts [TMLR 2023]☆317May 15, 2024Updated 2 years ago
- Official implementation for "MM-Eval: A Multilingual Meta-Evaluation Benchmark for LLM-as-a-Judge and Reward Models"☆20Oct 26, 2024Updated last year
- ☆18May 17, 2022Updated 4 years ago
- the benchmark for finance☆11Jul 4, 2023Updated 3 years ago
- Code for Preventing Language Models From Hiding Their Reasoning, which evaluates defenses against LLM steganography.☆25Jan 26, 2024Updated 2 years ago
- ☆134Jul 8, 2024Updated 2 years ago
- LLMs can generate feedback on their work, use it to improve the output, and repeat this process iteratively.☆815Oct 4, 2024Updated last year