augmented LLM with self reflection
☆144Nov 21, 2023Updated 2 years ago
Alternatives and similar repositories for awesome-llm-self-reflection
Users that are interested in awesome-llm-self-reflection are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- code for EMNLP 2024 paper: Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis☆12Nov 17, 2024Updated last year
- ☆24Nov 11, 2024Updated last year
- Self-Reflection in LLM Agents: Effects on Problem-Solving Performance☆102Nov 25, 2024Updated last year
- Let's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning with LLMs☆41Jan 30, 2024Updated 2 years ago
- [AAAI26] Trade-offs in Large Reasoning Models: An Empirical Analysis of Deliberative and Adaptive Reasoning over Foundational Capabilitie…☆11Feb 7, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Researchers who published code, models (in some cases), and demo apps (in few cases) along with their SOTA paper☆12Oct 19, 2023Updated 2 years ago
- Awesome LLM Self-Consistency: a curated list of Self-consistency in Large Language Models☆129Jul 20, 2025Updated last year
- ☆17Apr 26, 2024Updated 2 years ago
- The substitution of qsub.☆12Jan 25, 2019Updated 7 years ago
- Official repo for the TMLR paper "Discffusion: Discriminative Diffusion Models as Few-shot Vision and Language Learners"☆29Apr 27, 2024Updated 2 years ago
- ☆21Dec 7, 2024Updated last year
- 🤝 The code for "Can Large Language Model Agents Simulate Human Trust Behaviors?"☆119Jul 11, 2026Updated last month
- ☆13Apr 13, 2024Updated 2 years ago
- Llemma formal2formal (tactic prediction) theorem proving experiments☆20Oct 17, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- This is a collection of research papers for Self-Correcting Large Language Models with Automated Feedback.☆572Oct 28, 2024Updated last year
- Code for SafeMERGE (ICLR 2025).☆15Apr 1, 2025Updated last year
- Analyzing LLM Alignment via Token distribution shift☆17Jan 26, 2024Updated 2 years ago
- This repository contains the implementation of Concept Activation Regions, a new framework to explain deep neural networks with human con…☆17Oct 7, 2022Updated 3 years ago
- [ICML'24] TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks☆33Sep 20, 2024Updated last year
- ☆195Mar 8, 2026Updated 5 months ago
- From Chain-of-Thought prompting to OpenAI o1 and DeepSeek-R1 🍓☆3,681Apr 20, 2026Updated 4 months ago
- papers related to LLM-agent that published on top conferences☆319Apr 14, 2025Updated last year
- ☆16May 25, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An implementation of effective policy ensemble.☆16Jul 5, 2023Updated 3 years ago
- [ICLR 2025] Code for Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models☆26Apr 14, 2025Updated last year
- Open-source Human Feedback Library☆11Oct 25, 2023Updated 2 years ago
- ☆73Apr 2, 2024Updated 2 years ago
- [NeurIPS 2025@FoRLM] R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search☆17Jan 24, 2026Updated 7 months ago
- ☆23Aug 7, 2023Updated 3 years ago
- A Python 3 Bandit Visualization Package☆11Oct 16, 2017Updated 8 years ago
- Must-read Papers on LLM Agents.☆3,110Jul 27, 2026Updated last month
- [ACL'26] Official Repository for for paper "Data-Efficient RLVR via Off-Policy Influence Guidance"☆25Jul 26, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for NeurIPS'24 paper 'Grokked Transformers are Implicit Reasoners: A Mechanistic Journey to the Edge of Generalization'☆242Jul 19, 2025Updated last year
- EMNLP 2024: Model Editing Harms General Abilities of Large Language Models: Regularization to the Rescue☆37May 26, 2025Updated last year
- Awesome LLM papers, news and projects about learning to reason with LLM, OpenAI o1, reasonning techniques, chain-of-thought (COT), Large …☆28Oct 10, 2024Updated last year
- Adaptive Hardness Negative Sampling for Collaborative Filtering, AAAI2024☆13Dec 13, 2023Updated 2 years ago
- opentqa is a open framework of the textbook question answering, which includes xtqa, mcan, cmr, mfb, mutan.☆11Mar 27, 2021Updated 5 years ago
- ☆23Sep 28, 2023Updated 2 years ago
- A probabilitic model for contextual word representation. Accepted to ACL2023 Findings.☆25Oct 22, 2023Updated 2 years ago