[NeurIPS '23 Spotlight] Thought Cloning: Learning to Think while Acting by Imitating Human Thinking
☆268Jun 28, 2024Updated 2 years ago
Alternatives and similar repositories for Thought-Cloning
Users that are interested in Thought-Cloning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Guide Your Agent with Adaptive Multimodal Rewards (NeurIPS 2023 Accepted)☆33Sep 25, 2023Updated 2 years ago
- TART: A plug-and-play Transformer module for task-agnostic reasoning☆201Jun 22, 2023Updated 3 years ago
- Lamorel is a Python library designed for RL practitioners eager to use Large Language Models (LLMs).☆249Dec 11, 2025Updated 8 months ago
- ☆15Apr 26, 2025Updated last year
- SwiftSage: A Generative Agent with Fast and Slow Thinking for Complex Interactive Tasks☆327Oct 22, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for "Learning to Model the World with Language." ICML 2024 Oral.☆421Jan 7, 2026Updated 7 months ago
- OMNI: Open-endedness via Models of human Notions of Interestingness☆66Jan 28, 2025Updated last year
- ☆27May 7, 2025Updated last year
- [NeurIPS'23 Spotlight] "Mind2Web: Towards a Generalist Agent for the Web" -- the first LLM-based web agent and benchmark for generalist w…☆1,020Nov 5, 2025Updated 9 months ago
- Codes for "Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models".☆1,139Dec 23, 2023Updated 2 years ago
- STEVE-1: A Generative Model for Text-to-Behavior in Minecraft☆216Jun 4, 2024Updated 2 years ago
- Codes for Evolving Plastic ANNs☆15Dec 18, 2022Updated 3 years ago
- ☆21Oct 6, 2023Updated 2 years ago
- Official codebase for "SelFee: Iterative Self-Revising LLM Empowered by Self-Feedback Generation"☆226Jun 6, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Repo to reproduce the First-Explore paper results☆39May 6, 2026Updated 3 months ago
- ☆25Oct 11, 2024Updated last year
- A simple wrapper for OpenAI to log input/outputs.☆106Aug 28, 2023Updated 2 years ago
- ☆459Sep 27, 2023Updated 2 years ago
- ☆146May 2, 2024Updated 2 years ago
- This repository contains some of the code used in the paper "Training Language Models with Langauge Feedback at Scale"☆26Mar 30, 2023Updated 3 years ago
- Official repository for ACL 2025 paper "Model Extrapolation Expedites Alignment"☆75May 20, 2025Updated last year
- LLMs can generate feedback on their work, use it to improve the output, and repeat this process iteratively.☆818Oct 4, 2024Updated last year
- Official repo for NAACL 2024 Findings paper "LeTI: Learning to Generate from Textual Interactions."☆66Jun 29, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Official implementation of the DECKARD Agent from the paper "Do Embodied Agents Dream of Pixelated Sheep?"☆94May 23, 2023Updated 3 years ago
- Provably (and non-vacuously) bounding test error of deep neural networks under distribution shift with unlabeled test data.☆10Feb 27, 2024Updated 2 years ago
- ☆223Jun 6, 2023Updated 3 years ago
- Simple next-token-prediction for RLHF☆228Sep 30, 2023Updated 2 years ago
- Code for the arXiv preprint "The Unreasonable Effectiveness of Easy Training Data"☆48Jan 17, 2024Updated 2 years ago
- ☆55Sep 9, 2023Updated 2 years ago
- (NeurIPS '22) LISA: Learning Interpretable Skill Abstractions - A framework for unsupervised skill learning using Imitation☆29Feb 22, 2023Updated 3 years ago
- ☆1,063May 29, 2023Updated 3 years ago
- Code for LaMPP: Language Models as Probabilistic Priors for Perception and Action☆37Apr 3, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Controllability-Aware Unsupervised Skill Discovery (ICML 2023)☆30Jun 3, 2023Updated 3 years ago
- [ACL2023] We introduce LLM-Blender, an innovative ensembling framework to attain consistently superior performance by leveraging the dive…☆991Oct 22, 2024Updated last year
- Entailment self-training☆27May 30, 2023Updated 3 years ago
- Official Implementation of "Graph of Thoughts: Solving Elaborate Problems with Large Language Models"☆2,832Mar 24, 2026Updated 4 months ago
- Intrinsic Motivation from Artificial Intelligence Feedback☆136Nov 7, 2023Updated 2 years ago
- Code for Arxiv 2023: Improving Language Model Negociation with Self-Play and In-Context Learning from AI Feedback☆208May 24, 2023Updated 3 years ago
- Codes and files for the paper Are Emergent Abilities in Large Language Models just In-Context Learning☆33Jan 9, 2025Updated last year