Plancraft is a minecraft environment and agent suite to test planning capabilities in LLMs
☆31Nov 7, 2025Updated 9 months ago
Alternatives and similar repositories for plancraft
Users that are interested in plancraft are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-Word Probabilistic based supertokenizer☆15May 15, 2025Updated last year
- This repository contains code for the paper "Uncertainty Estimation and Calibration with Finite-State Probabilistic RNNs" (Wang, Lawrence…☆17Mar 8, 2021Updated 5 years ago
- ML framework to estimate Bayesian posteriors of galaxy morphological parameters☆12Jul 10, 2025Updated last year
- [ACL 2025 Findings] Text2World: Benchmarking Large Language Models for Symbolic World Model Generation☆29Feb 25, 2025Updated last year
- Uncertainty Quantification with Pre-trained Language Models: An Empirical Analysis☆15Oct 11, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of "LM-Infinite: Simple On-the-Fly Length Generalization for Large Language Models"☆40Nov 11, 2024Updated last year
- Benchmark Test-Time Scaling of General LLM Agents☆21Apr 14, 2026Updated 4 months ago
- RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment☆18Dec 19, 2024Updated last year
- ☆21Apr 3, 2026Updated 4 months ago
- Code for "Multi-Objective GFlowNets"☆20Jul 12, 2023Updated 3 years ago
- [NeurIPS 2024] The official implementation of "Image Copy Detection for Diffusion Models"☆18Oct 1, 2024Updated last year
- Code for "Can Retriever-Augmented Language Models Reason? The Blame Game Between the Retriever and the Language Model", EMNLP Findings 20…☆27Nov 2, 2023Updated 2 years ago
- [PAKDD-2021] Hierarchical Self Attention Based Autoencoder for Open-Set Human Activity Recognition☆17May 23, 2023Updated 3 years ago
- Applying Deep Reinforcement Learning for dialogue generation. aka chatbot☆13Apr 30, 2017Updated 9 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- We study toy models of skill learning.☆35Feb 3, 2026Updated 6 months ago
- [ICLR 26] The official code repository for the paper "Mirage or Method? How Model–Task Alignment Induces Divergent RL Conclusions".☆19Feb 9, 2026Updated 6 months ago
- ☆24Oct 27, 2023Updated 2 years ago
- ☆11Oct 7, 2024Updated last year
- The Conceptual Coverage Across Languages Benchmark for Text-to-Image Models☆12Oct 28, 2024Updated last year
- A reinforcement learning environment for the IGLU 2022 at NeurIPS☆36May 27, 2023Updated 3 years ago
- ☆12Feb 6, 2021Updated 5 years ago
- Control LLM☆23Apr 6, 2025Updated last year
- ☆10Nov 14, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆21Apr 27, 2026Updated 3 months ago
- Hospital simulator with pedestrians and robot☆15Oct 20, 2024Updated last year
- OneEdit: A Neural-Symbolic Collaboratively Knowledge Editing System.☆20Oct 14, 2024Updated last year
- Base repo for paper 'StyleMeUp: Towards Style-Agnostic Sketch-Based Image Retrieval'☆15Apr 27, 2022Updated 4 years ago
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 9 months ago
- ☆15Apr 6, 2026Updated 4 months ago
- Source code for our paper: "ARIA: Training Language Agents with Intention-Driven Reward Aggregation".☆30Aug 9, 2025Updated last year
- Code and data for "Medical Dialogue Generation via Dual Flow Modeling" (ACL 2023 Findings)☆14Nov 22, 2023Updated 2 years ago
- [ICLR 2025] Benchmarking Agentic Workflow Generation☆156Feb 19, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Hand Gesture Controlled Tello Drone using Python and OpenCV 2021☆12Jun 6, 2022Updated 4 years ago
- An open-source framework to benchmark and assess safety specifications of Reinforcement Learning problems.☆14Aug 25, 2023Updated 2 years ago
- Swire Dataset and Application Code☆17Jan 7, 2019Updated 7 years ago
- [ICLR 2025] Bridging and Modeling Correlations in Pairwise Data for Direct Preference Optimization☆12Jan 26, 2025Updated last year
- Code for L4DC 2022 paper: Joint Synthesis of Safety Certificate and Safe Control Policy Using Constrained Reinforcement Learning.☆14Jul 31, 2023Updated 3 years ago
- Code and data for the Nature Machine Intelligence paper "Knowledge graph-enhanced molecular contrastive learning with functional prompt".☆11May 16, 2023Updated 3 years ago
- ☆25Aug 11, 2026Updated last week