Official implementation for ICLR 2025 paper "Distilling Reinforcement Learning Algorithms for In-Context Model-Based Planning"
☆23Mar 5, 2025Updated last year
Alternatives and similar repositories for dicp
Users that are interested in dicp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Next-gen Foundation Model for Embodied AI☆34Apr 7, 2026Updated 5 months ago
- [NeurIPS 2025] Official codebase for T2MIR: Mixture-of-Experts Meets In-Context Reinforcement Learning.☆34Oct 26, 2025Updated 11 months ago
- Benchmark for evaluating the generalization capabilities of Multi-Objective Reinforcement Learning (MORL) algorithms.☆30Jun 6, 2025Updated last year
- 🔥 [ICLR 2026] Official implementation of Recurrent Action Transformer with Memory, an offline RL agent with memory mechanisms. https://s…☆26Nov 23, 2025Updated 10 months ago
- End-to-End Humanoid Robot Safe and Comfortable Locomotion Policy☆30Oct 23, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆17May 17, 2024Updated 2 years ago
- ☆13May 21, 2023Updated 3 years ago
- Minimal Transformer base in JAX. A single backbone for language modelling, diffusion, classification, etc...☆16May 28, 2025Updated last year
- Code for AAAI 2023 paper "Hypernetworks for Zero-shot Transfer in Reinforcement Learning"☆24Apr 26, 2023Updated 3 years ago
- Retrieval-Augmented Decision Transformer: External Memory for In-context RL☆26Oct 27, 2024Updated last year
- Implemention of the Decision-Pretrained Transformer (DPT) from the paper Supervised Pretraining Can Learn In-Context Reinforcement Learni…☆81May 28, 2024Updated 2 years ago
- [IEEE TPAMI] A Framework for Constrained Multi-Objective Reinforcement Learning☆21Apr 18, 2025Updated last year
- Awesome In-Context RL: A curated list of In-Context Reinforcement Learning - - —☆309Sep 8, 2025Updated last year
- Official Repository for "Scaling Multi-Agent Reinforcement Learning with Selective Parameter Sharing" (ICML2021)☆10Oct 26, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ICLR 25 Spotlight, Transformer-based Off-Policy Episodic RL☆17Feb 18, 2025Updated last year
- Latent Dynamics Mixture, NeurIPS 2021☆18Oct 25, 2022Updated 3 years ago
- [ICLR 2026] Meta-RL Induces Exploration in Language Agents☆49Feb 1, 2026Updated 7 months ago
- off-policy RL on long sequences☆172May 29, 2026Updated 3 months ago
- Decision Mamba: Reinforcement Learning via Sequence Modeling with Selective State Spaces☆54Apr 1, 2024Updated 2 years ago
- The Codebase of <Towards an Information Theoretic Framework of Context-Based Offline Meta-Reinforcement Learning> In NeurIPS 2024☆29Feb 20, 2025Updated last year
- ☆25Sep 23, 2024Updated 2 years ago
- This repository contains the official code for our NeurIPS 2021 publication "Robust Deep Reinforcement Learning through Adversarial Loss…☆33Jan 21, 2022Updated 4 years ago
- Related papers for Continual Reinforcement Learning.☆59Feb 8, 2026Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official implementation for paper 'Bridging Adaptivity and Safety'.☆28Jun 4, 2025Updated last year
- ☆17Sep 25, 2024Updated 2 years ago
- ☆14Sep 29, 2025Updated 11 months ago
- ☆16Jan 30, 2025Updated last year
- ☆38Apr 3, 2026Updated 5 months ago
- ☆10May 14, 2024Updated 2 years ago
- 🔥 Datasets and env wrappers for offline safe reinforcement learning☆141Nov 12, 2025Updated 10 months ago
- [ICML'23] Official PyTorch Implementation of NA2Q, and a comprehensive benchmark based on pymarl☆24Jan 14, 2024Updated 2 years ago
- This is the official code implementation of Bongard-OpenWorld (ICLR 2024).☆14Jan 6, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆13Aug 19, 2024Updated 2 years ago
- Official Implementation for "In-Context Reinforcement Learning from Noise Distillation"☆35Sep 18, 2024Updated 2 years ago
- The 4th rank system of the SemEval 2021 Task4.☆10May 7, 2022Updated 4 years ago
- DeltaProduct is a new linear recurrent neural network architecture that uses products of generalized Householder matrices as state-transi…☆18Oct 13, 2025Updated 11 months ago
- Official code for "World Models via Policy-Guided Trajectory Diffusion", TMLR 2024☆77Mar 22, 2024Updated 2 years ago
- ☆21Aug 25, 2026Updated last month
- Official Implementation for "In-Context Reinforcement Learning for Variable Action Spaces"☆93Feb 11, 2024Updated 2 years ago