Related papers for offline reforcement learning (we mainly focus on representation and sequence modeling and conventional offline RL)
☆19Apr 21, 2022Updated 4 years ago
Alternatives and similar repositories for Papers-of-Offline-RL
Users that are interested in Papers-of-Offline-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Codes accompanying the paper "Offline Reinforcement Learning with Value-Based Episodic Memory" (ICLR 2022 https://arxiv.org/abs/2110.0979…☆15Mar 9, 2022Updated 4 years ago
- Codes accompanying the paper "Believe What You See: Implicit Constraint Approach for Offline Multi-Agent Reinforcement Learning" (NeurIPS…☆75Oct 18, 2022Updated 3 years ago
- Author implementation of Monte Carlo Augmented Actor Critic in PyTorch☆18Oct 24, 2022Updated 3 years ago
- Safe Reinforcement Learning with Natural Language Constraints☆17Oct 24, 2021Updated 4 years ago
- Reinforcement learning - Batched Impala - PyTorch - Mario Kart☆14Jul 21, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code repository accompanying the Heuristic Guided RL NeurIPS'21 paper☆17Jan 3, 2022Updated 4 years ago
- ☆10Sep 19, 2023Updated 3 years ago
- Counterfactual explanations for Reinforcement Learning agents on Atari☆12Apr 3, 2023Updated 3 years ago
- The exact codes used by the team "liveinparis" at the kaggle football competition ranked 6th/1141☆57Dec 14, 2020Updated 5 years ago
- Code for Mildly Conservative Q-learning for Offline Reinforcement Learning (NeurIPS 2022)☆62Apr 29, 2024Updated 2 years ago
- [S&P 2024] Replication Package for "Mind Your Data! Hiding Backdoors in Offline Reinforcement Learning Datasets".☆33Dec 30, 2024Updated last year
- PRML Page-by-page配套资料,对PRML全书及各章节的review☆18Apr 16, 2024Updated 2 years ago
- (ICLR 2021) Learning to Represent Action Values as a Hypergraph on the Action Vertices☆23Jun 22, 2021Updated 5 years ago
- Generalised UDRL☆37May 12, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆10Oct 15, 2020Updated 5 years ago
- Official codebase for Exact Energy-Guided Diffusion Sampling via Contrastive Energy Prediction☆36Nov 3, 2023Updated 2 years ago
- Assignments for CS294-112.☆29Sep 11, 2019Updated 7 years ago
- Decision Transformer: A brand new Offline RL Pattern.☆38Jan 28, 2022Updated 4 years ago
- ☆27Apr 24, 2020Updated 6 years ago
- Web application where humans can play Overcooked with AI agents.☆61Dec 6, 2022Updated 3 years ago
- [ICLR 2024 Spotlight] Code for the paper "Decision ConvFormer: Local Filtering in MetaFormer is Sufficient for Decision Making"☆13Apr 22, 2024Updated 2 years ago
- Code for NeurIPS 2022 paper "Robust offline Reinforcement Learning via Conservative Smoothing"☆23Feb 15, 2023Updated 3 years ago
- PyTorch implementation of the implicit Q-learning algorithm (IQL)☆44Dec 17, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Simulation of car parking in different parking lots using Unity ML-Agents☆13Dec 16, 2023Updated 2 years ago
- ☆18Sep 23, 2022Updated 4 years ago
- Invariant Causal Prediction for Block MDPs☆44Jun 11, 2020Updated 6 years ago
- ☆12Feb 20, 2021Updated 5 years ago
- Challenges and Opportunities in Offline Reinforcement Learning from Visual Observations☆116Apr 16, 2026Updated 5 months ago
- Setup for Octo and some experiments with the model☆12Apr 11, 2024Updated 2 years ago
- This is an official implementation of the paper ``Building Math Agents with Multi-Turn Iterative Preference Learning'' with multi-turn DP…☆31Dec 5, 2024Updated last year
- Arena: A General Evaluation Platform and Building Toolkit for Single/Multi-Agent Intelligence. AAAI 2020.☆103Mar 6, 2025Updated last year
- The implementation of Discriminator Soft Actor Critic☆15Jan 25, 2020Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Codes for the paper "SIDE: State Inference for Partially Observable Cooperative Multi-Agent Reinforcement Learning"☆11Jun 24, 2022Updated 4 years ago
- source code of paper 'Auto-STGCN: Autonomous Spatial-Temporal Graph Convolutional Network Search Based on Reinforcement Learning and Exis…☆11Jan 26, 2021Updated 5 years ago
- Exploration by Random Network Distillation☆15Dec 30, 2018Updated 7 years ago
- Official code for ICLR 2024 paper, SEABO: A Simple Search-Based Method for Offline Imitation Learning☆14Jan 19, 2024Updated 2 years ago
- Datasets for data-driven deep reinforcement learning with PyBullet environments☆152Mar 19, 2021Updated 5 years ago
- Represented Value Function Approach for Large Scale Multi Agent Reinforcement Learning☆17Mar 11, 2020Updated 6 years ago
- Code for Latent Action Space for Offline Reinforcement Learning [CoRL 2020]☆54Oct 18, 2021Updated 4 years ago