Code for demonstration example-task in RUDDER blog
☆24May 19, 2020Updated 6 years ago
Alternatives and similar repositories for rudder-demonstration-code
Users that are interested in rudder-demonstration-code are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A practical step-by-step guide to applying RUDDER☆36Nov 12, 2019Updated 6 years ago
- RUDDER: Return Decomposition for Delayed Rewards☆49Sep 17, 2020Updated 5 years ago
- Learning from Indirect Observations☆11Jul 16, 2021Updated 5 years ago
- ☆31Jan 16, 2023Updated 3 years ago
- ☆13Dec 6, 2018Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for the benchmark containing dataset, models and metrics for productive concept learning -- a kind of compositional reasoning task t…☆19Jul 22, 2021Updated 5 years ago
- Deep direct reinforcement learning for financial signal representation and trading☆31Oct 7, 2020Updated 5 years ago
- [SIGKDD' 24] PyTorch implementation of Temporal Prototype-Aware Learning for Active Voltage Control on Power Distribution Networks☆14Jul 28, 2024Updated 2 years ago
- Code to reproduce results on toy tasks and companion blog for the paper.☆25Jun 8, 2022Updated 4 years ago
- ☆33Aug 30, 2024Updated 2 years ago
- Change-Based Exploration Transfer☆35Apr 24, 2022Updated 4 years ago
- TensorFlow implementation for our paper "Learning Long-Term Reward Redistribution via Randomized Return Decomposition"☆19Mar 17, 2022Updated 4 years ago
- Invariant Causal Prediction for Block MDPs☆44Jun 11, 2020Updated 6 years ago
- ♊ Minimal PyTorch Twin Delayed DDPG (TD3) implementation☆10Jun 20, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Trajectory-wise Multiple Choice Learning for Dynamics Generalization in Reinforcement Learning (NeurIPS 2020)☆38Oct 27, 2020Updated 5 years ago
- Official PyTorch implementation of "ACE:Off-Policy Actor-Critic with Causality-Aware Entropy Regularization"☆35May 13, 2024Updated 2 years ago
- ☆14Mar 5, 2024Updated 2 years ago
- OpenaAI Gym Franka Emika Panda robot environment based on PyBullet.☆12Sep 8, 2023Updated 2 years ago
- PyTorch implementation of R2D2 (Recurrent Reply Distributed DQN)☆13Nov 14, 2019Updated 6 years ago
- Code for Optimistic Exploration even with a Pessimistic Initialisation☆14Aug 4, 2020Updated 6 years ago
- Model-based Offline Policy Optimization re-implement all by pytorch☆44Sep 13, 2023Updated 2 years ago
- ☆14May 20, 2023Updated 3 years ago
- CausalWorld: A Robotic Manipulation Benchmark for Causal Structure and Transfer Learning☆249Nov 1, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Continual Reinforcement Learning in 3D Non-stationary Environments☆39Jun 16, 2019Updated 7 years ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated 2 years ago
- ☆15Oct 26, 2020Updated 5 years ago
- ☆12Mar 21, 2024Updated 2 years ago
- A tool to automatically label, classify, and count marine debris in your aerial imagery. Designed to automate the tedious parts of standi…☆11Nov 22, 2023Updated 2 years ago
- ☆10Aug 8, 2021Updated 5 years ago
- Pytorch code for "Learning Guidance Rewards with Trajectory-space Smoothing" (NeurIPS 2020)☆12Jul 7, 2021Updated 5 years ago
- The source code of our ACL paper "A Training-free and Reference-free Summarization Evaluation Metric via Centrality-weighted Relevance an…☆14May 6, 2023Updated 3 years ago
- Datasets for data-driven deep reinforcement learning with Atari (wrapper for datasets released by Google)☆128Aug 30, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆10Oct 15, 2020Updated 5 years ago
- A repository with code experimenting with the different Machine Learning algorithms for Point Clouds included with Open3D-ML☆15Oct 10, 2022Updated 3 years ago
- PyTorch Implementation of "Language as an Abstraction for Hierarchical Deep Reinforcement Learning" paper☆25Feb 14, 2022Updated 4 years ago
- Code for optimal execution☆12Oct 29, 2020Updated 5 years ago
- MATLAB code and data for the paper “Optimal energy management of offshore wind farms considering the combination of overplanting and dyna…☆18Jun 23, 2024Updated 2 years ago
- VibeRL is a Reinforcement Learning framework built essentially through vibe coding with Kimi K2.☆18Aug 24, 2026Updated last week
- Code associated with the NeurIPS19 paper "Weighted Linear Bandits in Non-Stationary Environments"☆17Nov 14, 2019Updated 6 years ago