Code for paper "Learning Multimodal Transition Dynamics for Model-Based Reinforcement Learning".
☆34May 24, 2018Updated 8 years ago
Alternatives and similar repositories for multimodal_varinf
Users that are interested in multimodal_varinf are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- In Progress : State of the art Distributed Distributional Deep Deterministic Policy Gradient algorithm implementation in pytorch.☆19Jun 15, 2018Updated 8 years ago
- RL framework for embodied agents based on PyTorch☆11Apr 11, 2019Updated 7 years ago
- Proportional-Derivative Neural Networks, as described in Temporally Efficient Deep Learning with Spikes☆16Sep 5, 2017Updated 9 years ago
- A comparison of parameter space noise methods for exploration in deep reinforcement learning☆30Mar 14, 2019Updated 7 years ago
- ☆32Nov 13, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆99Mar 24, 2023Updated 3 years ago
- Hypergraph Case-Based Reasoning☆11Jun 14, 2018Updated 8 years ago
- ☆69May 26, 2018Updated 8 years ago
- ROS integration for Franka Emika research robots in Cognitive Robotics TU Delft☆10Jan 18, 2023Updated 3 years ago
- NIPS 2017 Value Prediction Network☆166Jan 12, 2018Updated 8 years ago
- [ICLR 22] Value Gradient weighted Model-Based Reinforcement Learning.☆26Apr 15, 2023Updated 3 years ago
- Library that provides environments for planning problems☆17Apr 24, 2026Updated 5 months ago
- Convert sc2 environment to gym-atari and play some mini-games☆21Oct 28, 2017Updated 8 years ago
- Few-shot Bayesian Imitation Learning with Policies as Logic over Programs☆22Oct 19, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Distributed Priortized Experience Replay☆10Aug 8, 2018Updated 8 years ago
- ☆11Oct 14, 2019Updated 6 years ago
- Sequence to Sequence Learning Model☆14Jan 9, 2016Updated 10 years ago
- ☆85May 29, 2019Updated 7 years ago
- Causal decomposition analysis☆12Apr 17, 2022Updated 4 years ago
- Code Release for Task Agnostic Dynamics Priors for Deep Reinforcement Learning☆12Jun 13, 2019Updated 7 years ago
- Code to reproduce Supervised Policy Update (ICLR 2019)☆17Dec 8, 2022Updated 3 years ago
- Reinforcement Learning Seminar at the Chinese University of Hong Kong, Shenzhen, China.☆21Nov 17, 2023Updated 2 years ago
- Building Agents with Imagination: pytorch step-by-step implementation☆213Feb 22, 2019Updated 7 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Decision Transformer for offline single-agent autonomous highway driving☆29Jun 19, 2023Updated 3 years ago
- Python OpenSceneGraph bindings - rebirth.☆10Apr 25, 2016Updated 10 years ago
- Value Iteration Networks☆291Apr 21, 2017Updated 9 years ago
- On the model-based stochastic value gradient for continuous reinforcement learning☆58Mar 6, 2026Updated 6 months ago
- tensorflow reinforcement learning agents for OpenAI gym environments☆146Jul 21, 2017Updated 9 years ago
- Regularization Matters in Policy Optimization☆21Nov 1, 2021Updated 4 years ago
- A working implementation of the Categorical DQN (Distributional RL).☆95Apr 7, 2018Updated 8 years ago
- Toolbox and trackers for object pose-estimation. Based on the work CosyPose and MegaPose☆48Updated this week
- Code for "Planning with Learned Object Importance in Large Problem Instances using Graph Neural Networks" (AAAI 2021)☆18Jan 26, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Tutorial on how to build a crypto trading bot with Binance and Python☆13Apr 9, 2025Updated last year
- Code for "Exponential Family Estimation via Adversarial Dynamics Embedding" (NeurIPS 2019)☆14Nov 26, 2019Updated 6 years ago
- Ranking Policy Gradient☆23Nov 27, 2019Updated 6 years ago
- Official code for ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning (AAAI'24)☆16Feb 10, 2024Updated 2 years ago
- Code for a recurrent neural network (RNN) created by unfolding the sequential ISTA (SISTA) algorithm for sequential sparse coding☆14Mar 24, 2017Updated 9 years ago
- Deep Planning Network: Control from pixels by latent planning with learned dynamics☆380Oct 15, 2021Updated 4 years ago
- Improvements made to pietrolechthaler's and his group project titled: "UR5 Pick and Place Simulation in Ros/Gazebo", available in the nex…☆11May 10, 2023Updated 3 years ago