Shared autonomy via deep reinforcement learning
☆80Mar 24, 2023Updated 3 years ago
Alternatives and similar repositories for deepassist
Users that are interested in deepassist are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the paper "Residual Policy Learning for Shared Autonomy".☆18Apr 14, 2020Updated 6 years ago
- Inferring beliefs about dynamics from behavior☆30May 24, 2018Updated 8 years ago
- Code for the paper, "Learning Human Objectives by Evaluating Hypothetical Behavior"☆86Dec 13, 2019Updated 6 years ago
- Cooperative Multi Agent Reinforcement Learning with Human in the Loop☆14Apr 25, 2023Updated 3 years ago
- We implement MADDPG in a congestion env, and compare with several control groups to highlight the performance of MADDPG☆11Jul 14, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for the paper "Learning to Assist Humans without Inferring Rewards"☆20Jul 7, 2024Updated 2 years ago
- code for polite☆12Feb 28, 2024Updated 2 years ago
- Code for our paper: Hierarchical RL Using an Ensemble of Proprioceptive Periodic Policies☆16Feb 21, 2019Updated 7 years ago
- RL agent to play μRTS with Stable-Baselines3 and PyTorch☆28Jan 23, 2022Updated 4 years ago
- ☆16May 1, 2011Updated 15 years ago
- Implementation of POMDP algorithms on the tiger example, as described in Littman, Cassandra and Kaelbling (1994).☆17Aug 8, 2017Updated 9 years ago
- Spaced repetition through deep reinforcement learning☆68Jul 21, 2021Updated 5 years ago
- ☆47Jun 19, 2018Updated 8 years ago
- An implementation of the Latent Skill Embedding model☆10Feb 19, 2016Updated 10 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for "Diffusion Co-Policy for Synergistic Human-Robot Collaborative Tasks", 2023.☆18Aug 20, 2024Updated last year
- Codebase for the paper "To the Noise and Back: Diffusion for Shared Autonomy"☆38Nov 28, 2023Updated 2 years ago
- Implementation of modular composition network from https://arxiv.org/pdf/1711.11289.pdf☆25Dec 30, 2017Updated 8 years ago
- Implementation of NeurIPS 2018 paper "Meta-Gradient Reinforcement Learning"☆21Jul 19, 2022Updated 4 years ago
- ☆10Dec 15, 2024Updated last year
- MoveIt! config files for the Aldebaran NAO☆10Jan 20, 2017Updated 9 years ago
- solver for discrete Mixed Observable Markov Decision Processes☆11Oct 30, 2020Updated 5 years ago
- Robust Reinforcement Learning Benchmark☆13Sep 22, 2024Updated last year
- Residual policy learning☆83Apr 8, 2019Updated 7 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A simple example of randomized ensembled double q learning☆19Sep 3, 2021Updated 4 years ago
- PyTorch implementation of Memory Augmented Self-Play☆52Oct 26, 2020Updated 5 years ago
- ☆11Jan 13, 2026Updated 7 months ago
- ROS Stack for Nao Humanoid using the DCM (Device Communication Manager) Proxy☆11Mar 31, 2018Updated 8 years ago
- This repo gives an example of using a simple method of reinforcement learning to beat the Lunar Lander environment. The agent uses a comb…☆18Jul 27, 2018Updated 8 years ago
- [ICLR 2018] TensorFlow code for zero-shot visual imitation by self-supervised exploration☆204May 30, 2018Updated 8 years ago
- presentations☆44Dec 8, 2018Updated 7 years ago
- Modeling uncertainty information in deep learning☆22Jan 11, 2018Updated 8 years ago
- A library implementing the kernels for and experiments using extrinsic gauge equivariant vector field Gaussian Processes☆26Oct 28, 2021Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Infer how suboptimal agents are suboptimal while planning, for example if they are hyperbolic time discounters.☆25Sep 26, 2020Updated 5 years ago
- World Models applied to the Open AI Sonic Retro Contest☆78Jun 30, 2018Updated 8 years ago
- Code for making #GANterpretations☆23Nov 30, 2020Updated 5 years ago
- ☆17Jun 10, 2024Updated 2 years ago
- Implementation of ChoiceNet☆128Jul 31, 2018Updated 8 years ago
- (Engineering) Toward human-in-the-loop AI: Enhancing deep reinforcement learning via real-time human guidance for autonomous driving☆67Dec 1, 2022Updated 3 years ago
- 使用眼动仪收集数据,语义分割视频图片,查看用户注视点落在各个物体上的信息☆10Dec 11, 2019Updated 6 years ago