Some hard problems for reinforcement learning.
☆32Oct 5, 2018Updated 7 years ago
Alternatives and similar repositories for RL_acid
Users that are interested in RL_acid are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆11Nov 23, 2018Updated 7 years ago
- A collection of reading material for the Workshop on "Structure & Priors in Reinforcement Learning" (SPiRL) at ICLR 2019.☆13May 5, 2021Updated 5 years ago
- A 2 month Ego-vision Dataset with Autographer Wearable Camera and 2 users☆11Apr 28, 2020Updated 6 years ago
- Pytorch-based python library for continuous reinforcement learning and imitation learning [superseded by @osudrl/apex]☆13Mar 13, 2020Updated 6 years ago
- Reinforcement Learning via Latent State Decoding☆29Jun 12, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- CNN using C++ and CUDA☆16May 21, 2019Updated 7 years ago
- Leave No Trace is an algorithm for safe reinforcement learning.☆15Apr 30, 2018Updated 8 years ago
- ☆13Jul 9, 2018Updated 8 years ago
- ☆17May 16, 2018Updated 8 years ago
- Environments and Wrappers for CARLA, designed for ease of use with RL Tasks.☆23Oct 1, 2020Updated 5 years ago
- ☆58Mar 25, 2026Updated 4 months ago
- A 150-lines python code for Augmented Random Search (https://arxiv.org/abs/1803.07055) with numpy.☆71Dec 13, 2018Updated 7 years ago
- Pytorch Implementation of paper "Noisy Natural Gradient as Variational Inference"☆122Sep 1, 2018Updated 7 years ago
- Sample code for generative recurrent autoencoders.☆26Nov 12, 2016Updated 9 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Repository for code experimenting with RL and Solar Tracking☆13Apr 24, 2018Updated 8 years ago
- Implementation of robust adaptive control methods for the linear quadratic regulator☆36Dec 13, 2021Updated 4 years ago
- Docker containers of baseline agents for the Crafter environment☆30Dec 14, 2021Updated 4 years ago
- Modifiable OpenAI Gym environments for studying generalization in RL☆90Jan 22, 2019Updated 7 years ago
- PAC-Bayes generalization certificates for ICP☆22Nov 16, 2023Updated 2 years ago
- Experiments from "The Description Length of Deep Learning Models"☆10Aug 1, 2018Updated 8 years ago
- Code for our paper: Hierarchical RL Using an Ensemble of Proprioceptive Periodic Policies☆16Feb 21, 2019Updated 7 years ago
- Implementation of the LOSSGRAD optimization algorithm☆15Mar 21, 2019Updated 7 years ago
- Mac port of Torcs, The Open Racing Car Simulator☆11Jun 16, 2010Updated 16 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A3C style Option-Critic with deliberation cost☆40Jan 9, 2018Updated 8 years ago
- Dateset Reset Policy Optimization☆30Apr 12, 2024Updated 2 years ago
- Models built with TensorFlow☆26Dec 5, 2018Updated 7 years ago
- upgrade on pytorch seq2seq tutorial☆10Mar 11, 2019Updated 7 years ago
- Code for 'Contrastive Multi-Document Question Generation'☆11Oct 16, 2022Updated 3 years ago
- Implementation of Diversity Is All You Need (DIAYN) on top of Stable Baselines 3.☆13Jul 11, 2022Updated 4 years ago
- Reward Estimation for Variance Reduction in Deep Reinforcement Learning☆11May 8, 2018Updated 8 years ago
- Implementation of REBAR in PyTorch☆17Jul 18, 2018Updated 8 years ago
- LaTeX beamer template in corporate design of University of Amsterdam☆13Dec 7, 2015Updated 10 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Simple change of a3c to a2c☆15Jun 18, 2017Updated 9 years ago
- ☆10Feb 18, 2018Updated 8 years ago
- my public website☆12Jul 8, 2026Updated last month
- rllab's viskit with some added features☆73May 1, 2023Updated 3 years ago
- AI planners written in Python☆34Jun 2, 2021Updated 5 years ago
- ROS package providing Gazebo simulation of the Phantom X Hexapod robot.☆56Mar 26, 2021Updated 5 years ago
- Scalable MCTS for team scenarios☆17Jun 14, 2024Updated 2 years ago