TensorFlow implementation of "Playing hard exploration games by watching YouTube"
☆39Sep 15, 2019Updated 7 years ago
Alternatives and similar repositories for HardRLWithYoutube
Users that are interested in HardRLWithYoutube are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [WIP] Playing Hard Exploration Games by Watching YouTube (Aytar et al., 2018)☆12Jan 31, 2019Updated 7 years ago
- Trajectory-wise Multiple Choice Learning for Dynamics Generalization in Reinforcement Learning (NeurIPS 2020)☆38Oct 27, 2020Updated 5 years ago
- NYU GSAS PhD thesis template☆11May 14, 2020Updated 6 years ago
- ☆13Mar 26, 2019Updated 7 years ago
- MicroPython STM Read Protection Module☆11Nov 5, 2014Updated 11 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Deep reinforcement learning baselines base on OpenAI. More algorithms are included, such as Rainbow: Combining Improvements in Deep Rei…☆35Aug 23, 2018Updated 8 years ago
- Code accompanying "Learning What To Do by Simulating the Past", ICLR 2021.☆27May 4, 2021Updated 5 years ago
- 3rd placed submission to the NeurIPS MineRL competition 2019☆10Mar 24, 2023Updated 3 years ago
- Decoupling Dynamics and Reward for Transfer Learning☆16Sep 7, 2018Updated 8 years ago
- Visualizing the learned space-time attention using Attention Rollout☆42Apr 1, 2022Updated 4 years ago
- ☆10Jul 20, 2023Updated 3 years ago
- This is an implementation of Deep Q Learning (DQN) playing Breakout from OpenAI's gym with Keras.☆29Feb 7, 2018Updated 8 years ago
- [IJCAI'20][ICLR'19 Workshop] Flow-based Intrinsic Curiosity Module. Playing SuperMario with RL agent and FICM!☆105Dec 8, 2022Updated 3 years ago
- Code for the Reset-free Trial and Error learning paper (RTE) experiments☆10Jan 3, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The repository is for Reinforcement-Learning Uncertainty research, in which we investigate various uncertain factors in RL.☆23Jun 16, 2023Updated 3 years ago
- Experiments to train transformer network to master reinforcement learning environments.☆32Mar 14, 2021Updated 5 years ago
- ☆42Oct 30, 2021Updated 4 years ago
- pix2pix and Cycle GAN architectures for image style transfer☆13May 27, 2021Updated 5 years ago
- Code companion of Multi-task Learning for Aggregated Data using Gaussian Processes paper☆11Apr 6, 2020Updated 6 years ago
- On the Feasibility of Cross-Task Transfer with Model-Based Reinforcement Learning☆16Apr 30, 2023Updated 3 years ago
- Controllability-Aware Unsupervised Skill Discovery (ICML 2023)☆30Jun 3, 2023Updated 3 years ago
- Companion code to CoRL 2018 paper: E Bıyık, D Sadigh. "Batch Active Preference-Based Learning of Reward Functions". Conference on Robot L…☆30May 29, 2019Updated 7 years ago
- Public Repo for the paper "Overcoming The Spectral-Bias of Neural Value Approximation"☆11May 25, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- An attempt to reverse engineer custom file formats used by the game Outlaws from LucasArts.☆16Aug 23, 2026Updated last month
- ☆12Aug 28, 2020Updated 6 years ago
- Python Library for Dynamic Movement Primitives with Reinforcement Learning☆14Jun 21, 2022Updated 4 years ago
- Integrating opencv with mujoco.☆12Mar 25, 2025Updated last year
- Exploration by Random Network Distillation☆15Dec 30, 2018Updated 7 years ago
- Introduction to Gaussian Processes☆11Jan 13, 2024Updated 2 years ago
- [ICLR 2023] Choreographer: a world-model-based agent that discovers and learns unsupervised skills in latent imagination, and it's able t…☆43Jun 18, 2024Updated 2 years ago
- PyTorch implementation of both discrete and continuous ACER☆25Jan 27, 2019Updated 7 years ago
- ☆39Aug 25, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- MaxSum is an algorithm about Distributed Constraint Optimization Problems (DCOPs)☆11Jan 15, 2018Updated 8 years ago
- Linear Algebra for Machine Learning Book Exercises☆13May 19, 2019Updated 7 years ago
- Source code for "Multi-objective Model-based Policy Search for Data-efficient Learning with Sparse Rewards" (CoRL 2018)☆13Oct 8, 2018Updated 7 years ago
- Hands-On Reinforcement Learning with TensorFlow & TRFL☆14Jan 18, 2021Updated 5 years ago
- Dream to Control: Learning Behaviors by Latent Imagination☆624Sep 10, 2021Updated 5 years ago
- Adding Dreamer-v3's implementation tricks to CleanRL's PPO☆17May 19, 2023Updated 3 years ago
- Submission code of UEFDRL team to NeurIPS 2019 MineRL challenge (5th place)☆13Nov 13, 2020Updated 5 years ago