RainBow, Tensorflow
☆49Mar 28, 2018Updated 8 years ago
Alternatives and similar repositories for RainBow
Users that are interested in RainBow are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deep reinforcement learning baselines base on OpenAI. More algorithms are included, such as Rainbow: Combining Improvements in Deep Rei…☆35Aug 23, 2018Updated 8 years ago
- Rainbow: Combining Improvements in Deep Reinforcement Learning☆1,673Jan 13, 2022Updated 4 years ago
- Tensorflow implementation for "Noisy network for exploration"☆19Aug 2, 2017Updated 9 years ago
- 蚂蚁金服-用户精确定位比赛☆12Oct 7, 2017Updated 8 years ago
- Efficient Exploration through Bayesian Deep Q-Networks☆38Feb 14, 2018Updated 8 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆10May 29, 2018Updated 8 years ago
- Metropolis-Hastings GAN in Tensorflow for enhanced generator sampling☆19Mar 29, 2019Updated 7 years ago
- Research into controllers for 2d and 3d Active Ragdolls (using MujocoUnity+ml_agents)☆33Aug 21, 2018Updated 8 years ago
- Tensorflow Implementation for "Noisy network for exploration"☆32Jul 17, 2017Updated 9 years ago
- The lite edition of 微信跳一跳(JumpJump) developed by Unity with AI developed by ml-agents.☆33Jan 30, 2018Updated 8 years ago
- A TensorFlow implementation of DeepMind's A Distributional Perspective on Reinforcement Learning.(C51-DQN)☆57Aug 25, 2017Updated 9 years ago
- Some code for tutorials following https://gym.openai.com/docs/rl☆15Jul 3, 2016Updated 10 years ago
- treelite runtime binding in Rust☆12Jun 12, 2025Updated last year
- C51-DDQN in Keras☆126Nov 8, 2017Updated 8 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆13Mar 31, 2024Updated 2 years ago
- Hyperbolic SVM in Python☆12Jun 21, 2022Updated 4 years ago
- Code for the paper "A Boolean Task Algebra For Reinforcement Learning"☆10Dec 8, 2022Updated 3 years ago
- Ranking Policy Gradient☆23Nov 27, 2019Updated 6 years ago
- Deep Q Learning Neural Network Tekken 7 bot☆15Dec 26, 2017Updated 8 years ago
- ☆11Oct 10, 2023Updated 2 years ago
- Forex Prediction using Lstm☆12Nov 29, 2020Updated 5 years ago
- Learning To Stop While Learning To Predict☆36Nov 2, 2022Updated 3 years ago
- The source code of the paper 'Dynamic Knowledge Routing Network For Target-Guided Open-Domain Conversation'☆24Mar 24, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PyTorch implementation of the Munchausen Reinforcement Learning Algorithms M-DQN and M-IQN☆46Oct 4, 2020Updated 5 years ago
- ☆11Feb 22, 2019Updated 7 years ago
- ☆11Dec 1, 2020Updated 5 years ago
- Repository for codes of 'Deep Reinforcement Learning'☆218Oct 4, 2019Updated 6 years ago
- DQN implemented in keras with Dueling Network and Prioritized Experience Replay☆16Nov 21, 2018Updated 7 years ago
- Using multiple sensor modalities to improve exploration for robotic manipulation tasks with sparse rewards☆10Sep 17, 2019Updated 6 years ago
- In this project, I explore various machine learning techniques including Principal Component Analysis (PCA), Support Vector Machines (SVM…☆11Dec 5, 2022Updated 3 years ago
- An example on how Unity3D can interact with a python server either local (on the same machine) or remote (on different machines)☆11Dec 7, 2014Updated 11 years ago
- Reinforcement Learning Enhanced Quantum-inspired Algorithm for Combinatorial Optimization☆16Feb 19, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆10Feb 28, 2019Updated 7 years ago
- RLtime is a reinforcement learning library focused on state-of-the-art q-learning algorithms and features☆143Sep 23, 2019Updated 6 years ago
- A codebase for experimenting with various approaches to action priors.☆18Jul 14, 2018Updated 8 years ago
- [CoRL2020] Learning obstacle representations for neural motion planning☆29Dec 4, 2020Updated 5 years ago
- This code implements the decycling and dismantling procedures devoloped in Braunstein, Alfredo, Luca Dall'Asta, Guilhem Semerjian, and L…☆18Nov 15, 2016Updated 9 years ago
- c++ implementation of alphagozero☆15May 29, 2018Updated 8 years ago
- Probabilistic Streaming Tensor Decomposition @ ICDM'2018☆12Apr 22, 2019Updated 7 years ago