A code reimplementation of DeepMind's "Multiagent Cooperation and Competition with Deep Reinforcement Learning" with Tensorflow
☆15Apr 27, 2018Updated 8 years ago
Alternatives and similar repositories for Tensorflow-DeepMind-Atari-Deep-Q-Learner-2Player
Users that are interested in Tensorflow-DeepMind-Atari-Deep-Q-Learner-2Player are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Markovian State and Action Abstractions for MDPs via Hierarchical MCTS within a POMDP Formulation☆11Jul 26, 2016Updated 10 years ago
- ☆10Nov 27, 2019Updated 6 years ago
- ☆13Apr 3, 2019Updated 7 years ago
- Research project - real-time multi-agent pursuit a moving target☆17Mar 13, 2021Updated 5 years ago
- Collaborative Deep Reinforcement Learning☆32Jul 29, 2017Updated 9 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Tensorflow implementation of DQN to control cart-pole from OpenAI gym environment☆14Sep 24, 2017Updated 8 years ago
- Faithful Python implementation of the paper "Towards Deep Symbolic Reinforcement Learning" by Garnelo et al.☆13Mar 23, 2021Updated 5 years ago
- Simulator for evaluating cloud/edge requests from connected vehicles and computing statistical analysis of the input network☆12May 7, 2018Updated 8 years ago
- Exploration by Random Network Distillation☆15Dec 30, 2018Updated 7 years ago
- Reinforcement Task Scheduling Project☆16Jun 3, 2019Updated 7 years ago
- 2019华为杯 第十六届研究生数学建模F题解决方案“基于最小代价与Dubins曲线的改进Dijkstra与A*算法航迹规划“源代码☆16Jan 16, 2020Updated 6 years ago
- Model-Free-Episodic-Control implementation.☆18Jun 3, 2019Updated 7 years ago
- Multiagent reinforcement learning simulation framework - Undergraduate thesis in Mechatronics Engineering at the University of Brasília☆69Sep 30, 2018Updated 7 years ago
- ☆30Aug 20, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Attentional Mechanism incorporated in Asynchronous Advantage Actor Critic a3c/a2c deep mind☆10Jan 9, 2018Updated 8 years ago
- ☆10Jul 1, 2019Updated 7 years ago
- Distributed DRL by Ray and TensorFlow Tutorial.☆10Dec 26, 2019Updated 6 years ago
- This is a Pytorch implementation of "Deep Low-Rank Subspace Clustering" (CVPRW 2020).☆15Sep 18, 2020Updated 5 years ago
- 面试必备基础知识☆12Mar 21, 2019Updated 7 years ago
- I added selfplay functionality to openai gyms☆10Jan 16, 2021Updated 5 years ago
- COSE: Configuring Serverless Functions using Statistical Learning☆10Jun 28, 2023Updated 3 years ago
- Option Critic with subgoal discovery by spectral decomposition of the Successor Features Matrix or clustering in Successor features space…☆24Nov 29, 2018Updated 7 years ago
- ☆17Mar 2, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 百度UIE抽 取模型torch版训练预测框架☆12Nov 20, 2024Updated last year
- simple keras implement for 《Memory Fusion Network for Multi-view Sequential Learning》☆14Apr 9, 2021Updated 5 years ago
- suPER is a collaborative multi-agent RL algorithm☆14Jun 11, 2024Updated 2 years ago
- Continuous Energy Minimization for Multitarget Tracking☆20Feb 9, 2022Updated 4 years ago
- ☆11Apr 12, 2020Updated 6 years ago
- a q-learning algorithms on packet routing.☆14Dec 1, 2018Updated 7 years ago
- Transmit Power Control Using Deep Neural Network for Underlay Device-to-Device Communication☆19Jan 8, 2019Updated 7 years ago
- Negative Update Intervals in Multi-Agent Deep Reinforcement Learning☆35May 14, 2019Updated 7 years ago
- ☆15Feb 8, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Java tool to translate VRP instances to VRP-REP unified format.☆11Nov 28, 2014Updated 11 years ago
- some Multiagent enviroment in 《Multi-agent Reinforcement Learning in Sequential Social Dilemmas》 and 《Value-Decomposition Networks For Co…☆131Jan 13, 2023Updated 3 years ago
- ☆16Dec 13, 2022Updated 3 years ago
- Code for the results of the Paper:☆21May 17, 2018Updated 8 years ago
- ☆18Aug 19, 2024Updated 2 years ago
- A basic face-swap implementation using OpenCV and dlib.☆17Feb 15, 2019Updated 7 years ago
- Code exploring the use of reward machines in the context of cooperative multi-agent reinforcement learning.☆14Apr 29, 2023Updated 3 years ago