Applying minimaxQ learning algorithm to 2 agents games
☆34Nov 27, 2017Updated 8 years ago
Alternatives and similar repositories for MinimaxQ-Learning
Users that are interested in MinimaxQ-Learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A fighter fly out trajectory time series data mining demo, I use agnes and k-means to clustering the flyout data samples into left, strai…☆13Aug 12, 2017Updated 8 years ago
- GAUSS EU project: Unmanned aerial vehicle Traffic Management (UTM) software development☆16Feb 7, 2022Updated 4 years ago
- Testing different RL algorithms for multi-agent environments. From SARSA, QLearning to Independent Q-Learning, Joint Action Learning and …☆12Mar 29, 2019Updated 7 years ago
- ☆12Mar 21, 2024Updated 2 years ago
- GPU Implementation of the STOMP algorithm for computing the matrix profile☆22Sep 27, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆12Nov 8, 2022Updated 3 years ago
- Agar.io OpenAI Gym Learning Environment☆12Sep 10, 2023Updated 2 years ago
- Application of Deep Reinforcement Learning to Supply Chain management. Reference: https://blog.griddynamics.com/deep-reinforcement-learni…☆12Jul 21, 2021Updated 5 years ago
- Code for our NeurIPS 2020 paper Improving Generalization in Reinforcement Learning with Mixture Regularization☆34Oct 22, 2020Updated 5 years ago
- pytorch implementation of DQN, NAF, DDPG☆13Jun 7, 2018Updated 8 years ago
- General solver for Hamilton-Jacobi-Bellman equations☆18Apr 23, 2017Updated 9 years ago
- Code for the paper "Functional Regularization for Reinforcement Learning via Learned Fourier Features"☆20Oct 2, 2022Updated 3 years ago
- Learning Multiaspect Traffic Couplings by Multirelational Graph Attention Networks for Traffic Prediction☆13Oct 7, 2022Updated 3 years ago
- convert triforce nand images to iso files☆15Sep 13, 2017Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation of MCTS algorithms in Munos (2014)☆13Aug 8, 2018Updated 7 years ago
- pytorch, noisy_distributional_double_dueling_PER_RNN_CNN...CartPole-v1 , Acrobot-v1, MountainCar-v0☆14Mar 19, 2018Updated 8 years ago
- Link to paper: https://www.ssrn.com/abstract=3804655☆14Jul 27, 2021Updated 5 years ago
- JAX implementations of various deep reinforcement learning algorithms.☆25Feb 2, 2025Updated last year
- ☆19Jul 12, 2026Updated 2 weeks ago
- Stationary distributions for arbitrary finite state Markov processes, including specializations for the Moran, Wright-Fisher, and other …☆22Aug 10, 2018Updated 7 years ago
- A python client library for microRTS.☆20Feb 5, 2020Updated 6 years ago
- Fictitious Self-play & Reinforcement Learning☆18Jan 26, 2018Updated 8 years ago
- Useful tools for the Baxter Research Robot☆10Mar 3, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The Official Implementation of Domain Adaptive Imitation Learning (DAIL)☆25Oct 26, 2020Updated 5 years ago
- Robotics in Python☆13Feb 21, 2023Updated 3 years ago
- Code to reproduce the NeurIPS 2019 paper "Generalization in Reinforcement Learning with Selective Noise Injection and Information Bottlen…☆52Jun 28, 2020Updated 6 years ago
- Learning from Guided Play: A Scheduled Hierarchical Approach for Improving Exploration in Adversarial Imitation Learning Source Code☆17Aug 23, 2024Updated last year
- Curiosity based exploration and playing in RL with Gym Robotics envs.☆12Sep 25, 2018Updated 7 years ago
- A Python implementation of the matrix profile algorithm☆16Feb 2, 2020Updated 6 years ago
- Classify time series data using motifs discovered from Sequitur processing of SAX discretized data.☆12Apr 20, 2017Updated 9 years ago
- Network Randomization: A Simple Technique for Generalization in Deep Reinforcement Learning / ICLR 2020☆57Apr 27, 2020Updated 6 years ago
- [IJCAI'23] Semantic-aware Generation of Multi-view Portrait Drawings (SAGE)☆10Feb 25, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The official implementation of DropGNN: Random Dropouts Increase the Expressiveness of Graph Neural Networks (NeurIPS 2021)☆26Jun 26, 2022Updated 4 years ago
- Implementation of DDPG (Modified from the work of Patrick Emami) - Tensorflow (no TFLearn dependency), Ornstein Uhlenbeck noise function,…☆64Apr 27, 2017Updated 9 years ago
- ☆25Aug 25, 2021Updated 4 years ago
- Supporting code for the paper "Portuguese Language Models and Word Embeddings: Evaluating on Semantic Similarity Tasks".☆11Dec 8, 2022Updated 3 years ago
- OpenAI Gym compatible reinforcement learning environment for Space Fortress https://arxiv.org/abs/1809.02206☆11Aug 30, 2024Updated last year
- HTML5 canvas based image editor for the web and ChromeOS☆12Apr 8, 2020Updated 6 years ago
- [NeurIPS 2024] Unsupervised Hierarchy-Agnostic Segmentation: Parsing Semantic Image Structure☆12Nov 27, 2025Updated 8 months ago