The purpose of this project is to research Artificial Intelligence and Reinforcement Learning. In the AI Arena, multiple agents can interact with a single environment. After sending its action, each each agent will receive a reward. This allows agents to learn, improve their behavior and to adapt to each other. Interesting phenomena can arise..…
☆39Oct 31, 2017Updated 8 years ago
Alternatives and similar repositories for AI_Arena
Users that are interested in AI_Arena are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Feasible target propagation code for the paper "Deep Learning as a Mixed Convex-Combinatorial Optimization Problem" by Friesen & Domingos…☆28Apr 12, 2018Updated 8 years ago
- IPython Magic Functions☆16Aug 14, 2017Updated 9 years ago
- CS 294: Deep Reinforcement Learning, Spring 2017 Berkeley☆11Feb 19, 2017Updated 9 years ago
- [INACTIVE] A bunch of articficial intelligence algorithms☆11May 14, 2016Updated 10 years ago
- An experiment with Thompson sampling and TD(0) on a grid world variant☆17Nov 8, 2013Updated 12 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- The code for the tutorial of the GrahpEdit and GraphNode nodes.☆16Mar 10, 2019Updated 7 years ago
- Introduction to Reinforcement Learning in Python☆13Oct 17, 2018Updated 7 years ago
- Code for ICLR 2019 paper "Efficient Augmentation via Data Subsampling"☆15Feb 20, 2019Updated 7 years ago
- Implementation of Unscented Fast SLAM algorithm for Applied Estimation (EL2320) - KTH☆10Jan 28, 2019Updated 7 years ago
- DataBright: Towards a Global Exchange for Decentralized Data Ownership and Trusted Computation☆13Jun 28, 2018Updated 8 years ago
- Model-Free Episodic Control☆14Jan 12, 2017Updated 9 years ago
- A minimalist SQL migration CLI for PostgreSQL.☆13Sep 8, 2025Updated last year
- This is a sample implementation of "TIMERS: Error-Bounded SVD Restart on Dynamic Networks"(AAAI 2018).☆12Jul 4, 2018Updated 8 years ago
- coding examples to Intro to RL☆13Apr 30, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Matlab code for learning doubly sparse dictionary on synthetic data. Details can be found in the paper "A Provable Approach for Double-Sp…☆11Mar 5, 2018Updated 8 years ago
- a motion detector for video; written with OpenCV☆12Nov 3, 2022Updated 3 years ago
- Contains an implementation of "Imitation Learning via Kernel Mean Embedding (2018, AAAI)"☆11Oct 2, 2018Updated 7 years ago
- growing interpretable part graphs on convnets via multi-shot learning, in AAAI 2017☆15May 28, 2017Updated 9 years ago
- Tools to help execute perfect maneuvers☆19May 25, 2021Updated 5 years ago
- This repository contains implementations of the paper VUSFA☆14Mar 31, 2021Updated 5 years ago
- Reproducing the reinforcement learning models used in "Emergence of Linguistic Communication from Referential Games with Symbolic and Pix…☆12Jun 23, 2018Updated 8 years ago
- we propose a novel and efficient cross-domain human parsing model to bridge the cross-domain differences in terms of visual appearance an…☆15Jan 9, 2018Updated 8 years ago
- Code Repo for paper Label Leakage and Protection in Two-party Split Learning (ICLR 2022).☆22Mar 12, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Comparison between Sarsa and Q-Learning algorithms on risk handling☆17Jul 10, 2017Updated 9 years ago
- Ancient two-player strategy race board game☆14Mar 19, 2024Updated 2 years ago
- Source code for the following paper(arXiv link): Improved Actor Relation Graph based Group Activity Recognition Zijian Kuang, Xinran Tie☆15Jan 19, 2022Updated 4 years ago
- Code for "Boosted Generative Models", AAAI 2018.☆20Dec 26, 2017Updated 8 years ago
- An easy to use evaluator for object detection performance metrics, such as mAP and AR☆17Jan 17, 2022Updated 4 years ago
- Notes for Deep Learning Papers☆19Sep 21, 2018Updated 8 years ago
- E2E ASR system☆14Oct 20, 2022Updated 3 years ago
- One page descriptions of board games☆11Jan 1, 2026Updated 8 months ago
- Investigations into simplified holdem poker☆12Oct 17, 2012Updated 13 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Proportional-Derivative Neural Networks, as described in Temporally Efficient Deep Learning with Spikes☆16Sep 5, 2017Updated 9 years ago
- Spatially explicit plant growth simulation☆11Sep 15, 2026Updated last week
- An implementation of AlphaZero for the board game Tak☆14Feb 11, 2023Updated 3 years ago
- ☆18Feb 22, 2018Updated 8 years ago
- ☆18Mar 26, 2022Updated 4 years ago
- Interactive documentation and programming with Scala, iPython notebook style.☆19Mar 9, 2016Updated 10 years ago
- HTML5 game made with Phaser☆11Jun 23, 2016Updated 10 years ago