Repository for Iterated Relearning: The Impact of Non-stationarity on Generalisation in Deep Reinforcement Learning
☆11Jun 8, 2020Updated 6 years ago
Alternatives and similar repositories for rl-iter
Users that are interested in rl-iter are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code to reproduce the NeurIPS 2019 paper "Generalization in Reinforcement Learning with Selective Noise Injection and Information Bottlen…☆52Jun 28, 2020Updated 6 years ago
- Implement Categorical Variational autoencoder using Pytorch☆15Apr 25, 2018Updated 8 years ago
- Tabula Rasa Tic-Tac-Toe☆10Jan 3, 2019Updated 7 years ago
- ☆16Aug 2, 2022Updated 4 years ago
- My wedding invitation☆12May 3, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Explore and Control with Adversarial Surprise☆10Jul 20, 2021Updated 5 years ago
- 经典BP神经网络预测上证指数(中国股市)☆14Mar 31, 2023Updated 3 years ago
- Pytorch implementation on OpenAI's Procgen ppo-baseline, built from scratch.☆31Sep 10, 2020Updated 5 years ago
- The open source of FeverBasketball environment for research purpose.☆11Mar 2, 2020Updated 6 years ago
- ☆11Jul 15, 2022Updated 4 years ago
- This repo contains the code of "Structure-Aware Transformer Policy for Inhomogeneous Multi-Task Reinforcement Learning".☆14May 20, 2022Updated 4 years ago
- Code for the paper Alpha Zero in Continuous Action Space (A0C) (https://arxiv.org/pdf/1805.09613.pdf)☆15Jan 19, 2021Updated 5 years ago
- Prioritized Sequence Experience Replay☆10Aug 16, 2021Updated 4 years ago
- Generalizable Implicit Hate Speech Detection using Contrastive Learning (COLING 2022)☆14Oct 9, 2022Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- 🟣 Explainable Ai interview questions and answers to help you prepare for your next machine learning and data science interview in 2026.☆12Jan 4, 2026Updated 7 months ago
- Neural Fictitious Self-Play in Leduc Holdem☆11Jul 4, 2018Updated 8 years ago
- Julia Implementation of the POMCP algorithm for solving POMDPs☆12Aug 6, 2021Updated 5 years ago
- ☆16May 11, 2023Updated 3 years ago
- ☆15Mar 25, 2018Updated 8 years ago
- My Submission for the OpenAI/NeurIPS ProcGen Competition☆11Nov 12, 2020Updated 5 years ago
- Adaptive LSTM from Breaking the Activation Function Bottleneck Through Adaptive Parameterization (https://arxiv.org/abs/1805.08574)☆27Apr 30, 2020Updated 6 years ago
- Automatic Data-Regularized Actor-Critic (Auto-DrAC)☆104Mar 24, 2023Updated 3 years ago
- Official codebase for Improving Computational Efficiency in Visual Reinforcement Learning via Stored Embeddings.☆21Mar 5, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ClusterGAN PyTorch implementation☆12Feb 24, 2020Updated 6 years ago
- My Homepage☆10Jun 26, 2026Updated last month
- Damn Vulnerable Chemical Process - Vinyl Acetate Monomer☆23Dec 13, 2015Updated 10 years ago
- Repository for ML Reproducibility Challenge 2020 for the Neurips paper, "The Value Equivalence Principle for Model-Based Reinforcement Le…☆18Apr 13, 2021Updated 5 years ago
- Code repository for the research project "You Play Ball, I Play Ball: Bayesian Multi-Agent Reinforcement Learning for Slime Volleyball", …☆17Nov 15, 2020Updated 5 years ago
- ☆26Apr 16, 2024Updated 2 years ago
- (ICML 2023) Feature learning in deep classifiers through Intermediate Neural Collapse: Accompanying code☆16Jul 27, 2023Updated 3 years ago
- super fast cpp implementation of longest common subsequence/substring☆23Oct 25, 2023Updated 2 years ago
- Novel Reinforcement Learning method for tackling goal-oriented robotics tasks with obstacles.☆40Mar 14, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Benchmarks for AutoAlbument - AutoML for Image Augmentation☆10Nov 5, 2023Updated 2 years ago
- Implementations of Curious Replay for model-based adaptation.☆43Jul 5, 2023Updated 3 years ago
- A PyTorch implementation of SEED, originally created by Google Research for TensorFlow 2.☆14Dec 8, 2020Updated 5 years ago
- Implementation of "Variational Inference for Monte Carlo Objectives"☆21Jul 27, 2020Updated 6 years ago
- Deep Variational Reinforcement Learning☆141Jun 21, 2022Updated 4 years ago
- Order Fulfillment by Multi-Agent Reinforcement Learning☆28Jun 12, 2026Updated 2 months ago
- Bayes-Adaptive Monte-Carlo Planning algorithm☆19Mar 5, 2013Updated 13 years ago