Pytorch implementation of Randomized Ensembled Double Q-learning (REDQ)
☆21Mar 12, 2021Updated 5 years ago
Alternatives and similar repositories for Randomized-Ensembled-Double-Q-learning-REDQ-
Users that are interested in Randomized-Ensembled-Double-Q-learning-REDQ- are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Author's PyTorch implementation of Randomized Ensembled Double Q-Learning (REDQ) algorithm.☆188Nov 14, 2024Updated last year
- Implementation of the Self Paced Reinforcement Learning Experiments☆19Sep 27, 2023Updated 3 years ago
- Code for the paper "D2RL: Deep Dense Architectures for Reinforcement Learning"☆40Jan 22, 2021Updated 5 years ago
- 2022 WSDM 爱奇艺用户留存预测赛 第三名方案☆14Jan 29, 2022Updated 4 years ago
- ur5 robot with robotiq parallel grippers for testing parallel grasping algorithms☆11Apr 10, 2016Updated 10 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Multi-objective reinforcement learning for covid-19 control☆12Aug 12, 2021Updated 5 years ago
- ☆10Nov 4, 2019Updated 6 years ago
- On the model-based stochastic value gradient for continuous reinforcement learning☆58Mar 6, 2026Updated 6 months ago
- ☆12Aug 15, 2020Updated 6 years ago
- Knock your images before you get stressed.☆11Aug 5, 2026Updated last month
- ☆12Sep 30, 2017Updated 8 years ago
- Machine learning to predict future number Covid19 Daily Cases (7-day moving average). Long Short Term Memory (LSTM) Predictor and Reinfor…☆14Feb 21, 2021Updated 5 years ago
- ☆12Jun 8, 2020Updated 6 years ago
- Machine Learning Tool to Forecast Grid Frequency and Scheduled Power Generation of Thermal Power Plant.☆17Nov 1, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Jun 23, 2017Updated 9 years ago
- Implement IMPALA architecture from Distributed Deep-RL Paper.☆15Oct 18, 2018Updated 7 years ago
- 🔍 Codebase for the ICML '20 paper "Ready Policy One: World Building Through Active Learning" (arxiv: 2002.02693)☆18Jul 6, 2023Updated 3 years ago
- ☆78Mar 15, 2021Updated 5 years ago
- A meta-population model for COVID19 in China☆11Jun 10, 2020Updated 6 years ago
- ☆13Jan 5, 2021Updated 5 years ago
- Benchmarking framework for Feature Selection and Feature Ranking algorithms 🚀☆19Apr 12, 2023Updated 3 years ago
- PGQ is an approach to combine Policy Gradient and Q-Learning. This repository will contain an implementation of PGQ.☆15Mar 9, 2017Updated 9 years ago
- Multi Agent Reinforcement Learning Environment For Aerial Unmanned Vehicles☆13Apr 13, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Deep Reinforcement Learning Agent to control traffic light providing emergency facilitation using real-time traffic data.☆10Feb 11, 2021Updated 5 years ago
- Co-Adaptation of Algorithmic and Implementational Innovations in Inference-based Deep Reinforcement Learning (NeurIPS2021)☆20Oct 25, 2021Updated 4 years ago
- Paper Collection of Reinforcement Learning Exploration covers Exploration of Muti-Arm-Bandit, Reinforcement Learning and Multi-agent Rein…☆37Nov 8, 2019Updated 6 years ago
- PyTorch implementation of Episodic Meta Reinforcement Learning on variants of the "Two-Step" task. Reproduces the results found in three …☆38Dec 12, 2020Updated 5 years ago
- Code for our paper "Active Perception using Light Curtains for Autonomous Driving", ECCV 2020☆10Dec 7, 2021Updated 4 years ago
- Official code for ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning (AAAI'24)☆16Feb 10, 2024Updated 2 years ago
- ☆15Apr 24, 2020Updated 6 years ago
- ☆118Apr 28, 2023Updated 3 years ago
- PyTorch implementation of MATD3☆13Apr 3, 2020Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Models built with TensorFlow☆26Dec 5, 2018Updated 7 years ago
- some RL algorithms☆19Dec 9, 2016Updated 9 years ago
- DIGIX2021 基于多目标优化的视频推荐 亚军方案☆26Oct 7, 2021Updated 4 years ago
- RL framework for embodied agents based on PyTorch☆11Apr 11, 2019Updated 7 years ago
- Code for the paper EpidemiOptim: A Toolbox for the Optimization of Control Policies in Epidemiological Models.☆16Mar 1, 2022Updated 4 years ago
- ☆20Oct 27, 2025Updated 11 months ago
- Native macOS control panel for Clawdbot☆15Mar 12, 2026Updated 6 months ago