Pytorch code for Arxiv Paper: Learning to learn: Meta-Critic Networks for Sample-Efficient Learning
☆57Apr 3, 2018Updated 8 years ago
Alternatives and similar repositories for meta-critic-networks
Users that are interested in meta-critic-networks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reproduction of the paper "Soft Q-Learning with Mutual Information Regularization" CoRL 2019.☆10Jan 10, 2019Updated 7 years ago
- A simple RNN meta-learner☆10Dec 17, 2018Updated 7 years ago
- Demo for the subjective interface☆14Mar 4, 2018Updated 8 years ago
- A chainer implementation of Memory Augmented Neural Network☆48Jul 15, 2016Updated 10 years ago
- ☆11Oct 19, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pytorch implementation of LOLA (https://arxiv.org/abs/1709.04326) using DiCE (https://arxiv.org/abs/1802.05098)☆98Aug 21, 2018Updated 7 years ago
- Reproducing Policy Distillation (DeepMind paper ICLR 2016)☆22Feb 17, 2020Updated 6 years ago
- Source code of "Variational Imitation Learning with Diverse-quality Demonstrations" in ICML 2020. This github repository includes python …☆20Aug 16, 2021Updated 5 years ago
- StarCraft AI bot☆62Dec 4, 2018Updated 7 years ago
- Active Learning with Partial Feedback, ICLR 2019☆11Apr 27, 2020Updated 6 years ago
- Implementation of the paper "Adaptive Skip Intervals: Temporal Abstraction for Recurrent Dynamical Models"☆24Sep 7, 2018Updated 7 years ago
- Deep Reinforcement Learning by using Phasic Policy Gradient in Pytorch & Tensorflow☆20Oct 5, 2021Updated 4 years ago
- A PyTorch implementation of the blocks from the _A Simple Neural Attentive Meta-Learner_ paper☆99Apr 17, 2018Updated 8 years ago
- Soft Actor-Critic☆162Mar 13, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Mind-aware Multi-agent Management Reinforcement Learning☆80Mar 6, 2019Updated 7 years ago
- Exploring algorithms in the domain of offline reinforcement learning (REM, Ensemble-DQN, DQN, ...)☆17Jul 7, 2020Updated 6 years ago
- Code for "Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks"☆2,727Jan 19, 2020Updated 6 years ago
- ☆10May 5, 2021Updated 5 years ago
- ppo-lstm-parallel☆49Mar 26, 2019Updated 7 years ago
- Code for training policies based on paper Coordinated Multi-Agent Imitation Learning☆26Aug 7, 2017Updated 9 years ago
- This repository contains implementations of the paper, Bayesian Model-Agnostic Meta-Learning.☆20Jan 19, 2023Updated 3 years ago
- meta-learning research☆159Jul 31, 2021Updated 5 years ago
- Safe Policy Improvement with Baseline Bootstrapping☆26May 5, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This course introduced me to three cutting-edge technologies for privacy-preserving AI: Federated Learning, Differential Privacy, and Enc…☆11Sep 2, 2019Updated 6 years ago
- Proximal Policy Optimization with Stein Control Variates:☆34Feb 12, 2018Updated 8 years ago
- An implementation of deep reinforcement learning TD3 algorithm with prioritized experience replay (PER) buffer☆26Aug 14, 2019Updated 7 years ago
- Reinforcement Learning with Model-Agnostic Meta-Learning in Pytorch☆884Dec 27, 2022Updated 3 years ago
- I used this paper as inspiration https://arxiv.org/pdf/1904.03367.pdf☆35Mar 10, 2026Updated 5 months ago
- This is a program to solve NER with HMM. The principles and details can refer to my blog: https://blog.csdn.net/weixin_41679411/article/d…☆11Nov 20, 2018Updated 7 years ago
- Personal Repo to keep track of RL papers☆31May 3, 2021Updated 5 years ago
- Meta Learning / Learning to Learn / One Shot Learning / Few Shot Learning☆2,656Nov 26, 2018Updated 7 years ago
- OpenCL Inference Engine for pytorch☆51Jan 13, 2018Updated 8 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Deep Reinforcement Learning with Self-Play☆12Nov 27, 2019Updated 6 years ago
- [TPAMI'22] TransFuser: Imitation with Transformer-Based Sensor Fusion for Autonomous Driving, [CVPR'21] Multi-Modal Fusion Transformer fo…☆15Sep 14, 2022Updated 3 years ago
- Accompanying code for "Deep Reinforcement Learning that Matters"☆154Sep 22, 2017Updated 8 years ago
- Reproduction of "Model-Agnostic Meta-Learning" (MAML) and "Reptile".☆189Mar 21, 2019Updated 7 years ago
- This is the implementation of the CULP classification algorithm. The paper introducing this algorithm - `Classification Using Link Predic…☆13Jun 17, 2024Updated 2 years ago
- ByteCamp 2019 高并发高可用秒杀系统设计与实现 工程赛道三等奖(字节跳动夏令营入营 Top 150 in 6000+,Team Top 3 in 16 ,秒杀赛场 Top 1)☆16Nov 16, 2022Updated 3 years ago
- Scalable Multi-Agent Reinforcement Learning☆15Dec 25, 2021Updated 4 years ago