TensorFlow implementation of asynchronous advantage actor-critic (A3C)
☆38Oct 20, 2021Updated 4 years ago
Alternatives and similar repositories for ocd-a3c
Users that are interested in ocd-a3c are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Jun 23, 2017Updated 9 years ago
- Tensorflow implementation of Asynchronous Advantage Actor Critic (A3C) from "Asynchronous Methods for Deep Reinforcement Learning".☆24Apr 20, 2017Updated 9 years ago
- A gym environment for Stuart Armstrong's model of a treacherous turn.☆18Jul 28, 2018Updated 8 years ago
- Keras implementation of guide actor-critic for continuous control☆11Mar 12, 2018Updated 8 years ago
- Models and training scripts for the English, German and Russian MAGEC systems described in R. Grundkiewicz, M. Junczys-Dowmunt: Minimally…☆12Jul 7, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A project copied from google-research which named motion-imitation was rewrited with PyTorch☆10Sep 30, 2022Updated 3 years ago
- A simple wrapper to analyse and visualise reinforcement learning agents' behaviour in the environment.☆14Jan 8, 2022Updated 4 years ago
- Reading notes & PyTorch experiments on OpenAI's "Spinning Up in DRL" tutorial.☆40Dec 8, 2022Updated 3 years ago
- My reproduction of various reinforcement learning algorithms (DQN variants, A3C, DPPO, RND with PPO) in Tensorflow.☆37Mar 24, 2023Updated 3 years ago
- PyTorch implementation of Munchausen Reinforcement Learning based on DQN and SAC. Handles discrete and continuous action spaces☆15Oct 3, 2021Updated 4 years ago
- ☆13May 29, 2018Updated 8 years ago
- A video labeling platform for training classification algorithms.☆15Mar 30, 2021Updated 5 years ago
- ☆18Oct 12, 2014Updated 11 years ago
- ASLA arecord (http://alsa.opensrc.org/Arecord/) wrapper for Node.js.☆10Jan 27, 2016Updated 10 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Deep Reinforcement Learning Algorithms Implementation in PyTorch☆27Feb 11, 2025Updated last year
- A C++ implementation of the asynchronous advantage actor-critic (A3C) algorithm☆23Mar 17, 2020Updated 6 years ago
- PyTorch Implementation of the Maximum a Posteriori Policy Optimisation☆84Nov 19, 2022Updated 3 years ago
- Implementation for ACER in tensorflow and sonnet by deepmind☆11Aug 28, 2017Updated 9 years ago
- Experiment utility code, specifically designed for use with Compute Canada.☆11Jan 27, 2025Updated last year
- Port of pybullet envs to gymnasium☆18Mar 4, 2025Updated last year
- Lightweight interface to AWS☆47Oct 15, 2019Updated 6 years ago
- [AAMAS 2023] Code for the paper "Automatic Noise Filtering with Dynamic Sparse Training in Deep Reinforcement Learning"☆12Feb 22, 2024Updated 2 years ago
- RL-Toolkit: A Research Framework for Robotics☆21Jan 22, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Cross-Platform object detection using TensorFlow Lite and OpenCV in C++☆18Apr 26, 2020Updated 6 years ago
- Example Code for the Conditional Action Trees Paper☆13May 24, 2021Updated 5 years ago
- Implementation of the two-step-task as described in "Prefrontal cortex as a meta-reinforcement learning system" and "Learning to Reinforc…☆60Mar 28, 2019Updated 7 years ago
- Source code for "A deep dive into reinforcement learning"☆13Dec 17, 2019Updated 6 years ago
- Code for the Black-DROPS algorithm: "Black-Box Data-efficient Policy Search for Robotics", IROS 2017/ICRA 2018☆65Nov 17, 2021Updated 4 years ago
- sign elf binaries with GPG☆17Oct 10, 2016Updated 9 years ago
- Matrix exponential in cuda for pytorch and tensorflow☆17Nov 26, 2018Updated 7 years ago
- Imagination Augmented Agents in TensorFlow☆20Oct 21, 2018Updated 7 years ago
- Part-of-speech tagger implemented using a feedforward network in TensorFlow☆14Jan 15, 2018Updated 8 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- GOAT(山羊)是中英文大语言模型,基于LlaMa进行SFT。☆12Apr 24, 2023Updated 3 years ago
- [AutoML'22] Bayesian Generational Population-based Training (BG-PBT)☆31Sep 16, 2022Updated 3 years ago
- my public website☆12Jul 8, 2026Updated last month
- Theano implementation of T1-T2 gradient-based method for tuning continuous hyperparameters.☆10Jun 20, 2016Updated 10 years ago
- This repository contains my implementation of a shape-constrained network which predicts up to 170 FPS☆12Feb 12, 2019Updated 7 years ago
- Paper notes for my PhD on Machine Learning (mostly focused on Reinforcement Learning)☆17Jul 22, 2019Updated 7 years ago
- FQF(Fully parameterized Quantile Function for distributional reinforcement learning) is a general reinforcement learning framework for At…☆48Sep 26, 2020Updated 5 years ago