TensorFlow implementation of asynchronous advantage actor-critic (A3C)
☆38Oct 20, 2021Updated 4 years ago
Alternatives and similar repositories for ocd-a3c
Users that are interested in ocd-a3c are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Jun 23, 2017Updated 9 years ago
- Implementation of Diversity Is All You Need (DIAYN) on top of Stable Baselines 3.☆13Jul 11, 2022Updated 4 years ago
- Keras implementation of guide actor-critic for continuous control☆11Mar 12, 2018Updated 8 years ago
- Models and training scripts for the English, German and Russian MAGEC systems described in R. Grundkiewicz, M. Junczys-Dowmunt: Minimally…☆12Jul 7, 2021Updated 5 years ago
- A project copied from google-research which named motion-imitation was rewrited with PyTorch☆10Sep 30, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Reading notes & PyTorch experiments on OpenAI's "Spinning Up in DRL" tutorial.☆40Dec 8, 2022Updated 3 years ago
- A simple wrapper to analyse and visualise reinforcement learning agents' behaviour in the environment.☆15Jan 8, 2022Updated 4 years ago
- Algorithms for Gradient TD updates☆19Feb 21, 2026Updated 6 months ago
- Deep Reinforcement Learning Algorithms Implementation in PyTorch☆27Feb 11, 2025Updated last year
- 패스트캠퍼스 텍스트마이닝을 위한 머신러닝 실습 자료실☆19Feb 21, 2020Updated 6 years ago
- A C++ implementation of the asynchronous advantage actor-critic (A3C) algorithm☆23Mar 17, 2020Updated 6 years ago
- Implementation for ACER in tensorflow and sonnet by deepmind☆11Aug 28, 2017Updated 9 years ago
- Experiment utility code, specifically designed for use with Compute Canada.☆11Jan 27, 2025Updated last year
- Port of pybullet envs to gymnasium☆18Mar 4, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Cooperation and Fairness in Multi-Agent Reinforcement Learning☆16Aug 6, 2025Updated last year
- Lightweight interface to AWS☆47Oct 15, 2019Updated 6 years ago
- OpenAI Gym environment for DART robotics simulator.☆22Apr 17, 2018Updated 8 years ago
- ☆17Mar 31, 2022Updated 4 years ago
- Implementation of the two-step-task as described in "Prefrontal cortex as a meta-reinforcement learning system" and "Learning to Reinforc…☆60Mar 28, 2019Updated 7 years ago
- Code for the Black-DROPS algorithm: "Black-Box Data-efficient Policy Search for Robotics", IROS 2017/ICRA 2018☆65Nov 17, 2021Updated 4 years ago
- Imagination Augmented Agents in TensorFlow☆20Oct 21, 2018Updated 7 years ago
- ☆10Oct 1, 2020Updated 5 years ago
- my public website☆12Jul 8, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A2C is a special case of PPO!☆23May 20, 2022Updated 4 years ago
- Paper notes for my PhD on Machine Learning (mostly focused on Reinforcement Learning)☆17Jul 22, 2019Updated 7 years ago
- FQF(Fully parameterized Quantile Function for distributional reinforcement learning) is a general reinforcement learning framework for At…☆48Sep 26, 2020Updated 5 years ago
- An implementation of DreamerV2 written in JAX, with support for running multiple random seeds of an experiment on a single GPU.☆18Jan 16, 2023Updated 3 years ago
- Alamofire with ReactiveSwift☆13Oct 28, 2017Updated 8 years ago
- Adding Dreamer-v3's implementation tricks to CleanRL's PPO☆16May 19, 2023Updated 3 years ago
- ☆18Jul 24, 2026Updated last month
- Library for model based RL in robotics☆37Sep 10, 2018Updated 8 years ago
- iOS navigation controller with progress bar and TweetBot 3-like back button for navigation☆13Dec 9, 2015Updated 10 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Implicit Normalizing Flows + Reinforcement Learning☆62May 31, 2019Updated 7 years ago
- Developing different methods for expanding a query/topic in information retrieval and choosing the best expanded query using similarity m…☆11May 17, 2017Updated 9 years ago
- [NeurIPS'24] "NeuralFuse: Learning to Recover the Accuracy of Access-Limited Neural Network Inference in Low-Voltage Regimes" by Hao-Lun …☆10Sep 18, 2025Updated last year
- The Tensorflow code and a DeepMind Lab wrapper for my article "Meta-Reinforcement Learning" on FloydHub.☆37Mar 28, 2019Updated 7 years ago
- A small package used to visualize gradient descent of test functions.☆19Aug 1, 2022Updated 4 years ago
- TensorFlow & Keras implementation of DQN with HER (Hindsight Experience Replay)☆40Jul 31, 2020Updated 6 years ago
- Library of models for Protein Function prediction (part of the 18th top solution out of 1625 teams in CAFA5)☆20May 23, 2025Updated last year