Implementation of TD Lambda algorithm, with neural network for value estimation
☆20Apr 16, 2018Updated 8 years ago
Alternatives and similar repositories for Deep-Watkins-Q-and-Actor-Critic
Users that are interested in Deep-Watkins-Q-and-Actor-Critic are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of True Online TD(lambda) with a Fourier Basis function approximator.☆13May 9, 2015Updated 11 years ago
- A reinforcement learning package implemented in Torch☆11Jan 24, 2016Updated 10 years ago
- ☆31Oct 24, 2023Updated 2 years ago
- Leveraging Recursive Gumbel-Max Trick for Approximate Inference in Combinatorial Spaces, NeurIPS 2021☆14Dec 11, 2021Updated 4 years ago
- ☆10Aug 15, 2016Updated 10 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆11Jul 25, 2021Updated 5 years ago
- Python code to automatically produce a summary of a piece of text.☆11Sep 8, 2016Updated 10 years ago
- JAX implementation of the T5 model: Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer☆24Jun 10, 2023Updated 3 years ago
- Dreamer on JAX☆16Jan 19, 2022Updated 4 years ago
- 关于Fault-Tolerant Federated Reinforcement Learning with Theoretical Guarantee这篇论文的详细代码解读☆12Dec 27, 2023Updated 2 years ago
- Code for optimal execution☆12Oct 29, 2020Updated 5 years ago
- Implement the model of Halperin and Feldshteyn for DJIA and SP500☆10Apr 4, 2019Updated 7 years ago
- ☆18Dec 11, 2015Updated 10 years ago
- RAD: Reinforcement Learning with Augmented Data (code for procgen experiments)☆19Mar 29, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- This repository provides several functions to generate and process race track maps containing specific local information, which is furthe…☆13Dec 2, 2021Updated 4 years ago
- LSTM based neural network that predicts the state of the vehicle in terms of position and velocity.☆14May 7, 2021Updated 5 years ago
- A classifier for cat and dog images (Response to Siraj's challenge of the week)☆38Feb 25, 2017Updated 9 years ago
- MPC package for solving optimal control problems☆19Jun 11, 2025Updated last year
- Implementation of Denoising Diffusion Probabilistic Models (DDPM) in JAX and Flax.☆22Oct 12, 2023Updated 2 years ago
- Django extensions for Hyperview mobile apps.☆23Feb 2, 2026Updated 7 months ago
- Equinor's collection of subsurface reservoir modelling scripts☆21Updated this week
- A collection of meta-learning algorithms in Jax☆24Sep 3, 2022Updated 4 years ago
- A Bachelor Thesis implementation of RRT, RRT* and Informed RRT*.☆14Jul 9, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Replicated the Alpha Go Zero paper but applied it to the game Santorini.☆13Jan 27, 2018Updated 8 years ago
- NLP Resources for Indian Languages☆10Nov 9, 2020Updated 5 years ago
- Code for the paper Dynamics Generalisation in Reinforcement Learning via Adaptive Context-Aware Policies (NeurIPS 2023). https://arxiv.or…☆15Nov 21, 2023Updated 2 years ago
- Energy-based Surprise Minimization for Multi-Agent Value Factorization☆12Oct 20, 2023Updated 2 years ago
- Code for experiments done for EMNLP2020.☆11Dec 8, 2022Updated 3 years ago
- A small and easy python-only decoder-only for AIS messages : AIVDM/AIVDO☆24Mar 14, 2019Updated 7 years ago
- ☆15Oct 6, 2019Updated 6 years ago
- Lets get started with Machine Learning☆18Mar 10, 2023Updated 3 years ago
- Code from the paper An actor-critic algorithm with policy gradients to solve the job shop scheduling problem using deep double recurrent …☆13Mar 20, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- WOS(web of science)网站文献爬取工具☆18Sep 7, 2018Updated 8 years ago
- Graph Convolutional Networks in JAX☆34Jan 3, 2021Updated 5 years ago
- Sim2Real Transfer for Deep Reinforcement Learning with Stochastic State Transition Delays, CORL-2020.☆26Jun 3, 2021Updated 5 years ago
- An PPO - LSTM based RL agent to solve the classic word game - Hangman☆15Nov 20, 2024Updated last year
- Tools to make git easier to use and to avoid the learning curve☆21Apr 3, 2019Updated 7 years ago
- This is the code for "DeepMind Reinforcement Learning" By Siraj Raval on Youtube☆84Sep 5, 2018Updated 8 years ago
- Using Word2Vec to explore semantic similarities between the entities of "A Song of Ice and Fire" ("Game of Thrones").☆24Jul 9, 2016Updated 10 years ago