Implementation of TD Lambda algorithm, with neural network for value estimation
☆20Apr 16, 2018Updated 8 years ago
Alternatives and similar repositories for Deep-Watkins-Q-and-Actor-Critic
Users that are interested in Deep-Watkins-Q-and-Actor-Critic are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A reinforcement learning package implemented in Torch☆11Jan 24, 2016Updated 10 years ago
- ☆10Aug 15, 2016Updated 9 years ago
- This repo contains a PyTorch implementation of a CNN model for multi-label Image classification model deployed on heroku.☆14Feb 28, 2021Updated 5 years ago
- Python code to automatically produce a summary of a piece of text.☆11Sep 8, 2016Updated 9 years ago
- SEC Form 13f Securities datasets☆14Apr 19, 2019Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆12Mar 17, 2024Updated 2 years ago
- JAX implementation of the T5 model: Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer☆24Jun 10, 2023Updated 3 years ago
- Dreamer on JAX☆16Jan 19, 2022Updated 4 years ago
- my pdf files☆12Aug 7, 2019Updated 6 years ago
- Code for optimal execution☆12Oct 29, 2020Updated 5 years ago
- Implement the model of Halperin and Feldshteyn for DJIA and SP500☆10Apr 4, 2019Updated 7 years ago
- ☆18Dec 11, 2015Updated 10 years ago
- This repository provides several functions to generate and process race track maps containing specific local information, which is furthe…☆13Dec 2, 2021Updated 4 years ago
- LSTM based neural network that predicts the state of the vehicle in terms of position and velocity.☆14May 7, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Use Logitech G27 steering wheel to remote control openxc-vehicle-simulator☆13Feb 18, 2014Updated 12 years ago
- A Higher-order HMM with EM algo.☆16May 4, 2022Updated 4 years ago
- Equinor's collection of subsurface reservoir modelling scripts☆21Updated this week
- A collection of meta-learning algorithms in Jax☆24Sep 3, 2022Updated 3 years ago
- BankHoldingCompanyData☆14Mar 11, 2026Updated 4 months ago
- NLP Resources for Indian Languages☆10Nov 9, 2020Updated 5 years ago
- Code for the paper Dynamics Generalisation in Reinforcement Learning via Adaptive Context-Aware Policies (NeurIPS 2023). https://arxiv.or…☆15Nov 21, 2023Updated 2 years ago
- Energy-based Surprise Minimization for Multi-Agent Value Factorization☆12Oct 20, 2023Updated 2 years ago
- Simple AI Agent Trained to play Hangman☆13Sep 27, 2019Updated 6 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Python API and analysis of Chicago's bikeshare☆10Dec 8, 2022Updated 3 years ago
- Lets get started with Machine Learning☆18Mar 10, 2023Updated 3 years ago
- An analysis, with a focus on demand forecasting, of transactional data associated with over 2.5 million customers and 31,868 SKUs over th…☆17Oct 4, 2020Updated 5 years ago
- 📻「我的收藏电台」写了个播放界面☆22Nov 8, 2022Updated 3 years ago
- $GIT_REV in your dokku env☆15Jun 28, 2018Updated 8 years ago
- ☆20Mar 28, 2023Updated 3 years ago
- Implementation of RRT, RRT-connect, RRT*, and PRM in c++☆13Oct 25, 2017Updated 8 years ago
- An PPO - LSTM based RL agent to solve the classic word game - Hangman☆15Nov 20, 2024Updated last year
- Indeed web crawler☆11Aug 14, 2018Updated 7 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Tools to make git easier to use and to avoid the learning curve☆21Apr 3, 2019Updated 7 years ago
- preprocessing of the MUC4 dataset☆11Aug 28, 2012Updated 13 years ago
- This is the code for "DeepMind Reinforcement Learning" By Siraj Raval on Youtube☆84Sep 5, 2018Updated 7 years ago
- Link to paper: https://www.ssrn.com/abstract=3804655☆14Jul 27, 2021Updated 4 years ago
- Reinforcement Learning via Latent State Decoding☆29Jun 12, 2023Updated 3 years ago
- Deep learning models for contextual multi-armed bandit setting☆13May 16, 2021Updated 5 years ago
- Code for demonstration example-task in RUDDER blog☆24May 19, 2020Updated 6 years ago