Value & Policy Iteration for the frozenlake environment of OpenAI
☆15May 14, 2019Updated 7 years ago
Alternatives and similar repositories for frozenlake
Users that are interested in frozenlake are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Flutter Hackathon '20 Project.☆13Apr 18, 2024Updated 2 years ago
- Slides, Code, and Exercises to support [R Quickstart tutorial](http://conferences.oreilly.com/strata/hadoop-big-data-ca/public/schedule/d…☆10Mar 25, 2016Updated 10 years ago
- ☆17Jun 5, 2026Updated last month
- Tensorflow implementation of Asynchronous Advantage Actor Critic (A3C) from "Asynchronous Methods for Deep Reinforcement Learning".☆24Apr 20, 2017Updated 9 years ago
- Lime: Explaining the predictions of any machine learning classifier☆16May 27, 2019Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A startup search engine made using embeddings built on crunchbase company descriptions☆11Dec 2, 2015Updated 10 years ago
- Generate text and predict next word for an initial piece of text using RNNs and LSTMs☆11Jun 27, 2017Updated 9 years ago
- Hands On Reinforcement Learning with Python[Video], Published by Packt☆13Jan 14, 2021Updated 5 years ago
- Implementation of Siamese Network using MXNet/Gluon☆10May 18, 2018Updated 8 years ago
- Deep Q-Network (DQN) to play classic Atari Games☆11Sep 18, 2017Updated 8 years ago
- A reinforcement learning agent that learns to solve mazes using Group Relative Policy Optimization (GRPO).☆12Feb 9, 2025Updated last year
- Bayesian FlowNetS in Tensorflow☆21Dec 20, 2017Updated 8 years ago
- Text generation for the Shakespeare model☆13Apr 26, 2017Updated 9 years ago
- Forecasting library in python☆13Sep 6, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Notebook from my blog☆15Apr 9, 2017Updated 9 years ago
- ☆15May 31, 2017Updated 9 years ago
- This is the code for "Iphone XS Supply Chain" By Siraj Raval on Youtube☆19Sep 18, 2018Updated 7 years ago
- 의사결정(DP) + 강화학습(RL) + 온라인광고(OA) + 파이썬웹(Pyweb)☆10Nov 30, 2016Updated 9 years ago
- ☆15Aug 24, 2022Updated 3 years ago
- ☆17Aug 24, 2022Updated 3 years ago
- Pytorch implementation of BEAR in "Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction"☆11Oct 29, 2019Updated 6 years ago
- This is the code for "Internet of Things Optimization" By Siraj Raval on Youtube☆30Sep 24, 2018Updated 7 years ago
- Analyze a real-time IPv4 packet stream and export metrics about the data flows☆14Jan 29, 2020Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A visualization system for RoboCup@Home robots☆10Jul 12, 2019Updated 7 years ago
- MoveIt! config files for the Aldebaran NAO☆10Jan 20, 2017Updated 9 years ago
- ☆20Aug 16, 2022Updated 3 years ago
- Tutorial repository for raisimGym☆29Nov 27, 2019Updated 6 years ago
- Causal discovery with typed directed acyclic graphs (t-DAG). This is a ServiceNow Research project that was started at Element AI.☆13Jul 6, 2023Updated 3 years ago
- Make sense of deep neural networks using TensorBoard☆20May 9, 2017Updated 9 years ago
- Quick-Data-Science-Experiments☆19Dec 12, 2017Updated 8 years ago
- The codebase for Inducing Causal Structure for Interpretable Neural Networks☆11Dec 3, 2021Updated 4 years ago
- ☆11Apr 4, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for Sibling Rivalry and experiments presented in associated paper☆18May 1, 2025Updated last year
- A repo to design basic Policy Gradient labs☆12Jul 6, 2023Updated 3 years ago
- Official Codebase for "Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control" (NeurIPS 2024)☆15Oct 29, 2024Updated last year
- performing sentiment analysis on the whatsapp chats.☆23Oct 17, 2017Updated 8 years ago
- python solver for tangram puzzles☆14Jan 3, 2019Updated 7 years ago
- Source code for paper Mroueh, Sercu, Rigotti, Padhi, dos Santos, "Sobolev Independence Criterion", NeurIPS 2019☆14Jun 17, 2024Updated 2 years ago
- Ros2 Docker Mac Netowrk☆15Jul 1, 2021Updated 5 years ago