Implementation of q-learning using TensorFlow
☆58May 9, 2017Updated 9 years ago
Alternatives and similar repositories for dqn
Users that are interested in dqn are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Notes on Logistic Regression and OWLQN☆26Apr 8, 2017Updated 9 years ago
- 《显卡就是开发板》 所提到的文档,代码和程序☆19Apr 10, 2018Updated 8 years ago
- Reimplementation of the clockwork recurrent neural network in Torch7☆14Feb 4, 2016Updated 10 years ago
- Mobile Ad Hoc (MANET) simulation and analysis using OMNET++.☆11Mar 3, 2020Updated 6 years ago
- LoRa AODV Routing Protocol implementation modifying FLoRa framework. It works on Omnet++☆12Mar 26, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆16Feb 1, 2022Updated 4 years ago
- simple example of gradient-based hyperparameter optimization using tensorflow☆19Feb 29, 2016Updated 10 years ago
- A reinforcement learning algorithm for congestion control, together with a realistic Omnet++ network simulation environment☆37Jul 20, 2023Updated 3 years ago
- An attempt at implementing ideas in "Learning to Transduce with Unbounded Memory" (http://arxiv.org/abs/1506.02516)☆11Jul 27, 2016Updated 10 years ago
- This repository implements Distilled Graph Attention Policy Networks (DGAPNs), a curiosity-driven reinforcement learning model to generat…☆21Jan 21, 2022Updated 4 years ago
- ClockworkRNN implementation using python and Theano☆17Apr 6, 2015Updated 11 years ago
- Urban UAV Mobility Model for NS3☆15Aug 2, 2022Updated 4 years ago
- ☆13May 3, 2017Updated 9 years ago
- Scalable Distributed LDA implementation for Spark & Glint☆29Sep 27, 2016Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Python implementation of tabular asynchronous actor critic☆11May 3, 2016Updated 10 years ago
- Code and data for the AAAI 2015 paper entitled: "Predicting the demographics of Twitter users from social evidence using website traffic …☆45Sep 18, 2019Updated 6 years ago
- Hierarchical Encoder Decoder for Dialog Modelling☆16May 20, 2015Updated 11 years ago
- A Pygame+Pymunk Carrom Simulation Testbed for reinforcement learning. [CS747][ Foundations of Intelligent and Learning Agents]☆15Jun 24, 2019Updated 7 years ago
- Simulation of LEACH routing protocol on OMNET++☆23Apr 16, 2023Updated 3 years ago
- Deep structured semantic model☆32May 5, 2016Updated 10 years ago
- Implementation of condnets☆16Apr 21, 2016Updated 10 years ago
- tools for alpha research☆23Dec 20, 2017Updated 8 years ago
- Long Short-Term Memory Recurrent Neural Networks☆26Jun 11, 2015Updated 11 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Python library for custom entity recognition using Sklearn CRF☆17Aug 1, 2017Updated 9 years ago
- record and share my reading everyday☆12Apr 1, 2016Updated 10 years ago
- generative models for speech☆20Jul 4, 2016Updated 10 years ago
- ☆10Jul 21, 2017Updated 9 years ago
- Backprop training of recurrent neural networks with Hebbian plastic connections☆20Jun 30, 2021Updated 5 years ago
- vue sudoku (数独)☆11Jan 5, 2023Updated 3 years ago
- Variational Bayes for NN in Torch7 (http://papers.nips.cc/paper/4329-practical-variational-inference-for-neural-networks.pdf)☆10Mar 23, 2015Updated 11 years ago
- Simplest Version of playing Atari with Deep Q Learning in Tensorflow☆155Oct 19, 2017Updated 8 years ago
- Lightweight, Portable, Flexible Distributed/Mobile Deep Learning with Dynamic, Mutation-aware Dataflow Dep Scheduler; for Python, R, Juli…☆13Dec 30, 2016Updated 9 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Using a paper from Google DeepMind I've developed a new version of the DQN using threads exploration instead of memory replay as explain …☆84Mar 4, 2016Updated 10 years ago
- a large scale lbfgs using a method in nips 2014 paper "Large-scale L-BFGS using MapReduce".☆13May 30, 2015Updated 11 years ago
- Benchmark of different RL algorithm☆13Dec 8, 2022Updated 3 years ago
- General experiments on Vanilla RNN and LSTM in Theano.☆16Aug 23, 2015Updated 11 years ago
- RL study guide — foundations through RLHF, DPO, GRPO, RLVR, agentic RL, and offline RL. Hand-written CS294 notes, 19 lecture drafts, 5 te…☆163Jul 1, 2026Updated 2 months ago
- UNSW's RoboCup Standard Platform League Team☆12Jun 18, 2022Updated 4 years ago
- AlphaGo代码☆11Apr 25, 2016Updated 10 years ago