Some notes and experience about David Silver's Reinforcement Learning Course
☆47Jun 24, 2019Updated 7 years ago
Alternatives and similar repositories for D.Silver_RL_Course
Users that are interested in D.Silver_RL_Course are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Unofficial Code for NeurIPS 2021 paper "Regret Minimization Experience Replay in Off-policy Reinforcement Learning"☆14May 24, 2021Updated 5 years ago
- solutions to David Silver's RL course project Easy21☆19Jun 28, 2016Updated 10 years ago
- Implementaion of the WWW paper Implicit User Awareness Modeling via Candidate Items for CTR Prediction in Search Ads☆18Apr 27, 2022Updated 4 years ago
- EDIS: Energy-guided DIffusion Sampling☆19Aug 10, 2024Updated 2 years ago
- The reproduce of Transformer architecture in paper "Attention is all your need"☆18May 15, 2020Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A python module designed for agile RL algorithm developing.☆26Jul 11, 2024Updated 2 years ago
- ☆10Jul 31, 2026Updated last month
- Code for recreating the results of our RSS 2020 paper, 'Learning Memory-Based Control for Human-Scale Bipedal Locomotion.'☆10Aug 18, 2022Updated 4 years ago
- This repository contains the scripts used during my participation on CIKM Cup 2016 (see http://cikmcup.org/ and https://competitions.coda…☆11Nov 4, 2016Updated 9 years ago
- Official implementation of ``Neural Pruning via Sparsity-indexed ODE: A Continuous Sparsity Viewpoint"☆11Jun 15, 2023Updated 3 years ago
- This library implements a model predictive control algorithm to generate walk trajectories with automatic foot step placement.☆14May 15, 2019Updated 7 years ago
- ☆10May 29, 2018Updated 8 years ago
- Customizable RecSys Simulator for OpenAI Gym☆26Dec 7, 2021Updated 4 years ago
- CK workflow, portable packages and other artifacts for the ReQuEST-ASPLOS'18 submission:☆12Jan 16, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Generate a weighted Voronoi diagram from a set of points and attribute values☆15May 30, 2017Updated 9 years ago
- Evaluate the Quality of Critique☆37Jun 1, 2024Updated 2 years ago
- Nankai Online Judge Front End☆16Oct 14, 2021Updated 4 years ago
- Implementation of Differential Learning Rate in Keras☆11Jun 4, 2019Updated 7 years ago
- Code for creating recurrent neural network with rotational dynamics. Model is discussed in detail in "Rotational Dynamics Reduce Interfer…☆17Jul 23, 2020Updated 6 years ago
- This is the code for G2MILP, a deep learning-based mixed-integer linear programming (MILP) instance generator.☆38Oct 3, 2024Updated last year
- CCKS 2020:面向金融领域的小样本跨类迁移事件抽取。该项目实现基于MRC的事件抽取方法☆39Oct 27, 2022Updated 3 years ago
- Notes for the Reinforcement Learning course by David Silver along with implementation of various algorithms.☆871Mar 31, 2022Updated 4 years ago
- ☆17Feb 21, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for the article "Automatic Temperature Control for Neural Machine Translation" (EMNLP 2018)☆14Apr 16, 2019Updated 7 years ago
- ☆13May 20, 2022Updated 4 years ago
- ☆12Dec 10, 2018Updated 7 years ago
- Reinforcement Learning Seminar at the Chinese University of Hong Kong, Shenzhen, China.☆21Nov 17, 2023Updated 2 years ago
- This is the companion code for the method reported in the paper "Learning game-theoretic models of multiagent trajectories using implicit…☆12Feb 8, 2021Updated 5 years ago
- 深度学习实战☆16Nov 12, 2019Updated 6 years ago
- ☆12Sep 23, 2024Updated last year
- [Findings of EMNLP22] From Mimicking to Integrating: Knowledge Integration for Pre-Trained Language Models☆19Mar 16, 2023Updated 3 years ago
- JAX implementation of the Mistral 7b v0.1 model☆13Mar 27, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Generate random terrain with Python☆14May 6, 2013Updated 13 years ago
- Pedestrian Intention and Trajectory Prediction Study Using PIEPredict and LSTM Models over PIE and Waymo-Open datasets.☆13May 30, 2021Updated 5 years ago
- Code for a multi-agent particle environment used in the paper "Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments"☆15Aug 30, 2021Updated 5 years ago
- A project on deep learning.☆14May 16, 2020Updated 6 years ago
- Benchmarking suite for MushroomRL Deep RL algorithms☆17Updated this week
- [ACL 2023] To Copy Rather Than Memorize: A Vertical Learning Paradigm for Knowledge Graph Completion☆11Feb 3, 2023Updated 3 years ago
- Evaluation of TD-MPC2.☆21Jan 21, 2024Updated 2 years ago