Implementation of Russo and Van Roy work on Information Directed Sampling (2017)
☆21Jan 18, 2019Updated 7 years ago
Alternatives and similar repositories for Information_Directed_Sampling
Users that are interested in Information_Directed_Sampling are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14May 30, 2019Updated 7 years ago
- ☆16Jun 10, 2022Updated 4 years ago
- ☆19Apr 15, 2026Updated 4 months ago
- Official implementation for the paper: "Shallow Updates for Deep Reinforcement Learning"☆18Nov 2, 2017Updated 8 years ago
- ☆10May 15, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆29May 27, 2024Updated 2 years ago
- INTeractive learning via REPresentatIon Discovery☆36Jun 2, 2024Updated 2 years ago
- Non official torchnet package for vision☆20Feb 4, 2017Updated 9 years ago
- Code for "Best arm identification in multi-armed bandits with delayed feedback", AISTATS 2018.☆20Apr 3, 2018Updated 8 years ago
- Reinforcement Learning implementations and research prototyping in TensorFlow☆81Apr 28, 2019Updated 7 years ago
- Demos for the MiniWoB++ benchmark☆21Feb 23, 2018Updated 8 years ago
- A ready-to-use Jemdoc-based website for research groups and similar organizations. It also contains a dynamic news/RSS-feed system which …☆11May 5, 2022Updated 4 years ago
- Code for Optimistic Exploration even with a Pessimistic Initialisation☆14Aug 4, 2020Updated 6 years ago
- An impulse engine written in Rust with WebAssembly☆15Aug 26, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implement DQN and DDQN algorithm on Atari games,such as BreakoutNoFrameskip-v4, PongNoFrameskip-v4,BoxingNoFrameskip-v4.☆15Jun 30, 2020Updated 6 years ago
- Code for 'Contrastive Multi-Document Question Generation'☆11Oct 16, 2022Updated 3 years ago
- Stable Diffusion in pure C/C++☆14May 29, 2024Updated 2 years ago
- ☆14May 23, 2021Updated 5 years ago
- Some starter code for training/testing some basic CNN models given our data.☆10Feb 15, 2017Updated 9 years ago
- Monte Carlo Tree Search for Markov decision processes using the POMDPs.jl framework☆81Nov 16, 2025Updated 9 months ago
- 📴 OffCon^3: SOTA PyTorch SAC and TD3 Implementations (arxiv: 2101.11331)☆25Jun 20, 2021Updated 5 years ago
- Official implementation for the paper "Quantum Bayesian Optimization" accepted to NeurIPS 2023.☆12Jan 7, 2024Updated 2 years ago
- Java framework for experimenting with a 2-D version of the voxel-based soft robots.☆20Mar 31, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Farming Environment Gym factory for Reinforcement Learning☆24Aug 22, 2025Updated last year
- ☆27May 17, 2019Updated 7 years ago
- Decoupled Reward-free ExplorAtion and Execution for Meta-reinforcement learning☆92Feb 13, 2023Updated 3 years ago
- ☆11Feb 20, 2017Updated 9 years ago
- Recurrent Additive Networks for Tensorflow☆16Jun 30, 2017Updated 9 years ago
- Some hard problems for reinforcement learning.☆32Oct 5, 2018Updated 7 years ago
- A python implementation of PROCLUS: PROjected CLUStering algorithm.☆10Jan 12, 2015Updated 11 years ago
- Recommendation engine and it's algorithms in python , R .☆12Oct 26, 2018Updated 7 years ago
- Reward shaping approach for instruction following settings, leveraging language at multiple levels of abstraction.☆21Mar 9, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ArXiv'18 implementation of amortized maximum likelihood (AML) for high-quality, weakly-supervised shape completion.☆11Nov 30, 2018Updated 7 years ago
- ☆12Nov 28, 2022Updated 3 years ago
- Retrieve information from DBLP and update BibTex files automatically☆53Jun 4, 2022Updated 4 years ago
- Mac port of Torcs, The Open Racing Car Simulator☆12Jun 16, 2010Updated 16 years ago
- For TAMP experiments using Drake☆13Jun 4, 2024Updated 2 years ago
- OpenAI gym environment for evolving morphologies of 2D virtual creatures.☆34Jul 26, 2023Updated 3 years ago
- Tutorials on how to use EAGERx☆16Aug 14, 2025Updated last year