Official implementation of the paper "Approximating two value functions instead of one: towards characterizing a new family of Deep Reinforcement Learning Algorithms": https://arxiv.org/abs/1909.01779 To appear at the next NeurIPS2019 DRL-Workshop
☆11Jul 14, 2021Updated 5 years ago
Alternatives and similar repositories for Deep-Quality-Value-Family
Users that are interested in Deep-Quality-Value-Family are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Documentation and guidelines for the Alan GPU cluster at the University of Liège.☆21Jul 19, 2023Updated 3 years ago
- Implementation prototype of the Deep Deterministic Off-Policy Gradient (DD-OPG) method.☆11Jun 12, 2019Updated 7 years ago
- Code repository for the generalized Galton board example in the paper "Mining gold from implicit models to improve likelihood-free infere…☆34Dec 2, 2019Updated 6 years ago
- Deep Collaboration Network in pytorch☆15Mar 15, 2018Updated 8 years ago
- This is a TensorFlow implementation of DeepMind's A Distributional Perspective on Reinforcement Learning.(C51-DDPG)☆11Sep 14, 2017Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- High-quality reference implementations of various algorithms for Inverse Reinforcement Learning☆13Jun 20, 2018Updated 8 years ago
- Sentiment Analysis via RNN, RNTN. Based on Stanford's Sentiment Analysis page.☆10Feb 5, 2015Updated 11 years ago
- Code for Policy Consolidation for Continual Reinforcement Learning☆10May 12, 2019Updated 7 years ago
- Emotiv SDK Community Edition☆13Oct 9, 2015Updated 10 years ago
- Bot for Minecraft environment☆13Jun 18, 2019Updated 7 years ago
- Information and Cyber Security Certifications☆17Feb 14, 2019Updated 7 years ago
- ☆14Aug 8, 2023Updated 2 years ago
- Code for reproducing the experiment results of the paper Imitation Learning with Sinkhorn Distances.☆14Aug 2, 2020Updated 5 years ago
- Research on Inverse Reinforcement Learning for self driving vehicles at UCLA☆13Nov 7, 2018Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A Benchmark Platform for Reinforcement Learning Based Dynamic Treatment Regime☆15Dec 7, 2024Updated last year
- Semantic alignment of astronomical data with natural language using multi-modal models. (Jax) Code associated with https://arxiv.org/abs/…☆17Oct 18, 2024Updated last year
- In Progress : State of the art Distributed Distributional Deep Deterministic Policy Gradient algorithm implementation in pytorch.☆19Jun 15, 2018Updated 8 years ago
- Code for "Training Generative Adversarial Networks with Binary Neurons by End-to-end Backpropagation"☆26Oct 30, 2019Updated 6 years ago
- Multimodal AI Assistant is an advanced chatbot designed for local interactions with LLMs, files, and multimodal functionalities. Built on…☆19Oct 23, 2024Updated last year
- The XENON1T raw data processor [deprecated]☆16Jan 3, 2021Updated 5 years ago
- Implicit Normalizing Flows + Reinforcement Learning☆62May 31, 2019Updated 7 years ago
- Procedural object generation for robotic manipulation☆11Oct 6, 2018Updated 7 years ago
- Train agents on MiniGrid from human demonstrations using Inverse Reinforcement Learning☆13Apr 15, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A ROS pipeline for GPU based feature detection, description and matching☆15Jul 2, 2018Updated 8 years ago
- Tutorial on Multi-Agent Reinforcement for Train Scheduling☆11May 18, 2020Updated 6 years ago
- A tensorflow implementation of hindsight experience replay☆17Apr 19, 2018Updated 8 years ago
- ShipAI simplified to medium tutorial☆16Dec 12, 2018Updated 7 years ago
- Repository for the paper "Adversarial Variational Optimization of Non-Differentiable Simulators"☆16Dec 17, 2018Updated 7 years ago
- ZeroMQ For Robot Control☆15Oct 6, 2016Updated 9 years ago
- High Performance Lane Detection using Computer Vision and CUDA.☆16Jun 14, 2017Updated 9 years ago
- Dockerization of FlytSIM, for easy deployment to linux, windows and mac☆15Feb 15, 2023Updated 3 years ago
- VAE + Quantile Networks for MNIST☆12Nov 29, 2018Updated 7 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for "Dynamic NeRFs for Soccer Scenes", by Lewin, Vandegar, Hoyoux, Barnich, and Louppe. (2023)☆25May 31, 2024Updated 2 years ago
- Bayesian Inverse Reinforcement Learning with simple environments☆19May 17, 2022Updated 4 years ago
- Code for the paper "Towards Reliable Simulation-Based Inference with Balanced Neural Ratio Estimation".☆15Nov 14, 2022Updated 3 years ago
- ☆19Feb 27, 2023Updated 3 years ago
- Normalizing flow models allowing for a conditioning context, implemented using Jax, Flax, and Distrax.☆21Mar 10, 2024Updated 2 years ago
- This is the official repository for the "Towards Vision-Language Mechanistic Interpretability: A Causal Tracing Tool for BLIP" paper acce…☆25Feb 16, 2026Updated 5 months ago
- Path planner that is able to create smooth, safe paths for the car to follow along a 3 lane highway with traffic.☆16Aug 28, 2017Updated 8 years ago