Deep Q Networks
β101Oct 18, 2018Updated 7 years ago
Alternatives and similar repositories for human-level-control-through-deep-reinforcement-learning
Users that are interested in human-level-control-through-deep-reinforcement-learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π Paper: Human-level control through deep reinforcement learning πΉοΈβ57May 9, 2024Updated 2 years ago
- a implement and derivation of "CCS-TA: quality-guaranteed online task allocation in compressive crowdsensing"β12Jun 12, 2022Updated 4 years ago
- A2C, ACKTR and A2T implementations for ViZDoomβ10Dec 18, 2017Updated 8 years ago
- Code for WACV 2021 Paper "Meta Module Network for Compositional Visual Reasoning"β43May 13, 2021Updated 5 years ago
- The official baseline implementations for Chronocept. Published at EACL 2026.β10Aug 4, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Python DQN for practicable portfolio managementβ13Sep 6, 2020Updated 5 years ago
- PyOblige is Python wrapper for OBLIGE - random level generator for Doomβ11Jul 2, 2018Updated 8 years ago
- This is a project for creating and using IL datasets based on HuggingFace weights with multithreads for performance, and benchmarkingβ13Aug 5, 2026Updated 3 weeks ago
- Snakemake workflow to benchmark scRNA-seq data simulatorsβ14Aug 10, 2022Updated 4 years ago
- Source Code for the ICML Paper "Curriculum Reinforcement Learning via Constrained Optimal Transport"β16Jun 9, 2022Updated 4 years ago
- Applying Reinforcement Learning in Minecraft - Project Malmo Tutorial (Mar 2021)β21Nov 8, 2023Updated 2 years ago
- [BMVC 2023 Oral] Boost Video Frame Interpolation via Motion Adaptationβ19Aug 22, 2024Updated 2 years ago
- Reinforcement learning with a network of spiking agentsβ22Jun 8, 2020Updated 6 years ago
- ACM MULTIMEDIA CONFERENCE 2020β11Jul 28, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ROS wrapper for AirDetβ12Jul 24, 2022Updated 4 years ago
- An MDP solver for Pacman to win games. Base game from Berkeley cs188xβ10Jan 23, 2018Updated 8 years ago
- A2C is a special case of PPO!β23May 20, 2022Updated 4 years ago
- DCIC22ζ°εδΈε½22-ηεͺεΎεεε²η«θ΅η¬¬εεζΉζ‘β14Jul 18, 2022Updated 4 years ago
- A curated list of awesome tools and resources used at Qurit.β11Dec 31, 2020Updated 5 years ago
- EVE: Efficient Vision-Language Pre-training with Masked Prediction and Modality-Aware MoEβ10Mar 1, 2024Updated 2 years ago
- β21Mar 22, 2023Updated 3 years ago
- Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and β¦β13Aug 29, 2024Updated 2 years ago
- Official code repository for the EMNLP 2021 paperβ26Jan 30, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Inverse Scaling in Test-Time Computeβ26Dec 3, 2025Updated 8 months ago
- PDiT: Interleaving Perception and Decision-making Transformers for Deep Reinforcement Learning. AAMAS 2024 (full paper with oral presentaβ¦β10Dec 27, 2023Updated 2 years ago
- some mixture of experts architecture implementationsβ28Mar 22, 2024Updated 2 years ago
- Code for "Adversarial and Perceptual Refinement Compressed Sensing MRI Reconstruction"β30Sep 17, 2018Updated 7 years ago
- 53 implementations of synthetic learning problems from Geoffrey Hinton's experimental papers (1981-2022). Pure numpy, laptop-runnable, paβ¦β35Updated this week
- β26Mar 27, 2022Updated 4 years ago
- awesome deep learning papers for reinforcement learningβ17Jan 10, 2018Updated 8 years ago
- [ICRA 2024] WLST: Weak Labels Guided Self-training for Weakly-supervised Domain Adaptation on 3D Object Detectionβ12Feb 6, 2024Updated 2 years ago
- Differentiable First-Order Logic Reasoning for Visual Question Answeringβ45Mar 7, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Raspberry Pi upload to Google Driveβ13Jun 10, 2017Updated 9 years ago
- Advanced_Data_Integration_Projectβ11Jul 31, 2018Updated 8 years ago
- Reinforcement Learning Seminar at the Chinese University of Hong Kong, Shenzhen, China.β21Nov 17, 2023Updated 2 years ago
- A Streamlit app to add structured tags to a dataset cardβ23Jun 30, 2022Updated 4 years ago
- β11Sep 29, 2021Updated 4 years ago
- MinHash implementation in Pythonβ12Aug 24, 2024Updated 2 years ago
- 6th place solutionβ22Feb 28, 2023Updated 3 years ago