rl on super-mario-bros
☆59Dec 23, 2020Updated 5 years ago
Alternatives and similar repositories for Supermariobros-PPO-pytorch
Users that are interested in Supermariobros-PPO-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- KDSS is the framework for knowledge distillation from LLMs☆12Nov 5, 2025Updated 11 months ago
- Code Repository for the NeurIPS 2024 Paper "Toward Efficient Inference for Mixture of Experts".☆19Oct 30, 2024Updated last year
- Incremental Mobile User Profiling: Reinforcement Learning with Spatial Knowledge Graph for Modeling Event Streams☆15Jul 25, 2024Updated 2 years ago
- Implementation of "Temporal Recurrent Networks for Online Action Detection"☆23May 6, 2019Updated 7 years ago
- ☆21Mar 18, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for MLSys 2024 Paper "SiDA-MoE: Sparsity-Inspired Data-Aware Serving for Efficient and Scalable Large Mixture-of-Experts Models"☆23Apr 13, 2024Updated 2 years ago
- ☆22Feb 2, 2021Updated 5 years ago
- 良心软件推荐分享:eDiary本地记事本,Windows记笔记软件;OCR截图识别文字......☆11Jun 17, 2020Updated 6 years ago
- A learning-based scheme to capture external force/torque caused by payload of tethered-UAV system☆21May 27, 2025Updated last year
- real time saliency android app with ncnn implementation☆11Feb 9, 2021Updated 5 years ago
- Integrating opencv with mujoco.☆12Mar 25, 2025Updated last year
- An example application of neural network distillation to MNIST☆11Sep 29, 2016Updated 10 years ago
- CURLA: CURL x CARLA -- Robust end-to-end Autonomous Driving by combining Contrastive Learning and Reinforcement Learning☆17Feb 6, 2024Updated 2 years ago
- ☆14Jul 30, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Lua runtime implemented in Rust☆17Oct 22, 2023Updated 2 years ago
- Official code repository for ICCV 2021 paper: Gravity-Aware Monocular 3D Human Object Reconstruction☆16Oct 12, 2021Updated 4 years ago
- My implementation (PyTorch) for the paper SST: Single-Stream Temporal Action Proposals (http://vision.stanford.edu/pdf/buch2017cvpr.pdf).☆10Dec 8, 2022Updated 3 years ago
- ☆36Sep 15, 2017Updated 9 years ago
- A simple coroutine library written in Zig language.☆15Sep 30, 2023Updated 3 years ago
- An ENet Implementation in C# , Look C code//github.com/lsalzman/enet☆15Oct 26, 2014Updated 11 years ago
- ☆13Oct 22, 2024Updated last year
- ☆11May 18, 2019Updated 7 years ago
- Using Deep Reinforcement Learning Project Repository☆10Nov 21, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆14May 28, 2023Updated 3 years ago
- CVPR2022 update everyday!☆11Apr 12, 2022Updated 4 years ago
- Code for TGRS 2021 paper. Edge-Aware Multiscale Feature Integration Network for Salient Object Detection in Optical Remote Sensing Images…☆12Apr 6, 2022Updated 4 years ago
- ☆20Apr 28, 2020Updated 6 years ago
- 深度学习500问,以问答形式对常用的概率知识、线性代数、机器学习、深度学习、计算机视觉等热点问题进行阐述,以帮助自己及有需要的读者。 全书分为18个章节,50余万字。由于水平有限,书中不妥之处恳请广大读者批评指正。 未完待续............ 如有意合作,联系sc…☆13Oct 14, 2021Updated 4 years ago
- Data from my web crawling project that ranked universities using the PageRank algorithm on a network comprised of faculty relationships.☆11Nov 7, 2021Updated 4 years ago
- 《机器学习》(西瓜书)公式推导解析,在线阅读地址:https://datawhalechina.github.io/pumpkin-book☆15Apr 3, 2019Updated 7 years ago
- ☆13Feb 6, 2018Updated 8 years ago
- Verlet Integration in 3D☆15Sep 26, 2014Updated 12 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Reinforcement learning on dynamic knowledge graphs☆31Jan 15, 2025Updated last year
- ppo+action mask for atari tennis agent☆12Mar 2, 2023Updated 3 years ago
- Some tools to operate PaddlePaddle model☆76Apr 4, 2022Updated 4 years ago
- ☆11Mar 16, 2021Updated 5 years ago
- Python hotkey manager that works on Linux, MacOS, Windows and Cygwin.☆21Jan 13, 2025Updated last year
- Code used in the paper "Learning to Learn from Web Data through Deep Semantic Embeddings" ECCV 2018 MULA Workshop☆11Aug 1, 2018Updated 8 years ago
- ☆13Nov 21, 2025Updated 10 months ago