DreamSmooth: Improving Model-Based RL with Reward Smoothing (ICLR 2024)
☆12May 6, 2024Updated 2 years ago
Alternatives and similar repositories for dreamsmooth
Users that are interested in dreamsmooth are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Release of Multistep Quasimetric Estimation (MQE)☆20Mar 13, 2026Updated 5 months ago
- A PyTorch Implementation of PlaNet: A Deep Planning Network for Reinforcement Learning☆13Aug 31, 2020Updated 6 years ago
- MuZero for Combinatorial Action Spaces: open-source codebase for MA-Gumbel-AlphaZero, MA-Sampled-AlphaZero, MA-Gumbel-MuZero and MA-Sampl…☆24Jan 22, 2024Updated 2 years ago
- paper on dexpilot☆15Oct 14, 2019Updated 6 years ago
- Official implementation of "Cross-Domain Transfer via Semantic Skill Imitation", Pertsch et al., CoRL 2022☆15Dec 15, 2022Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- CRIL: Continual Robot Imitation Learning via Generative Dynamics Model☆21Mar 13, 2021Updated 5 years ago
- ☆12Oct 24, 2024Updated last year
- Evaluation of TD-MPC2.☆21Jan 21, 2024Updated 2 years ago
- Code for ICRA24 paper "Think, Act, and Ask: Open-World Interactive Personalized Robot Navigation" Paper//arxiv.org/abs/2310.07968 …☆31Jun 18, 2024Updated 2 years ago
- Official implementation of HEAD CoRL 2025☆18Aug 9, 2025Updated last year
- The process scheduler used in Linux kernel (since version 2.6.23), simulated using Python.☆11Feb 13, 2022Updated 4 years ago
- Simple power management for my NAS.☆14Nov 17, 2024Updated last year
- ☆25Sep 23, 2024Updated last year
- M^3PC: Test-Time Model Predictive Control for Pretrained Masked Trajectory Model, ICLR 2025☆19Mar 17, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Transformer-based World Models☆91Apr 4, 2023Updated 3 years ago
- Codebase for [Order Matters: Agent-by-agent Policy Optimization](https://openreview.net/forum?id=Q-neeWNVv1)☆32Nov 22, 2025Updated 9 months ago
- GRAM: Generalization in Deep RL with a Robust Adaptation Module☆15Jul 16, 2026Updated last month
- ☆15Dec 16, 2021Updated 4 years ago
- ☆47Jan 29, 2024Updated 2 years ago
- ☆12Oct 5, 2020Updated 5 years ago
- A movie recommendation system built using Scala, Spark and Hadoop☆16Jan 21, 2022Updated 4 years ago
- [ICML 2024] The algorithm of Reinforcement Learning with an Assistant Reward Agent (ReLara)☆19Aug 2, 2024Updated 2 years ago
- Pytorch Implementation of Learning Latent Dynamic Robust Representations for World Models☆26May 11, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆16Dec 5, 2024Updated last year
- Unlock Your Next Favorite Film! Our NLP-powered Movie Recommendation Web App delivers tailored suggestions based on cast, genres, and pro…☆23Oct 15, 2023Updated 2 years ago
- (NeurIPS '22) LISA: Learning Interpretable Skill Abstractions - A framework for unsupervised skill learning using Imitation☆29Feb 22, 2023Updated 3 years ago
- ☆11Nov 18, 2023Updated 2 years ago
- [ECCV 2024] 💐Official implementation of the paper "Diffusion Reward: Learning Rewards via Conditional Video Diffusion"☆121Jul 2, 2024Updated 2 years ago
- Some tutorial programs fiiting differential equation parameters using the DiffEqFlux and sciml packages for Julia.☆14Apr 23, 2020Updated 6 years ago
- ☆11Dec 10, 2020Updated 5 years ago
- Code for the paper "3D FlowMatch Actor: Unified 3D Policy for Single- and Dual-Arm Manipulation"☆41Aug 18, 2025Updated last year
- This docker has Acados, Casadi, neural mpc and ROS Noetic for NVIDIA Jetson. It has been tested on a Jetson Nano and Jetson Orin NX☆13Jun 21, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Bulk and single-cell Multi-Omics ground truth Simulator in R☆13Feb 10, 2026Updated 6 months ago
- Implementation of Reinforce for educational purposes.☆13Jun 12, 2023Updated 3 years ago
- Learning Optimal Policies Through Contact in Differentiable Simulation☆115May 3, 2024Updated 2 years ago
- Learning globally stable dynamical systems policies through imitation. A modification of the original work, focussing on waypoint-based i…☆14Oct 12, 2024Updated last year
- AC-DiT: Adaptive Coordination Diffusion Transformer for Mobile Manipulation☆51Feb 23, 2026Updated 6 months ago
- A lightweight driving simulator, written in Julia.☆19Sep 25, 2024Updated last year
- run tinygrad kernels on esp32☆14Nov 28, 2023Updated 2 years ago