Assignments for CS294-112.
☆30Sep 11, 2019Updated 6 years ago
Alternatives and similar repositories for homework
Users that are interested in homework are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementation of MATD3☆13Apr 3, 2020Updated 6 years ago
- V-MPO torch version with DMLab30 and GTrXL☆13Mar 1, 2021Updated 5 years ago
- Related papers for offline reforcement learning (we mainly focus on representation and sequence modeling and conventional offline RL)☆19Apr 21, 2022Updated 4 years ago
- Stochastic Variance Reduction Policy Gradient Estimation☆11Nov 6, 2018Updated 7 years ago
- Official repository of the paper Towards safe human-to-robot handovers of unknown containers, presented at the IEEE International Confere…☆11Nov 28, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- my solutions to Berkley's Deep Reinforcement Learning Class CS294-112.☆17Mar 16, 2020Updated 6 years ago
- ☆20Oct 27, 2025Updated 9 months ago
- DRLib:a Concise Deep Reinforcement Learning Library, Integrating HER, PER and D2SR for Almost Off-Policy RL Algorithms.☆564Apr 2, 2024Updated 2 years ago
- Solve BipedalWalkerHardcore-v2 with TD3☆99May 21, 2023Updated 3 years ago
- Tensorflow implementation of BootstrappedDQN using OpenAI baselines☆19Jan 12, 2021Updated 5 years ago
- ☆18Mar 19, 2019Updated 7 years ago
- Simulation environments for Multi-Objective Reinforcement Learning (MORL)☆17Aug 2, 2022Updated 4 years ago
- A plotter for reinforcement learning (RL)☆238Dec 8, 2021Updated 4 years ago
- Assignments for CS294-112.☆17Jul 13, 2018Updated 8 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆20Aug 15, 2023Updated 3 years ago
- Taming MAML: efficient unbiased meta-reinforcement learning☆30Sep 30, 2022Updated 3 years ago
- Implement many Sparse Reward algorithms in Gym Fetch environment☆89Jul 9, 2020Updated 6 years ago
- Public implementation of "Learning from Suboptimal Demonstration via Self-Supervised Reward Regression" from CoRL'21☆22May 20, 2021Updated 5 years ago
- Gated Transformer Model for Computer Vision☆25Jul 11, 2021Updated 5 years ago
- Codes accompanying the paper "Believe What You See: Implicit Constraint Approach for Offline Multi-Agent Reinforcement Learning" (NeurIPS…☆76Oct 18, 2022Updated 3 years ago
- Repository for our ICML 2019 paper: Curiosity-Bottleneck☆34Nov 21, 2022Updated 3 years ago
- Attention-based Curiosity-driven Exploration in Deep Reinforcement Learning☆29Nov 27, 2019Updated 6 years ago
- ☆24Apr 27, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆26Apr 12, 2018Updated 8 years ago
- My simplest implementations of common ML algorithms☆20Jul 23, 2023Updated 3 years ago
- A collection of Reinforcement Learning implementations with PyTorch☆23Mar 22, 2022Updated 4 years ago
- Multi-Objective Deep Reinforcement Learning☆45Jan 1, 2017Updated 9 years ago
- Code for NeurIPS 2022 paper "Robust offline Reinforcement Learning via Conservative Smoothing"☆23Feb 15, 2023Updated 3 years ago
- Value-Decomposition Multi-Agent Actor-Critics☆42Dec 8, 2022Updated 3 years ago
- ☆14May 30, 2019Updated 7 years ago
- ☆55Feb 28, 2024Updated 2 years ago
- Exact Pareto Optimal solutions for preference based Multi-Objective Optimization☆68Jun 22, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A simple parser libaray for URDF Files. That returns a robot object which can be used to access links, joints, transformation matrices, e…☆16Jan 19, 2022Updated 4 years ago
- ☆13Jul 2, 2025Updated last year
- CartPole-v0 via PPO with GAE, PyTorch☆22Feb 10, 2019Updated 7 years ago
- Generate expert demonstrations; GAIL(Generative Adversarial Imitation Learning); IRL(Inverse Reinforcement Learning)☆32Aug 11, 2021Updated 5 years ago
- ☆12Apr 1, 2025Updated last year
- ☆12Feb 20, 2021Updated 5 years ago
- Setup for Octo and some experiments with the model☆12Apr 11, 2024Updated 2 years ago