Implementation of SAC and TD3 based on various RNN and Transformer.
☆32Sep 28, 2024Updated last year
Alternatives and similar repositories for Recurrent-Offpolicy-RL
Users that are interested in Recurrent-Offpolicy-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementations for Offline Preference-Based RL (PbRL) algorithms☆21Mar 24, 2025Updated last year
- ☆16Dec 5, 2024Updated last year
- Predicting for Customers, whether they will buy car insurance or not.☆11Jan 29, 2021Updated 5 years ago
- Minimal RLHF implementation built on top of minGPT.☆32Jul 4, 2024Updated 2 years ago
- Official PyTorch code for "Recurrent Off-policy Baselines for Memory-based Continuous Control" (DeepRL Workshop, NeurIPS 21)☆93Nov 21, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code base for NeurIPS 2022 paper Curriculum Reinforcement Learning using Optimal Transport via Gradual Domain Adaptation.☆11Aug 21, 2023Updated 2 years ago
- Benchmarked implementations of Offline RL Algorithms.☆77Mar 4, 2025Updated last year
- Personalized Client-Edge-Cloud Hierarchical Federated Learning on Non-IID Data☆11Sep 7, 2023Updated 2 years ago
- [NeurIPS 2024] Official code for "Variational Distillation of Diffusion Policies into Mixture of Experts"☆17Dec 7, 2024Updated last year
- When Do Transformers Shine in RL? Decoupling Memory from Credit Assignment, NeurIPS 2023 (oral)☆73Apr 26, 2026Updated 2 months ago
- ☆27Apr 22, 2024Updated 2 years ago
- # Analyzing-Visualizing-Data-PowerBI ☆25Jan 16, 2024Updated 2 years ago
- RLA is a tool for managing your RL experiments automatically☆71Feb 7, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Implementations of Temporal Difference InfoNCE (TD InfoNCE)☆35Nov 13, 2023Updated 2 years ago
- ☆20Oct 27, 2025Updated 8 months ago
- Codebase for the paper "How Crucial is Transformer in Decision Transformer?". Containing experiments on different pendulum tasks and code…☆28Mar 24, 2023Updated 3 years ago
- Public sourcecode for Transformable Gaussian Reward Function for Robot Navigation with Deep Reinforcement Learning☆22Aug 7, 2024Updated last year
- Official Pytorch Implementation of "Zero-Shot Off-Policy Learning" (ICML 2026)☆25Feb 16, 2026Updated 5 months ago
- Implementation of ICLR 2025 paper "Q-Adapter: Customizing Pre-trained LLMs to New Preferences with Forgetting Mitigation"☆18Oct 5, 2024Updated last year
- Official implementation of the ICLR 2021 paper "Differentiable Trust Region Layers for Deep Reinforcement Learning"☆11Aug 23, 2023Updated 2 years ago
- Some notes and solutions to "Machine Learning" authored by Zhi-Hua Zhou☆11Jul 20, 2021Updated 5 years ago
- Simulation system for path planning evaluation☆13Dec 13, 2025Updated 7 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code for MOBILE: Model-Bellman Inconsistency Penalized Offline Policy Optimization☆22Apr 17, 2024Updated 2 years ago
- ☆17Oct 25, 2023Updated 2 years ago
- ☆10Sep 19, 2023Updated 2 years ago
- Author's PyTorch implementation of TD7 for online and offline RL☆169Sep 12, 2023Updated 2 years ago
- ☆10Mar 11, 2024Updated 2 years ago
- This repository combines visual srvoing projected to a null space operator and Nonlinear Model Predictive Control (NMPC). The controller …☆11Jun 21, 2026Updated last month
- Faster RCNN using TensorFlow☆10Jul 31, 2022Updated 3 years ago
- Author's implementation of ReBRAC, a minimalist improvement upon TD3+BC☆63Aug 3, 2023Updated 2 years ago
- ☆11Apr 8, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Author's PyTorch implementation of ICML'23 paper "Policy Regularization with Dataset Constraint for Offline Reinforcement Learning" for D…☆18Nov 8, 2024Updated last year
- Jax-Baseline is a Reinforcement Learning implementation using JAX and Flax/Haiku libraries, mirroring the functionality of Stable-Baselin…☆67Updated this week
- UAV-based path planning for efficient localization of non-uniformly distributed weeds using prior knowledge: A reinforcement-learning app…☆15Jul 1, 2025Updated last year
- Using DDPG agent to control UAV system with energy efficiency☆16Jan 7, 2023Updated 3 years ago
- MATLAB implementation of DQN for a navigation environment☆13Aug 13, 2020Updated 5 years ago
- (NeurIPS 2023) Residual Q-Learning: Offline and Online Policy Customization without Value☆35Mar 29, 2024Updated 2 years ago
- This is code to accompany the paper "Accelerating Exploration with Unlabeled Prior Data".☆26Dec 5, 2023Updated 2 years ago