Deep Transformer Q-Networks for Partially Observable Reinforcement Learning
☆176Jul 7, 2024Updated 2 years ago
Alternatives and similar repositories for DTQN
Users that are interested in DTQN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Asymmetric methods for partially observable reinforcement learning☆10Jun 9, 2025Updated last year
- An easy PyTorch implementation of "Stabilizing Transformers for Reinforcement Learning"☆183Feb 21, 2023Updated 3 years ago
- ☆32Dec 1, 2019Updated 6 years ago
- Clean baseline implementation of PPO using an episodic TransformerXL memory☆212Jun 18, 2024Updated 2 years ago
- Gridworld domains in the gym interface☆29Oct 2, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official PyTorch code for "Recurrent Off-policy Baselines for Memory-based Continuous Control" (DeepRL Workshop, NeurIPS 21)☆93Nov 21, 2023Updated 2 years ago
- ☆10Apr 13, 2023Updated 3 years ago
- When Do Transformers Shine in RL? Decoupling Memory from Credit Assignment, NeurIPS 2023 (oral)☆73Apr 26, 2026Updated 3 months ago
- ☆20Jun 25, 2023Updated 3 years ago
- ☆18Jul 10, 2022Updated 4 years ago
- Beamer theme for Northeastern University☆16Oct 22, 2024Updated last year
- Transformers are Meta-Reinforcement Learners - International Conference on Machine Learning (ICML) 2022☆69May 8, 2023Updated 3 years ago
- Transformer in RL for decision-making☆106Feb 13, 2023Updated 3 years ago
- Implementation of Q-Transformer, Scalable Offline Reinforcement Learning via Autoregressive Q-Functions, out of Google Deepmind☆407Jun 20, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Extreme Q-Learning: Max Entropy RL without Entropy☆88Feb 14, 2023Updated 3 years ago
- Official codebase for Decision Transformer: Reinforcement Learning via Sequence Modeling.☆2,823Apr 29, 2024Updated 2 years ago
- Deep recurrent Q learning on CartPole-v1 environment☆96Jan 15, 2024Updated 2 years ago
- Code for Equivariant Transporter Network☆23Apr 17, 2023Updated 3 years ago
- Simple (but often Strong) Baselines for POMDPs in PyTorch, ICML 2022☆347Apr 26, 2026Updated 3 months ago
- A Reinforcement Learning Project using PPO + Transformer☆94Jul 21, 2023Updated 3 years ago
- Official implementation for "Let Offline RL Flow: Training Conservative Agents in the Latent Space of Normalizing Flows", NeurIPS 2022, O…☆12Jan 31, 2023Updated 3 years ago
- The official implementation of "Transformer in Transformer as Backbone for Deep Reinforcement Learning"☆59Dec 27, 2023Updated 2 years ago
- A curated list of Decision Transformer resources (continually updated)☆914May 21, 2026Updated 2 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Implementation of Multi-Game Decision Transformers in PyTorch☆49Feb 11, 2023Updated 3 years ago
- Generalised UDRL☆37May 12, 2022Updated 4 years ago
- [ECCV2022] [T-PAMI] StARformer: Transformer with State-Action-Reward Representations.☆97May 21, 2023Updated 3 years ago
- Single-file SAC-N implementation on jax with flax and equinox. 10x faster than pytorch☆56May 21, 2023Updated 3 years ago
- Code for "Possibility Before Utility: Learning And Using Hierarchical Affordances" (ICLR 2022)☆14Mar 14, 2022Updated 4 years ago
- TransMix: Transformer-based Value Function Decomposition for Cooperative Multi-agent Reinforcement Learning☆11Oct 18, 2022Updated 3 years ago
- [NeurIPS 2022] Open source code for reusing prior computational work in RL.☆100Jul 5, 2023Updated 3 years ago
- ☆12Aug 28, 2020Updated 5 years ago
- Adaptive Attention Span for Reinforcement Learning☆136May 11, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official Release of Multistep Quasimetric Estimation (MQE)☆18Mar 13, 2026Updated 4 months ago
- A System-Oriented Wargame Framework for Adversarial ML☆10Apr 24, 2023Updated 3 years ago
- Minimal implementation of Decision Transformer: Reinforcement Learning via Sequence Modeling in PyTorch for mujoco control tasks in Open…☆294Jun 10, 2022Updated 4 years ago
- Challenges and Opportunities in Offline Reinforcement Learning from Visual Observations☆115Apr 16, 2026Updated 3 months ago
- Fine-tuned MARL algorithms on SMAC (100% win rates on most scenarios)☆12Jul 29, 2023Updated 2 years ago
- krazy grid world☆26Mar 2, 2020Updated 6 years ago
- [ICLR 2024] Closing the Gap between TD Learning and Supervised Learning - A Generalisation Point of View.☆25Apr 19, 2024Updated 2 years ago