☆16Aug 15, 2025Updated 11 months ago
Alternatives and similar repositories for mini_crossformer
Users that are interested in mini_crossformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Aug 13, 2025Updated 11 months ago
- Official implementation of A Mixture of Surprises for Unsupervised Reinforcement Learning☆23Nov 16, 2022Updated 3 years ago
- A simple wrapper to analyse and visualise reinforcement learning agents' behaviour in the environment.☆14Jan 8, 2022Updated 4 years ago
- MoDem-V2 combines the sample efficiency of the original MoDem with conservative exploration in order to quickly and safely learn manipula…☆25Apr 1, 2024Updated 2 years ago
- ☆26Apr 16, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆20Nov 13, 2022Updated 3 years ago
- Code for NeurIPS 2022 paper Exploiting Reward Shifting in Value-Based Deep RL☆29Oct 29, 2023Updated 2 years ago
- Slot-TTA shows that test-time adaptation using slot-centric models can improve image segmentation on out-of-distribution examples.☆26Jun 20, 2023Updated 3 years ago
- Training Multiple agents in the same environment to collaborate and compete with each other☆12Dec 1, 2019Updated 6 years ago
- Transformer-based World Models☆91Apr 4, 2023Updated 3 years ago
- suPER is a collaborative multi-agent RL algorithm☆14Jun 11, 2024Updated 2 years ago
- MjLab for RL research on the Digit robots☆20Mar 19, 2026Updated 4 months ago
- Synchronized Curriculum Learning for RL Agents☆123Jul 31, 2026Updated last week
- Learning to Coordinate Manipulation Skills via Skill Behavior Diversification (ICLR 2020)☆49Jun 22, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Robot Testing Framework (RTF)☆19Apr 30, 2021Updated 5 years ago
- Memory Replay with Data Compression (ICLR 2022)☆16Sep 26, 2023Updated 2 years ago
- Code for paper "Successor Uncertainties: Exploration and Uncertainty in Temporal Difference Learning" by David Janz*, Jiri Hron*, Przemys…☆21Feb 24, 2023Updated 3 years ago
- Object-Centric-Representation Library (OCRL): This repo is to explore OCR on various downstream tasks from supervised learning tasks to R…☆12Feb 23, 2024Updated 2 years ago
- Corax: Core RL in JAX☆41Feb 22, 2024Updated 2 years ago
- Starter project configuration for my COMP 303 Software Development course at McGill University☆12Sep 26, 2022Updated 3 years ago
- Modular Single-file Reinfocement Learning Algorithms Library☆38May 16, 2023Updated 3 years ago
- [ICML 2024] The algorithm of Reinforcement Learning with an Assistant Reward Agent (ReLara)☆18Aug 2, 2024Updated 2 years ago
- A python implementation of Tangled Program Graphs. A Reinforcement Learning Algorithm.☆20Dec 16, 2021Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Installing Sensable Phantom devices in Linux.☆10Jul 5, 2019Updated 7 years ago
- [CoRL 2021] A robotics benchmark for cross-embodiment imitation.☆60Oct 4, 2023Updated 2 years ago
- ☆13Nov 1, 2023Updated 2 years ago
- Official PyTorch implementation of "Entity-Centric Reinforcement Learning for Object Manipulation from Pixels", Haramati et al., ICLR 202…☆32Feb 22, 2026Updated 5 months ago
- ☆51Nov 20, 2025Updated 8 months ago
- Elrond NFT minting platform POC (Also check out: www.elven.tools)☆12Aug 10, 2023Updated 3 years ago
- ☆11Nov 18, 2023Updated 2 years ago
- Code to reproduce results in the paper "Learning to Predict Navigational Patterns from Partial Observations" (RA-L 2023)☆12Jun 30, 2023Updated 3 years ago
- ☆34Jun 9, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ELIGN: Expectation Alignment as a Multi-agent Intrinsic Reward☆20Dec 5, 2022Updated 3 years ago
- Improving Token-Based World Models with Parallel Observation Prediction (ICML 2024)☆14Feb 23, 2026Updated 5 months ago
- Unofficial implementation for Sigmoid Loss for Language Image Pre-Training☆11Sep 26, 2023Updated 2 years ago
- The implementation of ICLR 2023 paper "Discovering Generalizable Multi-agent Coordination Skills from Multi-task Offline Data".☆45Oct 31, 2024Updated last year
- solving ml10☆26Nov 10, 2023Updated 2 years ago
- I2Q: A Fully Decentralized Q-Learning Algorithm☆19Nov 10, 2022Updated 3 years ago
- ☆26Jun 22, 2022Updated 4 years ago