Official implementation of the paper "Learning and Planning Multi-Agent Tasks via a MoE-based World Model"
☆33Mar 1, 2026Updated 5 months ago
Alternatives and similar repositories for m3w-marl
Users that are interested in m3w-marl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- a novel algo for meta-MARL; 元-多智能体强化学习算法☆24Mar 1, 2026Updated 5 months ago
- Official Implementation of "Decentralized Transformers with Centralized Aggregation are Sample-Efficient Multi-Agent World Models" (accep…☆19Dec 14, 2025Updated 7 months ago
- [ECCV 2026]☆59Jun 18, 2026Updated last month
- ☆17Mar 10, 2025Updated last year
- 无人机动态覆盖控制;1. 实现了一个无人机点覆盖环境;2. 给出了无人机连通保持规则;3. 给出了基于MARL的控制算法☆94Sep 16, 2025Updated 10 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆20Mar 2, 2026Updated 5 months ago
- ☆12Aug 4, 2023Updated 3 years ago
- Code of the paper: Debiasing Meta-Gradient Reinforcement Learning by Learning the Outer Value Function☆13Apr 13, 2026Updated 3 months ago
- MENTOR is a highly efficient visual RL algorithm that excels in both simulation and real-world complex robotic learning tasks.☆28Jul 9, 2025Updated last year
- ☆26Feb 21, 2022Updated 4 years ago
- [IROS 2025 oral] Official implementation of NOLO: Navigate Only Look Once☆22Nov 13, 2025Updated 8 months ago
- [NeurIPS 2022] Symmetry Teleportation for Accelerated Optimization☆12Nov 10, 2022Updated 3 years ago
- Official implementation of SLAC☆20Jul 17, 2026Updated 3 weeks ago
- ☆34Nov 6, 2025Updated 9 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- add attention mechanism in InvertedResidual block about shuffleNetV2☆10Mar 2, 2024Updated 2 years ago
- Datasets with baselines for Offline MARL.☆225Nov 2, 2025Updated 9 months ago
- Path Plan for Delivery Unmanned Aerial Vehicle☆16Jul 6, 2023Updated 3 years ago
- ☆24Jan 23, 2026Updated 6 months ago
- [ECCV 2026] Promsa: Progressive Multimodal Search Agents for Knowledge-Based Visual Question Answering☆91Jul 7, 2026Updated last month
- Model-based Offline Policy Optimization re-implement all by pytorch☆43Sep 13, 2023Updated 2 years ago
- Code for Scalable Offline Model-Based RL with Action chunking☆31Feb 20, 2026Updated 5 months ago
- Single Cycle and Pipeline CPU of RISC-V Architecture designed for Digital Design and Computer Organization Experiments 2021, NJU☆14Jan 17, 2022Updated 4 years ago
- LAMARL: LLM-Aided Multi-Agent Reinforcement Learning for Cooperative Policy Generation☆52Jul 19, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆22Feb 4, 2026Updated 6 months ago
- UHD is a surface electromyogram (sEMG) signals database, including original sEMG signals, the starting points and the termination points …☆12Dec 4, 2019Updated 6 years ago
- The official repository for the paper "Real-world Reinforcement Learning from Suboptimal Interventions”.☆62Jul 29, 2026Updated 2 weeks ago
- A collection of matrix games in JAX☆14Apr 13, 2026Updated 3 months ago
- A benchmark for evaluating reinforcement learning algorithms that train the policies using imaginary rollouts from LLMs.☆15Nov 4, 2025Updated 9 months ago
- Github Repo for CARL: Cautious Adaptation for RL in Safety Critical Settings☆14Nov 22, 2022Updated 3 years ago
- Official implementation of Lookahead Exploration with Neural Radiance Representation for Continuous Vision-Language Navigation (CVPR'24 H…☆109Apr 2, 2025Updated last year
- Dual Contrastive Learning for Few-shot Medical Image Segmentation☆28Mar 2, 2023Updated 3 years ago
- Research Resources on Human-Machine Collaborative Sensing by AI-Driven Unmanned Vehicles(边缘环境人机协同群智感知与决策关键技术研究)☆18Jul 2, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICCV 2025] Official implementation of Open-World Skill Discovery from Unsegmented Demonstration Videos☆17Sep 4, 2025Updated 11 months ago
- Mastering Massive Multi-Task Reinforcement Learning via Mixture-of-Expert Decision Transformer☆22Sep 18, 2025Updated 10 months ago
- Sim-to-real RL for in-hand cube rotation with the LEAP Hand, built on Mjlab.☆33Feb 21, 2026Updated 5 months ago
- Coarse-to-fine Q-Network☆59Aug 6, 2024Updated 2 years ago
- Simple implementation of simulator environments for sim2sim, with a wrapped ROS2 environment for Unitree robots' real deployment with sam…☆18Jan 20, 2026Updated 6 months ago
- [CVPR 2025] Source codes for the paper "Collaborative Tree Search for Enhancing Embodied Multi-Agent Collaboration"☆17Nov 10, 2025Updated 9 months ago
- ☆18Jul 14, 2023Updated 3 years ago