Soft-QMIX: Integrating Maximum Entropy For Monotonic Value Function Factorization
☆15Jul 3, 2024Updated 2 years ago
Alternatives and similar repositories for Soft-QMIX
Users that are interested in Soft-QMIX are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- repo for TMLR'26 paper "Reconciling In-Context and In-Weight Learning via Dual Representation Space Encoding"☆25Mar 8, 2026Updated 4 months ago
- Learning-based agent for Google Research Football (足球游戏智能体)☆120Apr 20, 2023Updated 3 years ago
- ☆15Jun 2, 2024Updated 2 years ago
- The python code is provided for distribution-auction algorithm with connected networks and federal topology networks.☆12Dec 6, 2023Updated 2 years ago
- Fine-tuned MARL algorithms on SMAC (100% win rates on most scenarios)☆19Aug 20, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆21Aug 16, 2024Updated last year
- Algorithm that combines QMIX with SAC for Multi-Agent Reinforcement Learning.☆59May 20, 2022Updated 4 years ago
- ☆16Jun 15, 2026Updated last month
- 2018 Master's Project: Program to do online 3D reconstruction with pose correction of UAV. High localization and object dimensional accur…☆13Oct 20, 2018Updated 7 years ago
- [ICML' 24] The PyTorch implementation of our paper: "Individual Contributions as Intrinsic Exploration Scaffolds for Multi-agent Reinforc…☆25May 29, 2024Updated 2 years ago
- [NeurIPS'23] "ProBio: A Protocol-guided Multimodal Dataset for Molecular Biology Lab"☆16Dec 13, 2023Updated 2 years ago
- curriculum☆27Feb 7, 2023Updated 3 years ago
- Library for line coverage and arc rouing for single and multiple robots☆11Mar 27, 2023Updated 3 years ago
- ☆26Feb 20, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆43Aug 8, 2021Updated 4 years ago
- ZJU Robotics project of differential drive car path planning and trajectory planning based on the Client simulation platform (my freshman…☆10Dec 2, 2020Updated 5 years ago
- Multi UAV navigation using Deep Reinforcement Learning within a Multi TSP☆11Nov 29, 2024Updated last year
- Sirius, an efficient correction mechanism, which significantly boosts Contextual Sparsity models on reasoning tasks while maintaining its…☆21Sep 10, 2024Updated last year
- Code accompanying the ICML'24 paper "Feature Contamination: Neural Networks Learn Uncorrelated Features and Fail to Generalize"☆22Feb 13, 2025Updated last year
- ☆25Feb 21, 2022Updated 4 years ago
- TransMix: Transformer-based Value Function Decomposition for Cooperative Multi-agent Reinforcement Learning☆11Oct 18, 2022Updated 3 years ago
- 清华大学生物,医学,药学等相关专业的毕业论文latex模板。也适用于其他专业。适合本硕博毕业论文和博后报告。本模板在tuna协会的thuthesis项目基础上,增补了和生医药相关同学的内容,也增添了对latex新手更加友好的注释。☆27Sep 14, 2023Updated 2 years ago
- ☆43May 3, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- This is the official code for our paper entitled "Dynamic Deep Factor Graph for Multi-Agent Reinforcement Learning".☆10Updated this week
- ☆10Jun 13, 2025Updated last year
- Pytorch implementations of the multi-agent reinforcement learning algorithms, including QMIX, VDN, COMA, MADDPG, MATD3, FACMAC and MASoft…☆56Mar 10, 2025Updated last year
- This repository, "Autonomous Driving System On Various Platforms", details the exploration and implementation of autonomous driving syste…☆10Aug 16, 2021Updated 4 years ago
- Implementation of CBAA and CBBA algorithms as described by Choi, Brunet, How☆10Sep 11, 2021Updated 4 years ago
- Reference code for the paper ""Centroid-Guided Target-Driven Topology Control Method for UAV Ad-Hoc Networks Based on Tiny Deep Reinforce…☆13Oct 21, 2024Updated last year
- Qwen-Image's DiT inference with TensorRT-10☆21Oct 13, 2025Updated 9 months ago
- ☆18Oct 6, 2022Updated 3 years ago
- Code and files from a project regarding UAV path planning in a SAR situation. The project was done for the 8th semester of the Operations…☆10Dec 8, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆17Oct 25, 2023Updated 2 years ago
- ☆11Dec 5, 2020Updated 5 years ago
- PyTorch使用技巧和教程☆12Apr 17, 2023Updated 3 years ago
- path coverage algorithm for drones using reinforcement learning☆13Aug 13, 2020Updated 5 years ago
- Cooperative target search, UAV swarm, Path planning☆15Mar 13, 2025Updated last year
- The AI Arena: A framework for distributed multi-agent reinforcement learning☆14Aug 5, 2022Updated 3 years ago
- ☆10Jun 22, 2020Updated 6 years ago