Adopting reasonable strategies is challenging but crucial for an intelligent agent with limited resources working in hazardous, unstructured, and dynamic environments to improve the system utility, decrease the overall cost, and increase mission success probability. Deep Reinforcement Learning (DRL) helps organize agents' behaviors and actions b…
☆13Dec 28, 2022Updated 3 years ago
Alternatives and similar repositories for Bayesian-Soft-Actor-Critic
Users that are interested in Bayesian-Soft-Actor-Critic are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Bayesian Soft Actor Critic☆16Jan 6, 2023Updated 3 years ago
- The test code for the paper "Attention-based advantage actor-critic algorithm with prioritized experience replay for complex 2-D robotic …☆10Aug 7, 2022Updated 3 years ago
- 明朝那些事儿☆11May 31, 2022Updated 4 years ago
- Source code for "Congestion-aware Distributed Task Offloading in Wireless Multi-hop Networks Using Graph Neural Networks"☆14Oct 23, 2024Updated last year
- Seq-HGNN: Learning Sequential Node Representation on Heterogeneous Graph☆12Aug 2, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Explainability of Deep RL algorithms using graph networks and layer-wise relevance propagation.☆12Aug 20, 2024Updated last year
- Heterogeneous Causal Metapath Graph Neural Network for Gene-Microbe-Disease Association Prediction☆12Aug 19, 2024Updated last year
- ☆19Mar 27, 2025Updated last year
- ☆14Sep 6, 2023Updated 2 years ago
- A Multiplicative Value Function for Safe and Efficient Reinforcement Learning. IROS 2023.☆24Sep 24, 2023Updated 2 years ago
- Efficient Exploration through Bayesian Deep-Q Networks.☆18Mar 22, 2022Updated 4 years ago
- The purpose of this project is to implement machine learning methods to study resource allocation problems, that is how to share limited …☆16Jun 7, 2022Updated 4 years ago
- [ACCV 2022] Information and scripts for the Apron Dataset☆13Dec 21, 2022Updated 3 years ago
- This repository accompanies the following paper: A Workflow for Offline Model-Free Robotic RL☆13Nov 5, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆15Jan 11, 2024Updated 2 years ago
- Rigid body model of a simple humanoid robot. Model available : urdf + srdf☆10Nov 10, 2025Updated 8 months ago
- Project explores collaboration capabilities of VDN and IQL agents on a custom MARL Food Collector environment☆11Apr 6, 2022Updated 4 years ago
- Control barrier function based motion planning☆18Apr 4, 2020Updated 6 years ago
- Code for the CoRL 2019 paper AC-Teach: A Bayesian Actor-Critic Method for Policy Learning with an Ensemble of Suboptimal Teachers☆24Feb 15, 2023Updated 3 years ago
- Slither-in Inspired Snake Environment for OpenAI Gym (Part of Requests for Research 2.0)☆12Mar 5, 2018Updated 8 years ago
- Codex Skill fork of fast-context-mcp that calls Windsurf Devstral without an MCP server☆20Jun 5, 2026Updated last month
- [ICRA 2025]Robust Self-Reconfiguration for Fault-Tolerant Control of Modular Aerial Robot Systems☆28Jun 9, 2025Updated last year
- Code for Paper "Scaling Law for Large Wireless Models", accepted by AAAI 2026☆15Nov 19, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Multi-objective reinforcement learning for covid-19 control☆12Aug 12, 2021Updated 4 years ago
- Robust-Trajectory-Tracking-for-Quadrotor-UAVs-using-Sliding-Mode-Control☆17Mar 24, 2024Updated 2 years ago
- Source code for the IROS21 paper Efficient Task Planning for Mobile Manipulation: a Virtual Kinematic Chain Perspective☆11Aug 2, 2021Updated 4 years ago
- OpenAI gym environments for goal-conditioned and language-conditioned reinforcement learning☆14Jan 27, 2026Updated 5 months ago
- Code for AAAI Workshop WMAC "Paper Simulating Rumor Spreading in Social Networks using LLM agents"☆13Feb 20, 2025Updated last year
- Novel Vehicular Edge and Fog enabled Computation offloading framework☆16Sep 7, 2022Updated 3 years ago
- ☆18Mar 14, 2026Updated 4 months ago
- [IEEE TPAMI] A Framework for Constrained Multi-Objective Reinforcement Learning☆19Apr 18, 2025Updated last year
- Autonomous feature discovery for tabular GLM models — a minimal fork of Karpathy’s autoresearch.☆15Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Notes for paper reading.☆10Jun 22, 2026Updated last month
- Closed-loop simulator of complex behavior and learning based on reinforcement learning and deep neural networks☆15Mar 20, 2026Updated 4 months ago
- Learned User Representations in Online Social Networks (Twitter) using Temporal Dynamics of Information Diffusion.☆10Oct 15, 2018Updated 7 years ago
- tokio_codec for Session Initiation Protocol (SIP)☆12Apr 14, 2019Updated 7 years ago
- reproduce paper "Energy-Efficient Multi-UAV-Enabled Multiaccess Edge Computing Incorporating NOMA"☆17Apr 13, 2024Updated 2 years ago
- Learning Long-Horizon Robot Exploration Strategies for Multi-Object Search in Continuous Action Spaces. http://multi-object-search.cs.uni…☆14Nov 29, 2022Updated 3 years ago
- 🌌 Applications of Physics-Informed ML: A collection of notebooks from my Masters research, exploring how machine learning can solve scie…☆12Apr 29, 2026Updated 2 months ago