Official code for the paper: Continual Task Allocation in Meta-Policy Network via Sparse Prompting
☆23Feb 10, 2025Updated last year
Alternatives and similar repositories for CoTASP
Users that are interested in CoTASP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Source code of the ICML24 paper "Self-Composing Policies for Scalable Continual Reinforcement Learning" (selected for oral presentation)☆29Jul 20, 2024Updated 2 years ago
- Code to reproduce results from the paper: Prediction and Control in Continual Reinforcement Learning, NeurIPS 2023.☆13May 10, 2024Updated 2 years ago
- Code for Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning, AAAI 2025☆15Dec 19, 2024Updated last year
- Continual reinforcement learning baselines: experiment specifications, implementation of existing methods, and common metrics. Easily ext…☆137Jul 6, 2023Updated 3 years ago
- Codebase for the paper Aerial Diffusion: Text Guided Ground-to-Aerial View Translation from a Single Image using Diffusion Models☆13Oct 3, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- NLPBench: Evaluating NLP-Related Problem-solving Ability in Large Language Models☆10Oct 27, 2023Updated 2 years ago
- [ICML2022] "Identity-Disentangled Adversarial Augmentation for Self-Supervised Learning"☆10Jul 24, 2022Updated 4 years ago
- Implementation for ACL 2024 paper "Meta-Task Prompting Elicits Embeddings from Large Language Models"☆12Jul 25, 2024Updated 2 years ago
- Coverage path planning under wind conditions☆11Feb 29, 2020Updated 6 years ago
- Multi Agent Task sharing implementation using RRT algorithm. Implementation in MatLab☆12Oct 18, 2016Updated 9 years ago
- (ICML 2024) The official code for Value-Evolutionary-Based Reinforcement Learning☆20Jul 2, 2024Updated 2 years ago
- UR5 and robotiq85_gripper in gazebo☆10Apr 6, 2022Updated 4 years ago
- Matlab simulation of a fleet of terrestrial vehicles performing formation control and obstacle avoidance☆13Jul 26, 2022Updated 3 years ago
- Code for "Demonstration-free Autonomous Reinforcement Learning via Implicit and Bidirectional Curriculum" (ICML 2023)☆10Jul 6, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for "Routing Manifold Alignment Improves Generalization of Mixture-of-Experts LLMs"☆19Nov 6, 2025Updated 8 months ago
- ☆11Apr 20, 2021Updated 5 years ago
- Version 3.0.0 Pytorch implementations of DQN, DDQN, DDPG, SAC, Discrete SAC. With more features :)☆12Feb 16, 2023Updated 3 years ago
- This is a project using Pytorch to fulfill reinforcement learning on a simple game - Gridworld☆14Jul 13, 2020Updated 6 years ago
- Simulation of car parking in different parking lots using Unity ML-Agents☆13Dec 16, 2023Updated 2 years ago
- 四种启发式算法(模拟退火、遗传算法、禁忌搜索、蚁群算法)解决TSP(旅行商问题)实例☆15Dec 27, 2019Updated 6 years ago
- [ACL'25] Mosaic-IT: Cost-Free Compositional Data Synthesis for Instruction Tuning☆20Sep 27, 2025Updated 9 months ago
- Improvements made to pietrolechthaler's and his group project titled: "UR5 Pick and Place Simulation in Ros/Gazebo", available in the nex…☆11May 10, 2023Updated 3 years ago
- The official PyTorch code for AAAI'23 Paper "Sparse Coding in a Dual Memory System for Lifelong Learning"☆12Feb 15, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients☆20Jun 17, 2025Updated last year
- AutoHallusion Codebase (EMNLP 2024)☆23Dec 6, 2024Updated last year
- ☆14Oct 23, 2025Updated 9 months ago
- [COLM 2025] "C3PO: Critical-Layer, Core-Expert, Collaborative Pathway Optimization for Test-Time Expert Re-Mixing"☆21Apr 9, 2025Updated last year
- The implementation of Discriminator Soft Actor Critic☆15Jan 25, 2020Updated 6 years ago
- This repository includes code for our paper: ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning…☆15May 2, 2026Updated 2 months ago
- Code for the paper "Functional Regularization for Reinforcement Learning via Learned Fourier Features"☆20Oct 2, 2022Updated 3 years ago
- Related papers for offline reforcement learning (we mainly focus on representation and sequence modeling and conventional offline RL)☆18Apr 21, 2022Updated 4 years ago
- ☆13Jan 16, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ROS package to simulate UR5 robot using Gazebo and ROS☆14Jun 15, 2021Updated 5 years ago
- ICLR 2020 Meta Reinforcement Learning with Autonomous Inference of Subtask Dependencies☆18Jul 16, 2020Updated 6 years ago
- (NeurIPS 2025) LaRes: Evolutionary Reinforcement Learning with LLM-based Adaptive Reward Search☆29Mar 9, 2026Updated 4 months ago
- [UR 2023] Robust Route Planning with Distributional Reinforcement Learning in a Stochastic Road Network Environment☆22Jun 19, 2024Updated 2 years ago
- Environments to support https://github.com/sholtodouglas/learning_from_play and reinforcement learning for robotic manipulation.☆21Mar 28, 2021Updated 5 years ago
- (ICML 2024) The official code for EvoRainbow: Combining Improvements in Evolutionary Reinforcement Learning for Policy Search☆37Feb 3, 2026Updated 5 months ago
- OptiDICE: Offline Policy Optimization via Stationary Distribution Correction Estimation☆16Aug 3, 2023Updated 2 years ago