☆20Nov 13, 2023Updated 2 years ago
Alternatives and similar repositories for rllib
Users that are interested in rllib are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20Jan 9, 2025Updated last year
- ☆32Nov 13, 2023Updated 2 years ago
- Code for the paper: Causal Action Influence Aware Counterfactual Data Augmentation @ICML2024☆14Jul 19, 2024Updated 2 years ago
- Code for our paper: Online Variational Filtering and Parameter Learning☆20Dec 8, 2021Updated 4 years ago
- Model Primitive Hierarchical Reinforcement Learning☆13Dec 8, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Implementation of the paper "Movement Primitives via Optimization" (Dragan et al., 2016). It includes both the adaptation of trajectories…☆22Jan 28, 2018Updated 8 years ago
- Offline RL algoritms implemented in Stable Baselines3 (pytorch)☆11Dec 7, 2021Updated 4 years ago
- Implementation of the Model-Based Meta-Policy-Optimization (MB-MPO) algorithm☆45Nov 15, 2018Updated 7 years ago
- Python Package for EIT(Electric Impedance Tomography)-like problems using Gauss-Newton method.☆17Nov 5, 2025Updated 10 months ago
- Code for the paper "Harnessing Discrete Representations for Continual Reinforcement Learning"☆16Jun 16, 2024Updated 2 years ago
- My final project submission for the Meta Learning course at BITS Goa (conducted by TCS Research)☆16May 3, 2021Updated 5 years ago
- ☆16Jul 4, 2019Updated 7 years ago
- The Controllable Agent project trains RL Agents able to optimize any reward function specified in real time, without any further learning…☆81Jul 17, 2023Updated 3 years ago
- Thinker project☆16Sep 4, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆21Apr 12, 2024Updated 2 years ago
- Implementation of mutual learning model between VAE and GMM.☆29Oct 8, 2025Updated 11 months ago
- Offline Risk-Averse Actor-Critic (O-RAAC). A model-free RL algorithm for risk-averse RL in a fully offline setting☆37Feb 9, 2021Updated 5 years ago
- Repository for SIGIR'18 paper: "Ranking for Relevance and Display Preferences in Complex Presentation Layouts"☆16Aug 28, 2018Updated 8 years ago
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 6 months ago
- A Simulated Optimal Intrusion Response Game☆21Apr 3, 2022Updated 4 years ago
- Notes for the Neuroscience & AI Reading Course (SEM-I 2020-21) at BITS Pilani Goa Campus☆13Sep 30, 2020Updated 5 years ago
- Perception related packages☆19Dec 18, 2024Updated last year
- 📖The Big-&-Extending-Repository-of-Transformers: Pretrained PyTorch models for Google's BERT, OpenAI GPT & GPT-2, Google/CMU Transformer…☆16Jun 9, 2019Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official code for UnICORNN (ICML 2021)☆28Oct 1, 2021Updated 4 years ago
- Multi-Agent Reinforcement Learning on network-security☆22Apr 12, 2022Updated 4 years ago
- Building blocks for productive research☆73Sep 1, 2026Updated 3 weeks ago
- Low-rank adaptation of large language models (LoRA) for Segment Anything 2.☆18Oct 31, 2024Updated last year
- Scaling safe exploration to vision control☆15Feb 19, 2025Updated last year
- Windy GridWorlds environments compatible with OpenAI gym.☆15Jul 8, 2022Updated 4 years ago
- Code associated with "Anxiety, avoidance, and sequential evaluation"☆17Oct 26, 2021Updated 4 years ago
- Python package for Sparse Linear Regression (SLiR)☆21Jul 16, 2019Updated 7 years ago
- Hierarchical entity typing via multi-level learning to rank☆12Oct 13, 2020Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Dataset generation for NeuralGrasps https://arxiv.org/abs/2207.02959☆24Sep 26, 2024Updated last year
- PyTorch - Implicit Quantile Networks - Quantile Regression - C51☆22Jul 26, 2019Updated 7 years ago
- DNN Node Collection using Inference Helper in ROS2☆13Apr 24, 2022Updated 4 years ago
- ETIP: An element tagging problem for Chinese insurance policy analysis☆13Apr 15, 2019Updated 7 years ago
- ☆30Aug 25, 2022Updated 4 years ago
- 🎾 Multi-Agent Proximal Policy Optimization approach to a competitive reinforcement learning problem☆22Sep 25, 2022Updated 3 years ago
- ☆38May 18, 2021Updated 5 years ago