[AutoML'22] Bayesian Generational Population-based Training (BG-PBT)
☆31Sep 16, 2022Updated 4 years ago
Alternatives and similar repositories for bgpbt
Users that are interested in bgpbt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23Aug 19, 2022Updated 4 years ago
- Release code for ICML2020 Knowing The What But Not The Where in Bayesian Optimization☆15Mar 7, 2023Updated 3 years ago
- Toolkit of Causal Model-based Reinforcement Learning.☆33Jun 5, 2023Updated 3 years ago
- Code-base for the paper Spectral Normalisation for Deep Reinforcement Learning: An Optimisation Perspective.☆11Jun 26, 2021Updated 5 years ago
- Vectorization techniques for fast population-based training.☆58Apr 26, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Challenges and Opportunities in Offline Reinforcement Learning from Visual Observations☆116Apr 16, 2026Updated 5 months ago
- Official Codebase for Offline Reinforcement Learning from Images with Latent Space Models☆31Apr 30, 2021Updated 5 years ago
- ☆16Jul 16, 2024Updated 2 years ago
- 📴 OffCon^3: SOTA PyTorch SAC and TD3 Implementations (arxiv: 2101.11331)☆25Jun 20, 2021Updated 5 years ago
- Survival of the Most Influential Prompts: Efficient Black-Box Prompt Search via Clustering and Pruning (Zhou et al.; EMNLP 2023 Findings)☆17Feb 17, 2024Updated 2 years ago
- A2C is a special case of PPO!☆23May 20, 2022Updated 4 years ago
- ☆20Jun 14, 2022Updated 4 years ago
- Official implementation of NeurIPS22 paper “Multi-agent Dynamic Algorithm Configuration”☆26Mar 6, 2023Updated 3 years ago
- Policy Transfer across Visual and Dynamics Domain Gaps via Iterative Grounding (RSS 2021)☆12Oct 22, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 🧶 Minimal PyTorch Soft Actor Critic (SAC) implementation☆39Feb 19, 2022Updated 4 years ago
- The code for the paper *The Sensitivity of Counterfactual Fairness to Unmeasured Confounding* @ UAI 2019☆14Apr 4, 2020Updated 6 years ago
- Implementation of Tactical Optimistic and Pessimistic value estimation☆25Jul 18, 2023Updated 3 years ago
- Docker containers of baseline agents for the Crafter environment☆30Dec 14, 2021Updated 4 years ago
- Notebooks for managing NeurIPS 2014 and analysing the NeurIPS experiment.☆13May 22, 2024Updated 2 years ago
- This repository showcases how to use the DynamixelSDK C++ and Python APIs to control an Interbotix XSeries Arm.☆18Apr 28, 2021Updated 5 years ago
- Learning Laplacian Representations in Reinforcement Learning☆18Jan 2, 2021Updated 5 years ago
- Temporally Correlated Episodic Reinforcement Learning, ICLR 24☆12Apr 8, 2024Updated 2 years ago
- A project copied from google-research which named motion-imitation was rewrited with PyTorch☆10Sep 30, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML'24] Tackling Prevalent Conditions in Unsupervised Combinatorial Optimization: Cardinality, Minimum, Covering, and More☆14Jul 12, 2024Updated 2 years ago
- Official code for ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning (AAAI'24)☆16Feb 10, 2024Updated 2 years ago
- Quadruped Robot controller design and simulation on Webots☆12Apr 28, 2020Updated 6 years ago
- Official Implementation of NeurIPS'23 Paper "Cross-Episodic Curriculum for Transformer Agents"☆33Oct 12, 2023Updated 2 years ago
- ☆22Nov 8, 2021Updated 4 years ago
- ♊ Minimal PyTorch Twin Delayed DDPG (TD3) implementation☆10Jun 20, 2021Updated 5 years ago
- Official Implementation of `An Optimisation Framework for Unsupervised Environment Design` from RLC 2025☆17Nov 24, 2025Updated 9 months ago
- Scalable data valuation using optimal transport (ICLR 2025)☆14Jul 15, 2025Updated last year
- [TPAMI 2023] Learning Symbolic Model-Agnostic Loss Functions via Meta-Learning. Paper Link: https://arxiv.org/abs/2209.08907☆16Aug 5, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Codes accompanying the paper "Offline Reinforcement Learning with Value-Based Episodic Memory" (ICLR 2022 https://arxiv.org/abs/2110.0979…☆15Mar 9, 2022Updated 4 years ago
- Author's PyTorch implementation of ICML'23 paper "Policy Regularization with Dataset Constraint for Offline Reinforcement Learning" for D…☆17Nov 8, 2024Updated last year
- Test-Time Label-Shift Adaptation☆14May 24, 2023Updated 3 years ago
- ☆21Sep 14, 2026Updated last week
- ☆35Sep 14, 2021Updated 5 years ago
- Determinantal Point Processes in Julia☆12Feb 20, 2020Updated 6 years ago
- Codes for DATA: Differentiable ArchiTecture Approximation.☆11Jul 22, 2021Updated 5 years ago