Official implementation of Stackelberg PPO for morphology–control co-design.
☆18Mar 17, 2026Updated 5 months ago
Alternatives and similar repositories for StackelbergPPO
Users that are interested in StackelbergPPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025] Implementation of "FACTS: A Factored State-Space Framework For World Modelling"☆31Jun 2, 2025Updated last year
- Code of the paper "Universal Morphology Control via Contextual Modulation" at ICML 2023☆16Aug 3, 2023Updated 3 years ago
- Official Pytorch Implementation of "Zero-Shot Off-Policy Learning" (ICML 2026)☆25Feb 16, 2026Updated 6 months ago
- [ICLR 2025 Spotlight] Official PyTorch Implementation of "BodyGen: Advancing Towards Efficient Embodiment Co-Design"☆62Oct 21, 2025Updated 10 months ago
- ☆35Mar 26, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code for Scalable Offline Model-Based RL with Action chunking☆34Feb 20, 2026Updated 6 months ago
- [ICLR2026] The implementation of "Structural Prognostic Event Modeling for Multimodal Cancer Survival Analysis"☆18Jun 11, 2026Updated 3 months ago
- Extended code for "Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks" with CIFAR-fs classification☆13Apr 26, 2019Updated 7 years ago
- ☆18Mar 10, 2026Updated 6 months ago
- Implementing the base algorithms of Deep Reinforcement Learning in Python☆22Sep 30, 2023Updated 2 years ago
- The official implementation of Value Flows☆59Feb 27, 2026Updated 6 months ago
- Official PyTorch implementation of "Latent Reasoning in TRMs is Secretly a Policy Improvement Operator" (ICML 2026)☆26May 29, 2026Updated 3 months ago
- Combating Mode Collapse via Manifold Entropy Estimation☆11Apr 21, 2023Updated 3 years ago
- Decoupled Q-Chunking☆75May 3, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The official implementation of the paper "Affective Faces for Goal-Driven Dyadic Communication."☆15Jan 27, 2023Updated 3 years ago
- VC-FB and MC-FB algorithms from "Zero-Shot Reinforcement Learning from Low Quality Data" (NeurIPS 2024)☆29Jan 14, 2025Updated last year
- ☆18Oct 9, 2024Updated last year
- Code for "Routing Manifold Alignment Improves Generalization of Mixture-of-Experts LLMs"☆19Nov 6, 2025Updated 10 months ago
- Getting Started in Imitation Learning☆13Mar 3, 2025Updated last year
- Official Release of Multistep Quasimetric Estimation (MQE)☆20Mar 13, 2026Updated 6 months ago
- Steal the structure of any viral video — beat-map the script with Claude, rewrite it as yours, record with a teleprompter☆31Aug 25, 2026Updated 2 weeks ago
- Official code for the LoG2022 paper -- MSGNN: A Spectral Graph Neural Network Based on a Novel Magnetic Signed Laplacian.☆14Sep 3, 2026Updated last week
- Official code for paper "Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement Learning"☆16Jun 12, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This codebase is to reproduce the results of the paper "Grounded Test-Time Adaptation for LLM Agents".☆20Mar 4, 2026Updated 6 months ago
- vLLM deployment for Unsloth Qwen3.6-35B-A3B-NVFP4-Fast on NVIDIA DGX Spark☆67Jul 29, 2026Updated last month
- GCRL in JAX. Official repository for LEO (ICML 2026).☆32Jun 20, 2026Updated 2 months ago
- ☆22Dec 3, 2025Updated 9 months ago
- Corax: Core RL in JAX☆41Feb 22, 2024Updated 2 years ago
- [ACM MM 2025] ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion Model☆16Nov 13, 2025Updated 10 months ago
- A dataloader, but for JAX☆20May 17, 2024Updated 2 years ago
- Brax Viewer is a real-time, interactive web viewer for monitoring reinforcement learning (RL) policies☆17Nov 28, 2025Updated 9 months ago
- Pointax: PointMaze Environment for JAX☆28Oct 22, 2025Updated 10 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [MICCAI 2023] ECL: Class-Enhancement Contrastive Learning for Long-tailed Skin Lesion Classification☆30Apr 7, 2024Updated 2 years ago
- ☆18Jun 25, 2022Updated 4 years ago
- This is where I write RL related stuff from scratch☆10Dec 15, 2019Updated 6 years ago
- ☆31Apr 5, 2025Updated last year
- ☆14Mar 5, 2024Updated 2 years ago
- Unofficial Implementation of Null-text Inversion (https://arxiv.org/abs/2211.09794)☆12Nov 20, 2022Updated 3 years ago
- Source code for the paper: Hear Both Sides: Efficient Multi-Agent Debate via Diversity-Aware Message Retention☆18Aug 14, 2026Updated 3 weeks ago