This project includes code for using the AsyncWebRL and WebGym frameworks to train web agent models.
☆49Jun 9, 2026Updated 2 months ago
Alternatives and similar repositories for webgym
Users that are interested in webgym are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for paper OpenWebRL: Online Multi-Turn Reinforcement Learning for Visual Web Agents☆43Aug 17, 2026Updated last week
- [ACL'25 (Findings)] Explorer: Scaling Exploration-driven Web Trajectory Synthesis for Multimodal Web Agents☆29Feb 17, 2026Updated 6 months ago
- ☆20Sep 16, 2025Updated 11 months ago
- Throughput-oriented multi-turn inference engine for KernelBench [ICML '25]☆24May 27, 2025Updated last year
- ☆75Jun 10, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The official repo of VideoAgentTrek☆59Oct 24, 2025Updated 10 months ago
- ☆23Jun 10, 2025Updated last year
- EvoCUA: Evolving Computer Use Agent☆343Mar 31, 2026Updated 5 months ago
- SimKO: Simple Pass@K Policy Optimization☆31Oct 24, 2025Updated 10 months ago
- ☆30Sep 4, 2025Updated 11 months ago
- Code for Diagnosing Bottlenecks in Deep Q-learning. Contains implementations of tabular environments plus solvers.☆17May 14, 2019Updated 7 years ago
- [JMLR] Gradual Domain Adaptation: Theory and Algorithms☆11Jan 14, 2025Updated last year
- Train a tiny LLaMA model from scratch to repeat your words using Reinforcement Learning from Human Feedback (RLHF)☆18May 23, 2024Updated 2 years ago
- Scalable pipeline for synthesizing verifiable RLVR training data for computer-use agents☆193Aug 13, 2026Updated 2 weeks ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [NeurIPS 2025 Spotlight] Official repository for "Web-Shepherd: Advancing PRMs for Reinforcing Web Agents"☆59May 21, 2025Updated last year
- PhoneHarness runtime harness for mixed-action phone agents☆47Jun 17, 2026Updated 2 months ago
- [NeurIPS 2025 Spotlight] Scaling Computer-Use Grounding via UI Decomposition and Synthesis☆177Jun 18, 2026Updated 2 months ago
- ☆19Jun 5, 2018Updated 8 years ago
- ☆56Nov 26, 2024Updated last year
- ☆24Feb 16, 2022Updated 4 years ago
- The model, data and code for OpenMobile☆49Jul 9, 2026Updated last month
- ☆70Mar 6, 2025Updated last year
- Silly twitter torch implementations.☆48Oct 14, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICLR 2026] Meta-RL Induces Exploration in Language Agents☆46Feb 1, 2026Updated 6 months ago
- Orchard: An Open-Source Agentic Modeling Framework☆494Jul 31, 2026Updated last month
- code of IJCAI submission "Soft Hindsight Experience Replay"☆13Mar 23, 2020Updated 6 years ago
- SVIP: Towards Verifiable Inference of Open-Source Large Language Models☆15Jun 3, 2025Updated last year
- [CVPR 2026] Ego2Web: A Web Agent Benchmark Grounded in Egocentric Videos☆29Mar 25, 2026Updated 5 months ago
- A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models☆29Nov 25, 2024Updated last year
- This is the repository for paper EscapeBench: Pushing Language Models to Think Outside the Box☆18Dec 19, 2024Updated last year
- ☆10Oct 8, 2021Updated 4 years ago
- This repository contains the implementation of the PTR algorithm described in the paper: Pre-Training for Robots: Leveraging Diverse Mult…☆32Oct 26, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- An approach to utomatically generating browser environment with verifiable tasks☆68Mar 24, 2026Updated 5 months ago
- ☆25Nov 30, 2020Updated 5 years ago
- There are my Pytorch codes for charactering adversarial subspace using local intrinsic dimensionality.☆13Apr 26, 2022Updated 4 years ago
- 🕵 Code for our EMNLP 2025 Main paper: "FlashAdventure: A Benchmark for GUI Agents Solving Full Story Arcs in Diverse Adventure Games"☆27Apr 26, 2026Updated 4 months ago
- A very hacky set of functions for getting plotly to do what I want when doing mech interp research, designed to be compatible with PyTorc…☆15Jun 16, 2023Updated 3 years ago
- Awesome GUI Agent Paper List☆896Aug 17, 2026Updated last week
- Learning from preferences is a common paradigm for fine-tuning language models. Yet, many algorithmic design decisions come into play. Ou…☆32Apr 20, 2024Updated 2 years ago