ππͺ BrowserGym, a Gym environment for web task automation
β1,394Oct 5, 2026Updated this week
Alternatives and similar repositories for BrowserGym
Users that are interested in BrowserGym are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AgentLab: An open-source framework for developing, testing, and benchmarking web agents on diverse tasks, designed for scalability and reβ¦β646Updated this week
- WorkArena: How Capable are Web Agents at Solving Common Knowledge Work Tasks?β274Sep 28, 2026Updated last week
- Code repo for "WebArena: A Realistic Web Environment for Building Autonomous Agents"β1,620Nov 26, 2025Updated 10 months ago
- VisualWebArena is a benchmark for multimodal agents.β488Nov 9, 2024Updated last year
- Standardize benchmark wrapping so the community can wrap various otherwise-incompatible benchmarks uniformly and use them everywhere.β55Sep 30, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environmentsβ3,182Sep 14, 2026Updated 3 weeks ago
- TapeAgents is a framework that facilitates all stages of the LLM Agent development lifecycleβ318Dec 16, 2025Updated 9 months ago
- Drive OSS standards and tools for data curation and evaluation creation for state of the art AI agentsβ57Sep 30, 2026Updated last week
- An Illusion of Progress? Assessing the Current State of Web Agentsβ205Jun 25, 2026Updated 3 months ago
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.β434Aug 5, 2026Updated 2 months ago
- A verified version of the WebArena Benchmarkβ62Mar 8, 2026Updated 7 months ago
- Building Open LLM Web Agents with Self-Evolving Online Curriculum RLβ541Jun 6, 2025Updated last year
- An agent benchmark with tasks in a simulated software company.β792Nov 17, 2025Updated 10 months ago
- Awesome GUI Agent Paper Listβ908Updated this week
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- WebLINX is a benchmark for building web navigation agents with conversational capabilitiesβ164Aug 16, 2026Updated last month
- [NeurIPS'23 Spotlight] "Mind2Web: Towards a Generalist Agent for the Web" -- the first LLM-based web agent and benchmark for generalist wβ¦β1,031Nov 5, 2025Updated 11 months ago
- AWM: Agent Workflow Memoryβ477Dec 22, 2025Updated 9 months ago
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]β748Jul 29, 2025Updated last year
- Code for "WebVoyager: WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models"β1,126Mar 4, 2024Updated 2 years ago
- β41Jul 21, 2024Updated 2 years ago
- Code for the paper π³ Tree Search for Language Model Agentsβ225Jul 25, 2024Updated 2 years ago
- Setup scripts for the WebArena benchmarkβ22Jun 19, 2025Updated last year
- DoomArena is a Framework for Testing AI Agents Against Evolving Security Threatsβ64Sep 12, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML'24] SeeAct is a system for generalist web agents that autonomously carry out tasks on any given website, with a focus on large multβ¦β852Feb 3, 2025Updated last year
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,397Updated this week
- [NeurIPS 2022] πWebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agentsβ602Sep 6, 2024Updated 2 years ago
- COLM2026β39Jul 9, 2026Updated 3 months ago
- [NeurIPS'25 D&B] Mind2Web-2 Benchmark: Evaluating Agentic Search with Agent-as-a-Judgeβ114Sep 25, 2026Updated 2 weeks ago
- Democratizing Reinforcement Learning for LLMsβ5,860Updated this week
- OS-ATLAS: A Foundation Action Model For Generalist GUI Agentsβ455Apr 20, 2025Updated last year
- All-in-one Web Agent framework for post-training. Start building with a few clicks!β281Jul 7, 2025Updated last year
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ23,805Updated this week
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- β117Jul 2, 2024Updated 2 years ago
- Windows Agent Arena (WAA) πͺ is a scalable OS platform for testing and benchmarking of multi-modal AI agents.β907Apr 13, 2026Updated 5 months ago
- A Benchmark for Evaluating Safety and Trustworthiness in Web Agents for Enterprise Scenariosβ29Mar 12, 2026Updated 6 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnosticsβ2,818Aug 23, 2026Updated last month
- Code and implementations for the ACL 2025 paper "AgentGym: Evolving Large Language Model-based Agents across Diverse Environments" by Zhiβ¦β850May 30, 2026Updated 4 months ago
- π» A curated list of papers and resources for multi-modal Graphical User Interface (GUI) agents.β1,217Aug 17, 2025Updated last year
- [ICLR'25 Oral] UGround: Universal GUI Visual Grounding for GUI Agentsβ321Aug 24, 2026Updated last month