ππͺ BrowserGym, a Gym environment for web task automation
β1,335Jul 17, 2026Updated last month
Alternatives and similar repositories for BrowserGym
Users that are interested in BrowserGym are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AgentLab: An open-source framework for developing, testing, and benchmarking web agents on diverse tasks, designed for scalability and reβ¦β627Jul 17, 2026Updated last month
- WorkArena: How Capable are Web Agents at Solving Common Knowledge Work Tasks?β267Apr 25, 2026Updated 4 months ago
- Code repo for "WebArena: A Realistic Web Environment for Building Autonomous Agents"β1,590Nov 26, 2025Updated 9 months ago
- VisualWebArena is a benchmark for multimodal agents.β485Nov 9, 2024Updated last year
- Standardize benchmark wrapping so the community can wrap various otherwise-incompatible benchmarks uniformly and use them everywhere.β53Jul 17, 2026Updated last month
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environmentsβ3,113Updated this week
- TapeAgents is a framework that facilitates all stages of the LLM Agent development lifecycleβ318Dec 16, 2025Updated 8 months ago
- Drive OSS standards and tools for data curation and evaluation creation for state of the art AI agentsβ55Aug 11, 2026Updated 2 weeks ago
- An Illusion of Progress? Assessing the Current State of Web Agentsβ200Jun 25, 2026Updated 2 months ago
- A verified version of the WebArena Benchmarkβ51Mar 8, 2026Updated 5 months ago
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.β433Aug 5, 2026Updated 3 weeks ago
- Building Open LLM Web Agents with Self-Evolving Online Curriculum RLβ538Jun 6, 2025Updated last year
- An agent benchmark with tasks in a simulated software company.β771Nov 17, 2025Updated 9 months ago
- WebLINX is a benchmark for building web navigation agents with conversational capabilitiesβ163Aug 16, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Awesome GUI Agent Paper Listβ896Aug 17, 2026Updated last week
- [NeurIPS'23 Spotlight] "Mind2Web: Towards a Generalist Agent for the Web" -- the first LLM-based web agent and benchmark for generalist wβ¦β1,022Nov 5, 2025Updated 9 months ago
- AWM: Agent Workflow Memoryβ463Dec 22, 2025Updated 8 months ago
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]β728Jul 29, 2025Updated last year
- Code for "WebVoyager: WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models"β1,123Mar 4, 2024Updated 2 years ago
- β41Jul 21, 2024Updated 2 years ago
- Code for the paper π³ Tree Search for Language Model Agentsβ223Jul 25, 2024Updated 2 years ago
- Setup scripts for the WebArena benchmarkβ22Jun 19, 2025Updated last year
- DoomArena is a Framework for Testing AI Agents Against Evolving Security Threatsβ62Sep 12, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICML'24] SeeAct is a system for generalist web agents that autonomously carry out tasks on any given website, with a focus on large multβ¦β852Feb 3, 2025Updated last year
- SkyRL: A Modular Full-stack RL Library for LLMsβ2,210Updated this week
- [NeurIPS 2022] πWebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agentsβ587Sep 6, 2024Updated last year
- COLM2026β38Jul 9, 2026Updated last month
- [NeurIPS'25 D&B] Mind2Web-2 Benchmark: Evaluating Agentic Search with Agent-as-a-Judgeβ114May 17, 2026Updated 3 months ago
- Democratizing Reinforcement Learning for LLMsβ5,808Updated this week
- OS-ATLAS: A Foundation Action Model For Generalist GUI Agentsβ453Apr 20, 2025Updated last year
- All-in-one Web Agent framework for post-training. Start building with a few clicks!β280Jul 7, 2025Updated last year
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ23,204Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A Benchmark for Evaluating Safety and Trustworthiness in Web Agents for Enterprise Scenariosβ26Mar 12, 2026Updated 5 months ago
- β117Jul 2, 2024Updated 2 years ago
- Windows Agent Arena (WAA) πͺ is a scalable OS platform for testing and benchmarking of multi-modal AI agents.β891Apr 13, 2026Updated 4 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnosticsβ2,781Aug 23, 2026Updated last week
- Interaction-first method for generating demonstrations for web-agents on any websiteβ57Apr 29, 2025Updated last year
- π» A curated list of papers and resources for multi-modal Graphical User Interface (GUI) agents.β1,214Aug 17, 2025Updated last year
- [ICLR'25 Oral] UGround: Universal GUI Visual Grounding for GUI Agentsβ317Updated this week