[NeurIPS 2022] πWebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
β572Sep 6, 2024Updated last year
Alternatives and similar repositories for WebShop
Users that are interested in WebShop are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ALFWorld: Aligning Text and Embodied Environments for Interactive Learningβ810Feb 8, 2026Updated 5 months ago
- Code repo for "WebArena: A Realistic Web Environment for Building Autonomous Agents"β1,556Nov 26, 2025Updated 7 months ago
- [NeurIPS'23 Spotlight] "Mind2Web: Towards a Generalist Agent for the Web" -- the first LLM-based web agent and benchmark for generalist wβ¦β1,015Nov 5, 2025Updated 8 months ago
- ScienceWorld is a text-based virtual environment centered around accomplishing tasks from the standardized elementary science curriculum.β368Dec 3, 2025Updated 7 months ago
- [ICLR 2023] ReAct: Synergizing Reasoning and Acting in Language Modelsβ4,074Feb 6, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)β3,601Feb 8, 2026Updated 5 months ago
- VisualWebArena is a benchmark for multimodal agents.β484Nov 9, 2024Updated last year
- A collection of reinforcement learning environments for simple web interaction tasksβ393Updated this week
- Code and implementations for the ACL 2025 paper "AgentGym: Evolving Large Language Model-based Agents across Diverse Environments" by Zhiβ¦β817May 30, 2026Updated last month
- π AppWorld: A Controllable World of Apps and People for Benchmarking Function Calling and Interactive Coding Agent, ACL'24 Best Resourceβ¦β471Feb 17, 2026Updated 5 months ago
- Official implementation for "You Only Look at Screens: Multimodal Chain-of-Action Agents" (Findings of ACL 2024)β261Jul 16, 2024Updated 2 years ago
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-inβ¦β2,151Jun 9, 2026Updated last month
- Code and Data for Tau-Benchβ1,344Mar 18, 2026Updated 4 months ago
- WebLINX is a benchmark for building web navigation agents with conversational capabilitiesβ162Feb 11, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- An Analytical Evaluation Board of Multi-turn LLM Agents [NeurIPS 2024 Oral]β427May 20, 2024Updated 2 years ago
- [NeurIPS 2023 D&B] Code repository for InterCode benchmark https://arxiv.org/abs/2306.14898β253May 5, 2024Updated 2 years ago
- Code for the paper "LASER: LLM Agent with State-Space Exploration for Web Navigation"β35Sep 26, 2023Updated 2 years ago
- Official code for paper "SPA-RL: Reinforcing LLM Agent via Stepwise Progress Attribution"β89Sep 13, 2025Updated 10 months ago
- Research Code for "ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL"β208Apr 17, 2025Updated last year
- Code for Paper: Autonomous Evaluation and Refinement of Digital Agents [COLM 2024]β149Nov 26, 2024Updated last year
- [ICML 2024] Official repository for "Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models"β847Jul 30, 2024Updated last year
- ππͺ BrowserGym, a Gym environment for web task automationβ1,288Jul 17, 2026Updated last week
- An Illusion of Progress? Assessing the Current State of Web Agentsβ192Jun 25, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β121Apr 8, 2025Updated last year
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.β2,756Apr 14, 2026Updated 3 months ago
- [ICML'24 Spotlight] "TravelPlanner: A Benchmark for Real-World Planning with Language Agents"β531May 24, 2026Updated 2 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ22,649Updated this week
- Building Open LLM Web Agents with Self-Evolving Online Curriculum RLβ535Jun 6, 2025Updated last year
- A codebase for "Language Models can Solve Computer Tasks"β240May 1, 2024Updated 2 years ago
- A Universal Platform for Training and Evaluation of Mobile Interactionβ63Sep 24, 2025Updated 10 months ago
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRLβ5,150Nov 13, 2025Updated 8 months ago
- Official repo for paper DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning.β393Feb 22, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- FireAct: Toward Language Agent Fine-tuningβ296Oct 22, 2023Updated 2 years ago
- [ACL2025 Findings] Benchmarking Multihop Multimodal Internet Agentsβ54Feb 27, 2025Updated last year
- [ICLR'24 spotlight] An open platform for training, serving, and evaluating large language model for tool learning.β5,708May 21, 2025Updated last year
- [NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environmentsβ3,034Updated this week
- Towards Large Multimodal Models as Visual Foundation Agentsβ274Apr 24, 2025Updated last year
- β116Jul 2, 2024Updated 2 years ago
- Ο-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domainsβ1,657Updated this week