[NeurIPS 2022] πWebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
β589Sep 6, 2024Updated last year
Alternatives and similar repositories for WebShop
Users that are interested in WebShop are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ALFWorld: Aligning Text and Embodied Environments for Interactive Learningβ848Feb 8, 2026Updated 6 months ago
- Code repo for "WebArena: A Realistic Web Environment for Building Autonomous Agents"β1,599Nov 26, 2025Updated 9 months ago
- [NeurIPS'23 Spotlight] "Mind2Web: Towards a Generalist Agent for the Web" -- the first LLM-based web agent and benchmark for generalist wβ¦β1,023Nov 5, 2025Updated 9 months ago
- ScienceWorld is a text-based virtual environment centered around accomplishing tasks from the standardized elementary science curriculum.β385Aug 20, 2026Updated 2 weeks ago
- [ICLR 2023] ReAct: Synergizing Reasoning and Acting in Language Modelsβ4,149Feb 6, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)β3,712Feb 8, 2026Updated 6 months ago
- VisualWebArena is a benchmark for multimodal agents.β485Nov 9, 2024Updated last year
- A collection of reinforcement learning environments for simple web interaction tasksβ398Aug 13, 2026Updated 3 weeks ago
- Code and implementations for the ACL 2025 paper "AgentGym: Evolving Large Language Model-based Agents across Diverse Environments" by Zhiβ¦β840May 30, 2026Updated 3 months ago
- π AppWorld: A Controllable World of Apps and People for Benchmarking Function Calling and Interactive Coding Agent, ACL'24 Best Resourceβ¦β502Updated this week
- Official implementation for "You Only Look at Screens: Multimodal Chain-of-Action Agents" (Findings of ACL 2024)β263Jul 16, 2024Updated 2 years ago
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-inβ¦β2,276Jun 9, 2026Updated 2 months ago
- Code and Data for Tau-Benchβ1,419Mar 18, 2026Updated 5 months ago
- WebLINX is a benchmark for building web navigation agents with conversational capabilitiesβ163Aug 16, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An Analytical Evaluation Board of Multi-turn LLM Agents [NeurIPS 2024 Oral]β443May 20, 2024Updated 2 years ago
- [NeurIPS 2023 D&B] Code repository for InterCode benchmark https://arxiv.org/abs/2306.14898β255May 5, 2024Updated 2 years ago
- Code for the paper "LASER: LLM Agent with State-Space Exploration for Web Navigation"β37Sep 26, 2023Updated 2 years ago
- Official code for paper "SPA-RL: Reinforcing LLM Agent via Stepwise Progress Attribution"β93Sep 13, 2025Updated 11 months ago
- Research Code for "ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL"β208Apr 17, 2025Updated last year
- Code for Paper: Autonomous Evaluation and Refinement of Digital Agents [COLM 2024]β149Nov 26, 2024Updated last year
- [ICML 2024] Official repository for "Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models"β857Jul 30, 2024Updated 2 years ago
- ππͺ BrowserGym, a Gym environment for web task automationβ1,342Jul 17, 2026Updated last month
- An Illusion of Progress? Assessing the Current State of Web Agentsβ200Jun 25, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- β121Apr 8, 2025Updated last year
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnosticsβ2,789Aug 23, 2026Updated last week
- [ICML'24 Spotlight] "TravelPlanner: A Benchmark for Real-World Planning with Language Agents"β543May 24, 2026Updated 3 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ23,258Updated this week
- Building Open LLM Web Agents with Self-Evolving Online Curriculum RLβ539Jun 6, 2025Updated last year
- A codebase for "Language Models can Solve Computer Tasks"β240May 1, 2024Updated 2 years ago
- A Universal Platform for Training and Evaluation of Mobile Interactionβ64Sep 24, 2025Updated 11 months ago
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRLβ5,363Nov 13, 2025Updated 9 months ago
- Official repo for paper DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning.β396Feb 22, 2025Updated last year
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- FireAct: Toward Language Agent Fine-tuningβ296Oct 22, 2023Updated 2 years ago
- [ACL2025 Findings] Benchmarking Multihop Multimodal Internet Agentsβ54Feb 27, 2025Updated last year
- [ICLR'24 spotlight] An open platform for training, serving, and evaluating large language model for tool learning.β5,733May 21, 2025Updated last year
- [NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environmentsβ3,120Updated this week
- Towards Large Multimodal Models as Visual Foundation Agentsβ276Apr 24, 2025Updated last year
- β117Jul 2, 2024Updated 2 years ago
- Ο-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domainsβ1,938Updated this week