[NeurIPS 2022] πWebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
β581Sep 6, 2024Updated last year
Alternatives and similar repositories for WebShop
Users that are interested in WebShop are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ALFWorld: Aligning Text and Embodied Environments for Interactive Learningβ832Feb 8, 2026Updated 6 months ago
- Code repo for "WebArena: A Realistic Web Environment for Building Autonomous Agents"β1,577Nov 26, 2025Updated 8 months ago
- [NeurIPS'23 Spotlight] "Mind2Web: Towards a Generalist Agent for the Web" -- the first LLM-based web agent and benchmark for generalist wβ¦β1,019Nov 5, 2025Updated 9 months ago
- ScienceWorld is a text-based virtual environment centered around accomplishing tasks from the standardized elementary science curriculum.β378Aug 4, 2026Updated last week
- [ICLR 2023] ReAct: Synergizing Reasoning and Acting in Language Modelsβ4,107Feb 6, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)β3,668Feb 8, 2026Updated 6 months ago
- VisualWebArena is a benchmark for multimodal agents.β485Nov 9, 2024Updated last year
- A collection of reinforcement learning environments for simple web interaction tasksβ396Updated this week
- Code and implementations for the ACL 2025 paper "AgentGym: Evolving Large Language Model-based Agents across Diverse Environments" by Zhiβ¦β828May 30, 2026Updated 2 months ago
- π AppWorld: A Controllable World of Apps and People for Benchmarking Function Calling and Interactive Coding Agent, ACL'24 Best Resourceβ¦β483Feb 17, 2026Updated 5 months ago
- Official implementation for "You Only Look at Screens: Multimodal Chain-of-Action Agents" (Findings of ACL 2024)β262Jul 16, 2024Updated 2 years ago
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-inβ¦β2,218Jun 9, 2026Updated 2 months ago
- Code and Data for Tau-Benchβ1,381Mar 18, 2026Updated 4 months ago
- WebLINX is a benchmark for building web navigation agents with conversational capabilitiesβ163Feb 11, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- An Analytical Evaluation Board of Multi-turn LLM Agents [NeurIPS 2024 Oral]β436May 20, 2024Updated 2 years ago
- [NeurIPS 2023 D&B] Code repository for InterCode benchmark https://arxiv.org/abs/2306.14898β254May 5, 2024Updated 2 years ago
- Code for the paper "LASER: LLM Agent with State-Space Exploration for Web Navigation"β35Sep 26, 2023Updated 2 years ago
- Official code for paper "SPA-RL: Reinforcing LLM Agent via Stepwise Progress Attribution"β92Sep 13, 2025Updated 11 months ago
- Research Code for "ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL"β208Apr 17, 2025Updated last year
- Code for Paper: Autonomous Evaluation and Refinement of Digital Agents [COLM 2024]β149Nov 26, 2024Updated last year
- [ICML 2024] Official repository for "Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models"β849Jul 30, 2024Updated 2 years ago
- ππͺ BrowserGym, a Gym environment for web task automationβ1,315Jul 17, 2026Updated 3 weeks ago
- An Illusion of Progress? Assessing the Current State of Web Agentsβ194Jun 25, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β121Apr 8, 2025Updated last year
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.β2,769Jul 24, 2026Updated 3 weeks ago
- [ICML'24 Spotlight] "TravelPlanner: A Benchmark for Real-World Planning with Language Agents"β537May 24, 2026Updated 2 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Frameworkβ22,954Updated this week
- Building Open LLM Web Agents with Self-Evolving Online Curriculum RLβ538Jun 6, 2025Updated last year
- A codebase for "Language Models can Solve Computer Tasks"β239May 1, 2024Updated 2 years ago
- A Universal Platform for Training and Evaluation of Mobile Interactionβ63Sep 24, 2025Updated 10 months ago
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRLβ5,293Nov 13, 2025Updated 9 months ago
- Official repo for paper DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning.β393Feb 22, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- FireAct: Toward Language Agent Fine-tuningβ296Oct 22, 2023Updated 2 years ago
- [ACL2025 Findings] Benchmarking Multihop Multimodal Internet Agentsβ54Feb 27, 2025Updated last year
- [ICLR'24 spotlight] An open platform for training, serving, and evaluating large language model for tool learning.β5,728May 21, 2025Updated last year
- [NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environmentsβ3,081Updated this week
- Towards Large Multimodal Models as Visual Foundation Agentsβ274Apr 24, 2025Updated last year
- β117Jul 2, 2024Updated 2 years ago
- Ο-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domainsβ1,796Updated this week