☆26May 20, 2025Updated last year
Alternatives and similar repositories for MCPWorld
Users that are interested in MCPWorld are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Implementation of ARPO: End-to-End Policy Optimization for GUI Agents with Experience Replay☆163May 29, 2025Updated last year
- ☆29Apr 2, 2026Updated 5 months ago
- [ICCV 2025] GUIOdyssey is a comprehensive dataset for training and evaluating cross-app navigation agents. GUIOdyssey consists of 8,834 e…☆162Jan 3, 2026Updated 8 months ago
- Advanced GUI agents☆17Feb 3, 2026Updated 7 months ago
- ☆41Sep 22, 2021Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- LLMs in a tiny box, under 3 Watt☆18Dec 30, 2024Updated last year
- VeriWeb: Verifiable Long-Chain Web Benchmark for Agentic Information-Seeking☆88Jan 21, 2026Updated 7 months ago
- Implementations of Influential Recommender System☆13Oct 29, 2024Updated last year
- LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation☆70Aug 9, 2024Updated 2 years ago
- Graph Convolutional Module for Temporal Action Localization in Videos☆10Jul 4, 2020Updated 6 years ago
- ☆16Sep 8, 2025Updated last year
- Implementation of SLIM, a framework of dynamics skill lifecycle management for agentic reinforcement learning☆22May 12, 2026Updated 4 months ago
- ☆17Dec 9, 2022Updated 3 years ago
- [SIGIR 2026] "One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment"☆16Apr 21, 2026Updated 4 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆57Aug 19, 2025Updated last year
- ☆48Apr 11, 2024Updated 2 years ago
- The CodeInsight dataset is designed for code generation tasks, providing developers with expert-curated examples that bridge the gap betw…☆15Oct 22, 2024Updated last year
- ☆18Jul 10, 2025Updated last year
- Replication package for ISSTA2023 paper - Towards Efficient Fine-tuning of Pre-trained Code Models: An Experimental Study and Beyond☆23Apr 9, 2023Updated 3 years ago
- ☆37May 28, 2024Updated 2 years ago
- Code for 🌍 UI-Simulator: LLMs as Scalable, General-Purpose Simulators For Evolving Digital Agent Training☆21Oct 17, 2025Updated 11 months ago
- ToolCUA: Towards Optimal GUI-Tool Path Orchestration for Computer Use Agents☆62May 13, 2026Updated 4 months ago
- ProAct is a framework designed to enable Large Language Model (LLM) agents to perform accurate, multi-turn lookahead reasoning in interac…☆18Feb 11, 2026Updated 7 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆20Mar 6, 2023Updated 3 years ago
- Fault-aware neural code rankers☆32Dec 9, 2022Updated 3 years ago
- ☆11Mar 10, 2021Updated 5 years ago
- This is the tool released in ICSE 2024 paper "Domain Knowledge Matters: Improving Prompts with Fix Templates for Repairing Python Type Er…☆17Jun 5, 2023Updated 3 years ago
- ☆23Jun 16, 2026Updated 3 months ago
- AndroidWorld is an environment and benchmark for autonomous agents☆919Sep 9, 2026Updated last week
- Official Implementation of ConceptLM.☆28Mar 18, 2026Updated 6 months ago
- [CVPR 2024] KEPP: Why Not Use Your Textbook? Knowledge-Enhanced Procedure Planning of Instructional Videos☆12Sep 24, 2024Updated last year
- [ICLR'25] Code for KaSA, an official implementation of "KaSA: Knowledge-Aware Singular-Value Adaptation of Large Language Models"☆22Jan 16, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- LGEB: Benchmark of Language Generation Evaluation☆16Oct 21, 2022Updated 3 years ago
- (ICLR 2025) AgentRefine: Enhancing Agent Generalization through Refinement Tuning☆21Nov 22, 2025Updated 9 months ago
- [ICLR 2026] Computer Agent Arena: Toward Human-Centric Evaluation and Analysis of Computer-Use Agents☆68Feb 26, 2026Updated 6 months ago
- On Policy Distillation Build on top of Verl☆99Sep 3, 2026Updated 2 weeks ago
- This repository contains the code for the paper “Neuro-Symbolic Query Compiler”, accepted to the Findings of ACL 2025.☆19Oct 20, 2025Updated 10 months ago
- Reproducing R1 for Code with Reliable Rewards☆13Apr 9, 2025Updated last year
- ☆11Feb 5, 2021Updated 5 years ago