☆196Aug 17, 2026Updated last month
Alternatives and similar repositories for agent-data-protocol
Users that are interested in agent-data-protocol are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2024 Tutorial] Language Agents: Foundations, Prospects, and Risks☆10Nov 27, 2024Updated last year
- [NeurIPS'25 D&B] Mind2Web-2 Benchmark: Evaluating Agentic Search with Agent-as-a-Judge☆114May 17, 2026Updated 4 months ago
- An Autonomous Curriculum Reinforcement Learning framework that steers agents to continually learn in specific environments with zero huma…☆44Jun 7, 2026Updated 3 months ago
- ☆70Jun 27, 2025Updated last year
- [ICML 2026 Oral] Agent-native Mid-training for Software Engineering☆80Jun 7, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Following☆16Oct 31, 2024Updated last year
- Code, datasets, models for the paper "Automatic Evaluation of Attribution by Large Language Models"☆56Jul 3, 2023Updated 3 years ago
- Code for Paper: Training Software Engineering Agents and Verifiers with SWE-Gym [ICML 2025]☆742Jul 29, 2025Updated last year
- [ACL'24] Code and data of paper "When is Tree Search Useful for LLM Planning? It Depends on the Discriminator"☆54Feb 23, 2024Updated 2 years ago
- [ICLR 2026] The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution☆489Aug 18, 2026Updated last month
- Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents☆31Apr 16, 2026Updated 5 months ago
- ☆311Jul 1, 2026Updated 2 months ago
- Measuring how well CLI agents like Claude Code or Codex CLI can post-train base LLMs on a single H100 GPU in 10 hours☆567Sep 10, 2026Updated last week
- Code and implementations for the paper "AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcemen…☆869Feb 15, 2026Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code and pre-trained models for "ReasonBert: Pre-trained to Reason with Distant Supervision", EMNLP'2021☆28Feb 1, 2023Updated 3 years ago
- Your agent is powerful but it doesn't know you. VibeLens visualizes agent sessions, personalizes your agents, provides dashboard analytic…☆21Updated this week
- Official repo of Toucan: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments☆268Dec 16, 2025Updated 9 months ago
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆198Sep 3, 2026Updated 2 weeks ago
- Sotopia-π: Interactive Learning of Socially Intelligent Language Agents (ACL 2024)☆86May 7, 2024Updated 2 years ago
- Open-source test harness for AI agents. Stress-test production agents with adversarial multi-turn scenarios in CI☆28Aug 14, 2026Updated last month
- True Few-Shot BioIE: Benchmarking GPT-3 In-Context and Small PLM Fine-Tuning☆12Jul 6, 2022Updated 4 years ago
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆23Jun 2, 2026Updated 3 months ago
- GenEnv: Difficulty-Aligned Co-Evolution Between LLM Agents and Environment Simulators☆67Dec 23, 2025Updated 9 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆18Nov 1, 2024Updated last year
- An Empirical Study of Memorization in NLP (ACL 2022)☆13Jun 22, 2022Updated 4 years ago
- Official Implementation of Knowledge Flow Prompting☆35Oct 20, 2025Updated 11 months ago
- A Practitioner's Guide to M(eow)ti Turn Agentic ReinfOrcement learning☆85Jan 16, 2026Updated 8 months ago
- SkillWeaver is a framework to enable web agent self-improvement through environment exploration and skill synthesis.☆157Apr 14, 2025Updated last year
- Meta Agents Research Environments is a comprehensive platform designed to evaluate AI agents in dynamic, realistic scenarios. Unlike stat…☆557Aug 26, 2026Updated 3 weeks ago
- Official Repository of Personalized Visual Instruct Tuning☆34Mar 6, 2025Updated last year
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,673Updated this week
- [NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agents☆779Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆151Jan 4, 2024Updated 2 years ago
- ☆42May 26, 2026Updated 3 months ago
- DeepDive: Advancing Deep Search Agents with Knowledge Graphs and Multi-Turn RL☆349Jun 17, 2026Updated 3 months ago
- A Gym for Agentic LLMs☆510Jan 21, 2026Updated 8 months ago
- Bridging the Generalization Gap in Text-to-SQL Parsing with Schema Expansion☆13Jul 26, 2023Updated 3 years ago
- This is the oficial repository for "Safer-Instruct: Aligning Language Models with Automated Preference Data"☆17Feb 22, 2024Updated 2 years ago
- Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning☆456May 28, 2026Updated 3 months ago