SearchAgent-Zero: A Scalable Multi-Turn Search Agent RL Framework
☆157Aug 23, 2026Updated this week
Alternatives and similar repositories for SearchAgent-Zero
Users that are interested in SearchAgent-Zero are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026] Information Gain-based Policy Optimization: A Simple and Effective Approach for Multi-Turn Search Agents☆144Jul 14, 2026Updated last month
- Official data and code for the paper "VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents".☆15Mar 18, 2026Updated 5 months ago
- 🔍 OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diver…☆270Jul 30, 2026Updated 3 weeks ago
- 🌟 A curated list of papers, methods, and resources on long-horizon credit assignment for agentic RL.☆61Aug 6, 2026Updated 2 weeks ago
- AgentIR is a retriever specialized for Deep Research agents (COLM 2026).☆63Apr 16, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Awesome List for Agentic RL☆1,793Aug 11, 2026Updated last week
- [EMNLP 2026 Main] Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policies☆60Feb 6, 2026Updated 6 months ago
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,321Nov 13, 2025Updated 9 months ago
- The official repository of "A Comprehensive Survey on Reinforcement Learning-based Agentic Search: Foundations, Roles, Optimizations, Eva…☆290Jul 30, 2026Updated 3 weeks ago
- [ICLR 2025] COME: Test-time Adaption by Conservatively Minimizing Entropy☆23Mar 5, 2025Updated last year
- Agent-RRM: Exploring Reasoning Reward Model for Agents☆70Mar 17, 2026Updated 5 months ago
- [COLM 2026] A Scalable Agentic RL Training Framework for Deep Research Agent☆125Aug 7, 2026Updated 2 weeks ago
- ☆20Jan 18, 2026Updated 7 months ago
- Geometric Problem Solving Integrating FormalGeo Symbolic System and Hypergraph Neural Network.☆16Sep 23, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 北大软微2023秋季 软工期末必修课真题☆14Jan 21, 2024Updated 2 years ago
- Scaling Deep Research via Reinforcement Learning in Real-world Environments.☆797May 10, 2026Updated 3 months ago
- Agentic Learning Powered by AWorld☆125Jun 18, 2026Updated 2 months ago
- HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurren…☆74May 23, 2026Updated 3 months ago
- "QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks"☆248Aug 5, 2026Updated 2 weeks ago
- [ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodal…☆479Apr 7, 2026Updated 4 months ago
- A minimal implementation of Agentic RAG using GRPO☆17Jun 11, 2025Updated last year
- Scaling Test-time Training for LLM Reasoning☆30Apr 14, 2026Updated 4 months ago
- Harness for deep search agent☆109Jun 16, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,242Jun 9, 2026Updated 2 months ago
- DeepDive: Advancing Deep Search Agents with Knowledge Graphs and Multi-Turn RL☆342Jun 17, 2026Updated 2 months ago
- Executive Memory for Coherent Long-Horizon Reasoning!☆86Jan 14, 2026Updated 7 months ago
- SDD 规格驱动开发 + Harness 多 Agent 编排;支持在 OpenCode、Claude Code、Codex CLI 间切换执行引擎,完成从规格到落地的自动化开发任务。☆15Apr 8, 2026Updated 4 months ago
- ☆47Jan 19, 2026Updated 7 months ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆622Jun 12, 2026Updated 2 months ago
- ☆43Jan 27, 2026Updated 6 months ago
- REDSearch: A scalable, cost-efficient framework for long-horizon search agents. Features complex task synthesis, optimized mid-training, …☆141Feb 26, 2026Updated 5 months ago
- [ACL 2025 Findings] Implicit Reasoning in Transformers is Reasoning through Shortcuts☆18Mar 11, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Rank-DistiLLM: Closing the Effectiveness Gap Between Cross-Encoders and LLMs for Passage Re-Ranking☆25Apr 4, 2025Updated last year
- (ICML 2025) Rethinking Chain-of-Thought from the Perspective of Self-Training☆13Feb 15, 2025Updated last year
- ☆187Aug 13, 2026Updated last week
- Curated papers, taxonomy, benchmarks, and decision guides for credit assignment in reasoning and agentic LLM reinforcement learning.☆146Aug 3, 2026Updated 3 weeks ago
- [ACL 2026] WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering.☆15Jul 25, 2026Updated 3 weeks ago
- Description for MV-MATH☆15Jul 20, 2025Updated last year
- Starbucks: Improved Training for 2D Matryoshka Embeddings☆25Jun 30, 2025Updated last year