SearchAgent-Zero: A Scalable Multi-Turn Search Agent RL Framework
☆148Aug 2, 2026Updated this week
Alternatives and similar repositories for SearchAgent-Zero
Users that are interested in SearchAgent-Zero are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026] Information Gain-based Policy Optimization: A Simple and Effective Approach for Multi-Turn Search Agents☆131Jul 14, 2026Updated 3 weeks ago
- Official data and code for the paper "VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents".☆15Mar 18, 2026Updated 4 months ago
- 🔍 OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diver…☆260Updated this week
- 🌟 A curated list of papers, methods, and resources on long-horizon credit assignment for agentic RL.☆51Updated this week
- AgentIR is a retriever specialized for Deep Research agents.☆62Apr 16, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Awesome List for Agentic RL☆1,747Jul 23, 2026Updated last week
- [arxiv: 2512.19673] Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policies☆60Feb 6, 2026Updated 5 months ago
- Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL☆5,247Nov 13, 2025Updated 8 months ago
- [ICLR 2025] COME: Test-time Adaption by Conservatively Minimizing Entropy☆23Mar 5, 2025Updated last year
- A Scalable Agentic RL Training Framework for Deep Research Agent☆112Jul 9, 2026Updated 3 weeks ago
- Agent-RRM: Exploring Reasoning Reward Model for Agents☆70Mar 17, 2026Updated 4 months ago
- ☆20Jan 18, 2026Updated 6 months ago
- Geometric Problem Solving Integrating FormalGeo Symbolic System and Hypergraph Neural Network.☆16Sep 23, 2025Updated 10 months ago
- 北大软微2023秋季 软工期末必修课真题☆14Jan 21, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Scaling Deep Research via Reinforcement Learning in Real-world Environments.☆788May 10, 2026Updated 2 months ago
- Agentic Learning Powered by AWorld☆119Jun 18, 2026Updated last month
- HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurren…☆70May 23, 2026Updated 2 months ago
- "QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks"☆241Updated this week
- [ACL-2026] MMSearch-R1 is an end-to-end RL framework that enables LMMs to perform on-demand, multi-turn search with real-world multimodal…☆475Apr 7, 2026Updated 3 months ago
- A minimal implementation of Agentic RAG using GRPO☆17Jun 11, 2025Updated last year
- Scaling Test-time Training for LLM Reasoning☆27Apr 14, 2026Updated 3 months ago
- Harness for deep search agent☆108Jun 16, 2026Updated last month
- 一个生产级的深度研究 Agent 系统,从零构建多智能体编排、Red-Blue 对抗降噪、 语义级上下文压缩、跨 Agent 共享记忆四大核心能力,配套 165 次独立实验 + Bootstrap 统计显著性检验的完整评测体系。☆80May 11, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,174Jun 9, 2026Updated last month
- DeepDive: Advancing Deep Search Agents with Knowledge Graphs and Multi-Turn RL☆335Jun 17, 2026Updated last month
- Executive Memory for Coherent Long-Horizon Reasoning!☆86Jan 14, 2026Updated 6 months ago
- SDD 规格驱动开发 + Harness 多 Agent 编排;支持在 OpenCode、Claude Code、Codex CLI 间切换执行引擎,完成从规格到落地的自动化开发任务。☆15Apr 8, 2026Updated 3 months ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆602Jun 12, 2026Updated last month
- ☆30Jan 27, 2026Updated 6 months ago
- REDSearch: A scalable, cost-efficient framework for long-horizon search agents. Features complex task synthesis, optimized mid-training, …☆134Feb 26, 2026Updated 5 months ago
- [ACL 2025 Findings] Implicit Reasoning in Transformers is Reasoning through Shortcuts☆18Mar 11, 2025Updated last year
- Rank-DistiLLM: Closing the Effectiveness Gap Between Cross-Encoders and LLMs for Passage Re-Ranking☆25Apr 4, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- (ICML 2025) Rethinking Chain-of-Thought from the Perspective of Self-Training☆13Feb 15, 2025Updated last year
- ☆176Jul 2, 2026Updated last month
- Curated papers, taxonomy, benchmarks, and decision guides for credit assignment in reasoning and agentic LLM reinforcement learning.☆125Updated this week
- Description for MV-MATH☆15Jul 20, 2025Updated last year
- [ACL 2026] WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering.☆15Jul 25, 2026Updated last week
- Starbucks: Improved Training for 2D Matryoshka Embeddings☆25Jun 30, 2025Updated last year
- Official PyTorch implementation of "LPOI: Listwise Preference Optimization for Vision Language Models" (ACL 2025 Main)☆16Jun 19, 2026Updated last month