Official Repo: AutoResearchBench: Benchmarking AI Agents on Complex Scientific Literature Discovery
☆72Apr 24, 2026Updated 5 months ago
Alternatives and similar repositories for AutoResearchBench
Users that are interested in AutoResearchBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Executive Memory for Coherent Long-Horizon Reasoning!☆87Jan 14, 2026Updated 8 months ago
- Data Synthesis for Deep Research Based on Semi-Structured Data☆217Jul 14, 2026Updated 2 months ago
- ☆155Nov 17, 2025Updated 10 months ago
- Including 12+ cutting-edge agent systems across multiple research directions☆36Nov 10, 2025Updated 10 months ago
- [ICLR 2026] EditScore: Unlocking Online RL for Image Editing via High-Fidelity Reward Modeling☆260Mar 20, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A general memory system for agents, powered by deep-research☆862Mar 14, 2026Updated 6 months ago
- Advancing search on top of AI agents☆35Jun 9, 2026Updated 3 months ago
- 2022 USTC 011705 (OSH) Course Project of Runikraft Group☆13Jul 22, 2022Updated 4 years ago
- 🦞 ResearchClawBench: Evaluating AI Agents for Automated Research from Re-Discovery to New-Discovery☆266Sep 17, 2026Updated 2 weeks ago
- [ACL 2025 Oral] 🔥🔥 MegaPairs: Massive Data Synthesis for Universal Multimodal Retrieval☆250Nov 6, 2025Updated 10 months ago
- From Prompt Injection to Persistent Control: Defending Agentic Workspaces Against Trojan Backdoors☆20Sep 22, 2026Updated last week
- ☆13Nov 26, 2021Updated 4 years ago
- Implementation of "Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation"☆21Jul 31, 2023Updated 3 years ago
- Repo for WWW 2022 paper: Progressively Optimized Bi-Granular Document Representation for Scalable Embedding Based Retrieval☆16Mar 1, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- SPAR: multi-agent scholarly retrieval with query decomposition, query evolution, and citation-aware exploration.☆28Aug 25, 2026Updated last month
- A Self-Evolving Framework Through Executable Subagent Accumulation and Reuse☆64Jul 10, 2026Updated 2 months ago
- GISA: A Benchmark for General Information-Seeking Assistant☆37Updated this week
- ☆137Sep 3, 2026Updated 3 weeks ago
- ☆13Jun 18, 2019Updated 7 years ago
- ☆74Feb 22, 2023Updated 3 years ago
- Official data and code for the paper "VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents".☆14Mar 18, 2026Updated 6 months ago
- BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent (ACL 2026 Main)☆366May 28, 2026Updated 4 months ago
- SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation. A typed knowledge graph unifies data synthe…☆26Jul 8, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆22Apr 3, 2024Updated 2 years ago
- An all-in-one framework for Ad-hoc Information Retrieval.☆18Apr 3, 2024Updated 2 years ago
- DeepResearch Bench II (DRB2) is the follow-up to DeepResearch Bench, with a stronger focus on measuring the gap between deep research sys…☆94Sep 11, 2026Updated 3 weeks ago
- [EMNLP'26 Findings] OPD-Evolver☆45Jun 17, 2026Updated 3 months ago
- ☆25Jul 24, 2023Updated 3 years ago
- IKEA: Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent☆72May 13, 2025Updated last year
- [ACL'26 Findings] Recovered in Translation: Efficient Pipeline for Automated Translation of Benchmarks and Datasets☆20Jun 27, 2026Updated 3 months ago
- [ICLR 2025] Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning (SASR)☆12Aug 26, 2025Updated last year
- Benchmarking Language Agents Under Controllable and Extreme Context Growth☆60Apr 29, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- AIRS-Bench: an AI Research Science benchmark for quantifying the end-to-end AI research abilities of LLM agents☆121May 5, 2026Updated 4 months ago
- About Code release for "FlashBias: Fast Computation of Attention with Bias" (NeurIPS 2025), https://arxiv.org/abs/2505.12044☆34Nov 17, 2025Updated 10 months ago
- ☆17May 15, 2025Updated last year
- ☆14Oct 3, 2024Updated 2 years ago
- Code and pre-trained models for "ReasonBert: Pre-trained to Reason with Distant Supervision", EMNLP'2021☆28Feb 1, 2023Updated 3 years ago
- DLLM-Searcher has been accepted by SIGIR 2026! 🥳☆34Jan 23, 2026Updated 8 months ago
- [ACL 2024] "Understanding and Patching Compositional Reasoning in LLMs"☆14Aug 8, 2026Updated last month