"DeepResearch-Eval: An End-to-End Evaluation Framework for DeepResearch Systems"
☆51Oct 16, 2025Updated 11 months ago
Alternatives and similar repositories for DeepResearch-Eval
Users that are interested in DeepResearch-Eval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- "FastAgent: Simple, Fast, and Strong LLM Agents"☆58Feb 10, 2026Updated 7 months ago
- Build the memory your agent actually needs — automatically.☆24Jul 3, 2026Updated 2 months ago
- [EMNLP2025] "RecGPT: A Foundation Model for Sequential Recommendation"☆61Oct 14, 2025Updated 11 months ago
- Memory Governance Protocol☆60Jul 20, 2026Updated 2 months ago
- [ACM TIST] "LLM4Urban: Urban Computing in the Era of Large Language Models"☆60Apr 4, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [WSDM'2025] "MixRec: Heterogeneous Graph Collaborative Filtering"☆19Dec 19, 2024Updated last year
- Upskill your model: Flash price. Pro performance.☆95Jul 26, 2026Updated last month
- [KDD 2026] "Breaking Information Cocoons: A Hyperbolic Graph-LLM Framework for Exploration and Exploitation in Recommender Systems"☆16Jan 29, 2025Updated last year
- "ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"☆32Dec 15, 2025Updated 9 months ago
- TailorRec: Recommendation for AI-Generated Content☆31Jun 29, 2026Updated 2 months ago
- [ACL 2026 Oral] "LightReasoner: Can Small Language Models Teach Large Language Models Reasoning?"☆611May 22, 2026Updated 4 months ago
- "VideoAgent: All-in-One Agentic Framework for Video Understanding, Editing, and Remaking"☆31Oct 17, 2025Updated 11 months ago
- [WSDM'2025] "DiffGraph: Heterogeneous Graph Diffusion Model"☆77Jun 12, 2025Updated last year
- "DeepInnovator: AI Research Assistant - Idea Spark & Scientific Discovery"☆290Mar 6, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ACL2025] "RecLM: Recommendation Instruction Tuning"☆112Jun 2, 2025Updated last year
- [CIKM'2024] "EasyST: A Simple Framework for Spatio-Temporal Prediction"☆19Sep 17, 2024Updated 2 years ago
- "AnyTool: Universal Tool-Use Layer for AI Agents"☆691Feb 28, 2026Updated 6 months ago
- [WWW'2024] "GraphPro: Graph Pre-training and Prompt Learning for Recommendation"☆72Jun 10, 2024Updated 2 years ago
- "OpenCity: Open Spatio-Temporal Foundation Models for Traffic Prediction"☆182Apr 3, 2026Updated 5 months ago
- Code for ICLR'24 Paper "Decoupling Weighing and Selecting for Integrating Multiple Graph Pre-training Tasks"☆11Mar 12, 2024Updated 2 years ago
- 智能合同审查agent是一个基于人工智能的合同风险分析与建议系统。该应用使用Streamlit构建用户界面,集成了目前领先的AI大模型来自动分析合同文档,识别潜在风险,并提供专业的修改建议。该系统还提供了合同合规性评分功能,帮助用户评估合同的法律合规性。☆17Aug 8, 2025Updated last year
- "CatchMe: Make Your AI Agents Truly Personal"☆505Updated this week
- "Your Fully-Automated Personal AI Assistant"☆1,743Oct 16, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [Recsys'2023] "RCL: Multi-Relational Contrastive Learning for Recommendation"☆15Sep 6, 2023Updated 3 years ago
- [ICML'2023] "GraphST: Spatial-Temporal Graph Learning with Adversarial Contrastive Adaptation"☆57Dec 24, 2024Updated last year
- Pushing Test-Time Scaling Limits of Deep Search with Asymmetric Verification☆23Oct 8, 2025Updated 11 months ago
- ☆16Jan 19, 2026Updated 8 months ago
- [CIKM'2023] "CL4ST: Spatio-Temporal Meta Contrastive Learning"☆23Jun 17, 2024Updated 2 years ago
- [EMNLP2025] "GraphAgent: Agentic Graph Language Assistant"☆370Feb 8, 2025Updated last year
- [ICDE'2024] "GraphAug: Graph Augmentation for Recommendation"☆23Sep 17, 2024Updated 2 years ago
- [WWW'2024] "PromptMM: Multi-Modal Knowledge Distillation for Recommendation with Prompt-Tuning"☆51Mar 20, 2024Updated 2 years ago
- [ACL 2026] "OpenPhone: Mobile Agentic Foundation Models for AI Phone"☆975Jul 14, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- MiroEval: A benchmark and evaluation framework for deep research agents — 100 tasks (70 text, 30 multimodal) assessed across synthesis qu…☆50Jul 6, 2026Updated 2 months ago
- ☆72Jan 4, 2026Updated 8 months ago
- Text2SQL-Eval is a Text-to-SQL evaluating component for LLM trained on an open-source training dataset.☆15Jan 12, 2024Updated 2 years ago
- DeepResearchEval: An Automated Framework for Deep Research Task Construction and Agentic Evaluation.☆143Feb 10, 2026Updated 7 months ago
- [CIKM'2024] "RecDiff: Diffusion Model for Social Recommendation"☆89Jun 16, 2025Updated last year
- [EMNLP2026] "VideoAgent: All-in-One Agentic Framework for Video Understanding and Editing, and Remaking"☆1,896Jul 22, 2026Updated 2 months ago
- Code release for "CURIE: Evaluating LLMs On Multitask Scientific Long Context Understanding and Reasoning", ICLR 2025☆34Apr 21, 2025Updated last year