EurekAgent: an autonomous research system for metric-driven tasks, built with Claude Code. Define the problem and metric. Get breakthrough results.
☆84Aug 10, 2026Updated last month
Alternatives and similar repositories for EurekAgent
Users that are interested in EurekAgent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [KDD 2025] AtomR: Atomic Operator-Empowered Large Language Models for Heterogeneous Knowledge Reasoning☆15May 27, 2025Updated last year
- [NIPS 2025 DB Spotlight] AGENTIF: Benchmarking Instruction Following of Large Language Models in Agentic Scenarios☆41Dec 1, 2025Updated 9 months ago
- Codes for paper SoAy: A Service-oriented APIs Applying Framework of Large Language Models☆27Jul 14, 2025Updated last year
- LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric Rewards☆40Jun 1, 2026Updated 3 months ago
- ☆12Apr 25, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Repository of paper "Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis" (ACL 2025 Main)☆20Jul 19, 2025Updated last year
- Code for our TKDE paper "Understanding WeChat User Preferences and “Wow” Diffusion"☆20Aug 29, 2024Updated 2 years ago
- The official repo for the paper "Externalizing Research Synthesis and Validation in AI Scientists through a Research Harness"☆45Jul 21, 2026Updated last month
- [ACL 2025] Agentic Reward Modeling: Integrating Human Preferences with Verifiable Correctness Signals for Reliable Reward Systems☆134Jun 11, 2025Updated last year
- 🌿 DeepPrune: Parallel Scaling without Inter-trace Redundancy☆21Apr 20, 2026Updated 4 months ago
- Code and dataset for the ACL 2021 paper "TWAG: A Topic-guided Wikipedia Abstract Generator"☆21Aug 9, 2021Updated 5 years ago
- [KDD24-ADS] R-Eval: A Unified Toolkit for Evaluating Domain Knowledge of Retrieval Augmented Large Language Models☆11Apr 9, 2024Updated 2 years ago
- ☆37Jul 13, 2026Updated 2 months ago
- Source code for EMNLP 2023 paper "Probabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex Questions".☆23Mar 21, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Resource, Evaluation and Detection Papers for ChatGPT☆456Mar 21, 2024Updated 2 years ago
- The repo for our paper: Enhancing LLM-Based Agents via Global Planning and Hierarchical Execution (NCIIP 2025 Best Paper)☆17Aug 18, 2025Updated last year
- Official repository of the video reasoning benchmark MMR-V. Can Your MLLMs "Think with Video"? [ICLR26]☆40Jun 23, 2025Updated last year
- Group project "Algorithms for large-scale optimal transport". Implement ADMMs and Sinkhorn's Algorithms.☆11Jan 28, 2019Updated 7 years ago
- ☆10Apr 21, 2023Updated 3 years ago
- 微博转发量预测大赛第四名☆11Dec 19, 2021Updated 4 years ago
- This repository contains the code and data for the paper "Chaining the Evidence: Robust Reinforcement Learning for Deep Search Agents wit…☆75Apr 8, 2026Updated 5 months ago
- A TensorFlow implementation of dependency-based word embeddings (dependency-based word2vec)☆12Jan 26, 2016Updated 10 years ago
- ☆62Oct 29, 2024Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- FlowPIE: Test-Time Scientific Idea Evolution with Flow-Guided Literature Exploration☆20Apr 26, 2026Updated 4 months ago
- Official implementation of ICLR 2026 paper "LUMINA: Detecting Hallucinations in RAG System with Context–Knowledge Signals"☆19Jan 31, 2026Updated 7 months ago
- REverse-Engineered Reasoning for Open-Ended Generation☆98Sep 10, 2025Updated last year
- Code for COLM 2026 Paper "Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs"☆33Jul 1, 2026Updated 2 months ago
- ☆27May 1, 2026Updated 4 months ago
- RapidIn: Scalable Influence Estimation for Large Language Models (LLMs). The implementation for paper "Token-wise Influential Training Da…☆22Mar 10, 2026Updated 6 months ago
- Generalist Multitask Transformer Encoders☆94Updated this week
- some machine learning algorithm☆14Jul 31, 2019Updated 7 years ago
- ☆43Jun 11, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- SimKO: Simple Pass@K Policy Optimization☆31Oct 24, 2025Updated 10 months ago
- Agent-native graph orchestration for Codex, Claude, and skill-compatible agents☆39Jul 18, 2026Updated 2 months ago
- Xlore2.0 Code[BaiduExtractor, HudongExtractor, WikiExtractor, XloreData, XloreWeb]☆12Apr 5, 2017Updated 9 years ago
- Data pipeline for HRM-Text pretraining☆74May 21, 2026Updated 3 months ago
- Source code and dataset for EMNLP 2022 paper "MAVEN-ERE: A Unified Large-scale Dataset for Event Coreference, Temporal, Causal, and Subev…☆92Aug 26, 2023Updated 3 years ago
- ☆57Aug 1, 2023Updated 3 years ago
- AutoMem: Automated Learning of Memory as a Cognitive Skill☆135Jul 3, 2026Updated 2 months ago