This is the code repo for the paper AceSearcher: Bootstrapping Reasoning and Search for LLMs via Reinforced Self-Play (NeurIPS 2025 Spotlight).
☆25Sep 29, 2025Updated 11 months ago
Alternatives and similar repositories for AceSearcher
Users that are interested in AceSearcher are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆30Apr 8, 2025Updated last year
- ☆17Jan 26, 2024Updated 2 years ago
- [ACL 2024 Findings] This is the code for our paper "Knowledge-Infused Prompting: Assessing and Advancing Clinical Text Data Generation wi…☆43Jun 23, 2024Updated 2 years ago
- [EMNLP'24] MedAdapter: Efficient Test-Time Adaptation of Large Language Models Towards Medical Reasoning☆36Dec 26, 2024Updated last year
- Replayable Nova-Seed → MARK → Mini Sovereign → AGI Jobs → Evidence Docket → vNext loop☆16Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for "BERTifying the Hidden Markov Model for Multi-Source Weakly Supervised Named Entity Recognition"☆32Jun 20, 2023Updated 3 years ago
- [ACL 2024] This is the code for our paper ”RAM-EHR: Retrieval Augmentation Meets Clinical Predictions on Electronic Health Records“.☆42Sep 19, 2024Updated 2 years ago
- ☆24Jun 13, 2024Updated 2 years ago
- Official repo for "Binary Retrieval-augmented Reward Mitigates Hallucinations"☆16Nov 13, 2025Updated 10 months ago
- ☆22Dec 18, 2025Updated 9 months ago
- MCPL: Multi-modal Collaborative Prompt Learning for Medical Vision-Language Model (Initial Version)☆13Apr 17, 2024Updated 2 years ago
- 🎯Awesome-AgenticRAG_DeepResearch: A curated list of resources on Agentic RAG & DeepResearch. 学习 参考关于AgenticRAG、DeepResearch的发展相关论文☆34Jul 23, 2026Updated 2 months ago
- [NeurIPS 2025] Think or Not? Selective Reasoning via Reinforcement Learning for Vision-Language Models☆61Sep 29, 2025Updated 11 months ago
- [ICLR'26] MedAgentGYM: Training LLM Agents for Code-Based Medical Reasoning at Scale☆134Apr 12, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2025] Reasoning MLLM, Share-GRPO, advantage vanishing, sparse reward☆38Sep 19, 2025Updated last year
- The codes for our ACL'22 paper: PRBOOST: Prompt-Based Rule Discovery and Boosting for Interactive Weakly-Supervised Learning.☆36Mar 18, 2022Updated 4 years ago
- Neurlps 2025☆19Mar 9, 2026Updated 6 months ago
- Curriculum-RLAIF is a data-centric curriculum learning framework for reward model training in RLAIF-based LLM alignment☆23Apr 18, 2026Updated 5 months ago
- [EMNLP'24] EHRAgent: Code Empowers Large Language Models for Complex Tabular Reasoning on Electronic Health Records☆143Dec 26, 2024Updated last year
- ☆13May 15, 2024Updated 2 years ago
- State-of-the-art neural tagger and lemmatizer for ancient languages☆16Aug 13, 2026Updated last month
- QBRC Somatic Mutation Calling Pipeline☆16Feb 8, 2022Updated 4 years ago
- Perform facts checks on your conversations with LLMs to catch fake-news, misleading information, and LLMs confusion.☆11Apr 22, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Code for "Neural Retrievers are Biased Towards LLM-Generated Content"☆14Oct 18, 2024Updated last year
- ☆15Nov 18, 2025Updated 10 months ago
- We release Open Meditron, a fully open, clinician-audited medical training corpus and evaluation protocol that closes the open-vs-closed …☆17Aug 3, 2026Updated last month
- [ACL 2025] Removal of Hallucination on Hallucination: Debate-Augmented RAG☆44Aug 4, 2025Updated last year
- This is the official repo for Towards Uncertainty-Aware Language Agent.☆31Aug 15, 2024Updated 2 years ago
- [EMNLP25 Main]The official code of "Gradient-Attention Guided Dual-Masking Synergetic Framework for Robust Text-based Person Retrieval"☆26Mar 30, 2026Updated 5 months ago
- MyPhoneBench: Do Phone-Use Agents Respect Your Privacy?☆24Apr 3, 2026Updated 5 months ago
- [NeurIPS 2024 Datasets and Benchmark Track Oral] MedCalc-Bench: Evaluating Large Language Models for Medical Calculations☆96Dec 18, 2025Updated 9 months ago
- [EMNLP 2024 Findings] ICON: Improving Inter-Report Consistency in Radiology Report Generation via Lesion-aware Mixup Augmentation☆19Dec 11, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [NeurIPS 2024] A task generation and model evaluation system for multimodal language models.☆71Nov 27, 2024Updated last year
- this is based on the paper Chain-of-Retrieval Augmented Generation☆15Mar 29, 2025Updated last year
- Med-Scout: Curing MLLMs' Geometric Blindness in Medical Perception via Geometry-Aware RL Post-Training☆20Feb 8, 2026Updated 7 months ago
- Toward Expert-Level Medical Text Validation with Language Models☆19Oct 23, 2025Updated 11 months ago
- Repository for the paper: Aligning LLMs to Ask Good Questions A Case Study in Clinical Reasoning☆18Feb 21, 2025Updated last year
- Hands-on repository for fine-tuning Large Language Models (LLMs) in the clinical domain with tutorials☆18Jul 10, 2026Updated 2 months ago
- CheXficient☆16Aug 28, 2026Updated 3 weeks ago