A general framework used on evaluating the performance of large language models (LLMs) based on the peer review mechanism among LLMs
☆19Aug 3, 2024Updated 2 years ago
Alternatives and similar repositories for PRE
Users that are interested in PRE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SIGIR'22 paper: Axiomatically Regularized Pre-training for Ad hoc Search☆23May 24, 2023Updated 3 years ago
- An evaluation framework to test AI in a trial-and-error process. It is a simplified Natural Selection test.☆22Mar 11, 2025Updated last year
- A smart skill search engine for agents with multi-field retrieval and quality signals.☆23Apr 15, 2026Updated 5 months ago
- ☆16Jul 25, 2025Updated last year
- Code for AAAI 2024 paper Wikiformer☆20Dec 21, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- LegalOne: A Family of Foundation Models for Reliable Legal Reasoning☆78Feb 3, 2026Updated 7 months ago
- word2vec java版本的一个实现☆10Apr 24, 2016Updated 10 years ago
- Code for JuDGE, SIGIR 2025 Long Paper☆37Aug 7, 2025Updated last year
- ☆29Jul 25, 2025Updated last year
- Large Language Models as Evaluators for Recommendation Explanations (RecSys 2024 Reproducibility)☆21Aug 13, 2025Updated last year
- Repo. for RLCF.☆15Apr 1, 2024Updated 2 years ago
- Code for I3 Retriever, accepted by CIKM'23.☆53Oct 22, 2023Updated 2 years ago
- The repo for our paper: Enhancing LLM-Based Agents via Global Planning and Hierarchical Execution (NCIIP 2025 Best Paper)☆17Aug 18, 2025Updated last year
- CIKM 2022: Evaluating Interpolation and Extrapolation Performance of Neural Retrieval Models☆10Aug 4, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Nov 9, 2021Updated 4 years ago
- ☆29Mar 10, 2026Updated 6 months ago
- ☆13May 11, 2021Updated 5 years ago
- Repository of paper "Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis" (ACL 2025 Main)☆20Jul 19, 2025Updated last year
- The official repo for our SIGIR'23 Full paper: Constructing Tree-based Index for Efficient and Effective Dense Retrieval☆28Jun 7, 2023Updated 3 years ago
- A little Python script to generate Cohen's Kappa and Weighted Kappa measures for inter-rater reliability☆23Sep 17, 2019Updated 7 years ago
- Code for KERM: Incorporating Explicit Knowledge in Pre-trained Language Models for Passage Re-ranking, accepted at SIGIR 2022.☆19Oct 31, 2022Updated 3 years ago
- collecting publicly available distillation datasets based on DepSeek-R1☆29Mar 12, 2025Updated last year
- The official repo for our SIGIR'23 Full paper: Structure-aware Pre-trained Language Model for Legal Case Retrieval☆98May 9, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code to reproduce THUIR‘s submissions for COLIEE 2023 Task1 and Task2☆28May 12, 2023Updated 3 years ago
- RapidIn: Scalable Influence Estimation for Large Language Models (LLMs). The implementation for paper "Token-wise Influential Training Da…☆22Mar 10, 2026Updated 6 months ago
- WSDM'22 Best Paper: Learning Discrete Representations via Constrained Clustering for Effective and Efficient Dense Retrieval☆119Aug 7, 2024Updated 2 years ago
- The implementation of deep learning models for EEG classification.☆15Apr 21, 2023Updated 3 years ago
- Large Visual Language Model(LVLM), Large Language Model(LLM), Multimodal Large Language Model(MLLM), Alignment, Agent, AI System, Survey☆21Jul 27, 2025Updated last year
- 采用bert进行事件抽取,[cls]进行事件分类,最后一层向量进行序列标注,两个任务同时训练。☆12Jun 7, 2021Updated 5 years ago
- deepspeed+trainer简单高效实现多卡微调大模型☆132May 27, 2023Updated 3 years ago
- A collection of recent open-source math datasets for training and evaluating Math LLMs☆34Apr 26, 2026Updated 4 months ago
- ☆16May 8, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- personalized product search with product reviews☆17Feb 1, 2023Updated 3 years ago
- Hybrid List Aware Transformer Reranking☆19Oct 25, 2022Updated 3 years ago
- ☆47Apr 9, 2025Updated last year
- ☆12Jul 4, 2022Updated 4 years ago
- Click models by c++☆21Jan 20, 2021Updated 5 years ago
- Truly Conversational Search is the next logic step in the journey to generate intelligent and useful AI. To understand what this may mean…☆115Jun 12, 2023Updated 3 years ago
- 基于区块链的商品溯源系统☆10Mar 11, 2021Updated 5 years ago