PULSE-EVAL
☆24Jan 12, 2024Updated 2 years ago
Alternatives and similar repositories for PULSE-EVAL
Users that are interested in PULSE-EVAL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 一个用于训练句子embedding的工具,支持Cosent以及Simcse、infonce☆24Jun 17, 2025Updated last year
- PULSE: Pretrained and Unified Language Service Engine☆498Dec 26, 2023Updated 2 years ago
- Counting-Stars (★)☆83Nov 24, 2025Updated 8 months ago
- 中国执业医师、药师、护士资格考试数据集和ChatGPT评估☆16Mar 13, 2026Updated 4 months ago
- ☆10Dec 28, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Official implementation of OpenTab (ICLR2024)☆14Mar 27, 2024Updated 2 years ago
- We systematically studied the influencing factors when LLM generates benchmarks,By using our code, you can generate high-quality QA datas…☆20May 20, 2025Updated last year
- [Nature Communications] The official code for "Quantifying the Reasoning Abilities of LLMs on Real-world Clinical Cases".☆70Nov 7, 2025Updated 8 months ago
- CMB, A Comprehensive Medical Benchmark in Chinese☆249Mar 27, 2025Updated last year
- LogicBench is a natural language question-answering dataset consisting of 25 different reasoning patterns spanning over propositional, fi…☆40May 2, 2024Updated 2 years ago
- Code and data to support Bamman et al. (2020), "A Dataset of Literary Coreference" (LREC)☆11Dec 8, 2022Updated 3 years ago
- ☆11Nov 8, 2023Updated 2 years ago
- Github repo for Peifeng's internship project☆13Nov 7, 2023Updated 2 years ago
- Training and testing code from our CVPR 2023 paper "Are Deep Neural Networks SMARTer than Second Graders?"☆11Aug 10, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Chain of Images for Intuitively Reasoning☆10Nov 29, 2023Updated 2 years ago
- ☆10Jun 11, 2023Updated 3 years ago
- ☆30Nov 5, 2024Updated last year
- 使用深度学习模型LSTM和ConvLSTM结合Attention,对金融衍生品的成交持仓比指标进行预测☆19Jan 7, 2022Updated 4 years ago
- The official code of TACL 2022, "Break, Perturb, Build: Automatic Perturbation of Reasoning Paths Through Question Decomposition".☆12Oct 18, 2021Updated 4 years ago
- ☆18Jan 12, 2024Updated 2 years ago
- The code and datasets of our ACM MM 2024 paper "Hallu-PI: Evaluating Hallucination in Multi-modal Large Language Models within Perturbed …☆11Sep 27, 2024Updated last year
- Code for "Demonstration-free Autonomous Reinforcement Learning via Implicit and Bidirectional Curriculum" (ICML 2023)☆10Jul 6, 2023Updated 3 years ago
- ☆10Nov 16, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [MICCAI'23] Text-guided Foundation Model Adaptation for Pathological Image Classification☆141Dec 26, 2023Updated 2 years ago
- PromptCBLUE: a large-scale instruction-tuning dataset for multi-task and few-shot learning in the medical domain in Chinese☆394Jan 23, 2024Updated 2 years ago
- get the media stream from Dahua/Haikang IPC SDK, and demux the stream to vedio and audio ES☆14Nov 15, 2015Updated 10 years ago
- The official data and code for EMNLP 2023 main conference paper: CRT-QA: A Dataset of Complex Reasoning Question Answering over Tabular D…☆13May 19, 2025Updated last year
- ☆11Jul 31, 2024Updated last year
- 基于LLM实现CHIP2021-Task3中文临床术语标准化任务,准确率约70%。☆16Dec 16, 2024Updated last year
- ☆13Nov 12, 2024Updated last year
- MedLSAM: Localize and Segment Anything Model for 3D Medical Images☆522Apr 30, 2024Updated 2 years ago
- This is the project for IRM methods☆12Sep 13, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code and notebooks and data for the paper "Domain Specific Question Answering Over Knowledge Graphs Using Logical Programming and Large L…☆12Jan 23, 2024Updated 2 years ago
- Cog wrapper for playgroundai/playground-v2.5-1024px-aesthetic☆17Nov 25, 2024Updated last year
- ☆34Feb 9, 2025Updated last year
- I don't want to maintain this project, the code probably won't compile or run. Archived.☆13Feb 25, 2024Updated 2 years ago
- ☆18Feb 20, 2025Updated last year
- MathEval is a benchmark dedicated to the holistic evaluation on mathematical capacities of LLMs.☆87Nov 15, 2024Updated last year
- ☆10Nov 29, 2024Updated last year