[NeurIPS 2024] Code and Data Repo for Paper "Embedding Trajectory for Out-of-Distribution Detection in Mathematical Reasoning"
☆28May 28, 2024Updated 2 years ago
Alternatives and similar repositories for OOD-Math-Reasoning
Users that are interested in OOD-Math-Reasoning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2023] Code and Data Repo for Paper "Element-aware Summary and Summary Chain-of-Thought (SumCoT)"☆56Jan 21, 2024Updated 2 years ago
- [ACL 2024] Can Watermarks Survive Translation? On the Cross-lingual Consistency of Text Watermark for Large Language Models☆49Jun 4, 2024Updated 2 years ago
- [ICLR 2025] Code and Data Repo for Paper "Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation"☆105Dec 19, 2024Updated last year
- [WMT 2022] Implementation of TAL-SJTU's system for WMT22 English-Livonian☆23May 4, 2023Updated 3 years ago
- [ACL 2022] Bridging the Data Gap between Training and Inference for Unsupervised Neural Machine Translation☆31Oct 6, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- {DeepL, Google, WMT-Best, davinci-003, turbo, gpt-4} × {En-De, En-Cs, En-Ru, En-Zh, De-Fr, En-Ja, Uk-En, Uk-Cs, En-Hr, En-Ha, En-Is}☆14Jun 18, 2023Updated 3 years ago
- Fast spectral clustering, described in the NeurIPS'23 paper "Fast and Simple Spectral Clustering in Theory and Practice"☆18Jun 19, 2025Updated last year
- [ICML2024]Adaptive decoding balances the diversity and coherence of open-ended text generation.☆19Jun 2, 2024Updated 2 years ago
- R-Judge: Benchmarking Safety Risk Awareness for LLM Agents (EMNLP Findings 2024)☆114Jan 11, 2026Updated 8 months ago
- Official repository for Decentralized Arena via Collective LLM Intelligence☆18May 19, 2025Updated last year
- [TACL 2024] MAPS enables LLMs🤖 to mimic the human😁 translation process.☆147Jun 7, 2024Updated 2 years ago
- Findings of ACL 2021☆24May 8, 2021Updated 5 years ago
- The rule-based evaluation subset and code implementation of Omni-MATH☆29Dec 23, 2024Updated last year
- Reliable Source Approximation: Source-Free Domain Adaptation for Vestibular Schwannoma MRI Segmentation☆11Dec 28, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- source code for NeurIPS'24 paper "HaloScope: Harnessing Unlabeled LLM Generations for Hallucination Detection"☆70Apr 11, 2025Updated last year
- Official repository for ACL 2025 paper "ProcessBench: Identifying Process Errors in Mathematical Reasoning"☆194May 20, 2025Updated last year
- The source code (Pytorch version) of paper "Multi-modality augmented Prototypical Network for Fault Diagnosis"☆12Aug 26, 2024Updated 2 years ago
- EMNLP'23 survey: a curation of awesome papers and resources on refreshing large language models (LLMs) without expensive retraining.☆135Dec 12, 2023Updated 2 years ago
- This the implementation of LeCo☆33Jan 20, 2025Updated last year
- [ICASSP 2022] Official PyTorch Implementation for "Attention Probe: Vision Transformer Distillation in the Wild" (ICASSP 2022)☆11Jan 23, 2022Updated 4 years ago
- ☆25Oct 11, 2024Updated last year
- ☆30Jun 19, 2023Updated 3 years ago
- A mobile GUI search engine using a vision-language model☆15May 5, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ACL 2024 Findings] CriticBench: Benchmarking LLMs for Critique-Correct Reasoning☆31Mar 5, 2024Updated 2 years ago
- ☆23Nov 15, 2022Updated 3 years ago
- ☆10Dec 21, 2022Updated 3 years ago
- ☆20Nov 3, 2024Updated last year
- ☆15Jan 14, 2026Updated 8 months ago
- Subspace clustering guided unsupervised feature selection☆14Feb 29, 2020Updated 6 years ago
- Code for training a language model reaction predictor. (To accompany our paper on the OOD evaluation of reaction predictors).☆12Jan 13, 2025Updated last year
- ☆33Nov 14, 2025Updated 10 months ago
- This is the official repo for the Liver Kidney Stomach dataset introduced in the paper: SOS: Selective Objective Switch for Rapid Immunof…☆17Aug 26, 2020Updated 6 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆70Jun 18, 2025Updated last year
- Source code and additional results for GLOD issues☆12Jan 19, 2023Updated 3 years ago
- Official Repo for SparseLLM: Global Pruning of LLMs (NeurIPS 2024)☆71Mar 27, 2025Updated last year
- [ICLR 2026] "Co-rewarding: Stable Self-supervised RL for Eliciting Reasoning in Large Language Models"☆60Feb 4, 2026Updated 7 months ago
- The official repo for the paper "Teacher Forcing Recovers Reward Functions for Text Generation"☆31May 27, 2023Updated 3 years ago
- ☆15Mar 2, 2023Updated 3 years ago
- [NeurIPS 2024] "Can Language Models Perform Robust Reasoning in Chain-of-thought Prompting with Noisy Rationales?"☆41Jul 18, 2025Updated last year