Code, Data and Model for Paper "Learning from Peers in Reasoning Models"
☆26May 13, 2025Updated last year
Alternatives and similar repositories for LeaP
Users that are interested in LeaP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A list of Numerical Multimodal reasoning papers and their implementation☆11May 13, 2024Updated 2 years ago
- Code for COLING 2022 long paper: Answering Numerical Reasoning Questions in Table-Text Hybrid Contents with Graph-based Encoder and Tree-…☆22Dec 15, 2022Updated 3 years ago
- Code and Model for NeurIPS 2024 Spotlight Paper "Stacking Your Transformers: A Closer Look at Model Growth for Efficient LLM Pre-Training…☆44Oct 16, 2024Updated last year
- Implementation of SelfExtend from the paper "LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning" from Pytorch and Zeta☆13Nov 11, 2024Updated last year
- Code and Data for Paper "AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning"☆54Sep 4, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The official implementation of Preference Data Reward-Augmentation.☆18May 1, 2025Updated last year
- [ICLR 2025] Bridging and Modeling Correlations in Pairwise Data for Direct Preference Optimization☆12Jan 26, 2025Updated last year
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"☆17Jun 30, 2025Updated last year
- A comprehensive paper list of Table-based Question Answering.☆40Sep 1, 2023Updated 3 years ago
- ☆14Sep 22, 2025Updated last year
- NeurIPS 2025: Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs☆66Nov 21, 2025Updated 10 months ago
- 🔥 [ICML'26] ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs☆36Aug 14, 2026Updated last month
- Your finetuned model's back to its original safety standards faster than you can say "SafetyLock"!☆11Oct 16, 2024Updated last year
- Towards Fine-grained Audio Captioning with Multimodal Contextual Cues☆90Jan 4, 2026Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code and Data for EMNLP 2023 Paper "MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Langu…☆14Apr 7, 2025Updated last year
- ☆34Oct 13, 2025Updated 11 months ago
- ☆14Apr 1, 2024Updated 2 years ago
- ☆20Mar 18, 2026Updated 6 months ago
- ☆25Feb 18, 2025Updated last year
- [NeurIPS 2026] When Reasoning Meets Its Laws☆38Jan 2, 2026Updated 8 months ago
- MyPhoneBench: Do Phone-Use Agents Respect Your Privacy?☆24Apr 3, 2026Updated 5 months ago
- We introduce EMMET and unify model editing with popular algorithms ROME and MEMIT.☆29Dec 16, 2024Updated last year
- Reinforcing Long-Term Performance in Recommender Systems with User-Oriented Exploration Policy (SIGIR 2024)☆14Oct 6, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP Main 2026]VTC-R1: Vision-Text Compression for Efficient Long-Context Reasoning.☆26Jul 20, 2026Updated 2 months ago
- SQuAD Question Generation module based on T5-large☆17Aug 26, 2022Updated 4 years ago
- ACDiT: Interpolating Autoregressive Conditional Modeling and Diffusion Transformer☆42Jan 29, 2026Updated 7 months ago
- Code for SafeMERGE (ICLR 2025).☆15Apr 1, 2025Updated last year
- [ACL 2025 (Findings)] DEMO: Reframing Dialogue Interaction with Fine-grained Element Modeling☆22Dec 16, 2024Updated last year
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- Python library providing a simple, fully supervised sentence embedding technique for textual adversarial attacks.☆13Dec 13, 2023Updated 2 years ago
- ☆47Mar 4, 2025Updated last year
- ☆31Sep 12, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for the paper "Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns"☆19Mar 15, 2024Updated 2 years ago
- Evaluating the faithfulness of long-context language models☆30Oct 21, 2024Updated last year
- RL with Experience Replay☆59Jul 27, 2025Updated last year
- Official code implementation for the ACL 2025 paper: 'Dynamic Scaling of Unit Tests for Code Reward Modeling'☆26May 16, 2025Updated last year
- [ACL 2024] Masked Thought: Simply Masking Partial Reasoning Steps Can Improve Mathematical Reasoning Learning of Language Models☆27Jul 9, 2024Updated 2 years ago
- [ICML 2025] Official code of "AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization"☆32Jan 10, 2026Updated 8 months ago
- 复旦研究生抢课脚本☆12Feb 14, 2022Updated 4 years ago