☆20Mar 18, 2026Updated 4 months ago
Alternatives and similar repositories for SplitReason
Users that are interested in SplitReason are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15May 27, 2025Updated last year
- ☆34Oct 13, 2025Updated 9 months ago
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"☆17Jun 30, 2025Updated last year
- CoMM: Collaborative Multi-Agent, Multi-Reasoning-Path Prompting for Complex Problem Solving (NAACL 2024 Findings))☆16Apr 26, 2024Updated 2 years ago
- Continuous Pipelined Speculative Decoding☆21May 25, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- RL with Experience Replay☆59Jul 27, 2025Updated 11 months ago
- Distributed Optimization Infra for learning CLIP models☆31Oct 3, 2024Updated last year
- Code for paper: Optimizing Length Compression in Large Reasoning Models☆29Oct 20, 2025Updated 9 months ago
- Archer2.0 evolves from its predecessor by introducing ASPO, which overcomes fundamental PPO-Clip limitations to prevent premature converg…☆31Oct 10, 2025Updated 9 months ago
- PoC for "SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning" [NeurIPS '25]☆74Oct 2, 2025Updated 9 months ago
- ☆65Mar 30, 2026Updated 3 months ago
- ☆11Feb 16, 2024Updated 2 years ago
- Official Implementation of wd1☆32Sep 25, 2025Updated 9 months ago
- ☆46Sep 27, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- APAR: LLMs Can Do Auto-Parallel Auto-Regressive Decoding☆14Jul 22, 2024Updated 2 years ago
- ☆15Nov 7, 2024Updated last year
- ☆20May 14, 2025Updated last year
- [NeurIPS 2025] The implementation of paper "On Reasoning Strength Planning in Large Reasoning Models"☆31Jul 6, 2025Updated last year
- [EMNLP 2025 Main] AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time☆89Jun 10, 2025Updated last year
- When Reasoning Meets Its Laws☆38Jan 2, 2026Updated 6 months ago
- [ICLR 2025] "GraphEval: A Lightweight Graph-Based LLM Framework for Idea Evaluation", Tao Feng, Yihang Sun, Jiaxuan You☆18Mar 18, 2025Updated last year
- ☆24Feb 18, 2025Updated last year
- We introduce EMMET and unify model editing with popular algorithms ROME and MEMIT.☆29Dec 16, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code, Data and Model for Paper "Learning from Peers in Reasoning Models"☆26May 13, 2025Updated last year
- An Ultra-Long Output Reinforcement Learning Approach☆23Jul 31, 2025Updated 11 months ago
- ☆18Mar 2, 2026Updated 4 months ago
- [EMNLP 2025] LightThinker: Thinking Step-by-Step Compression☆165Jun 22, 2026Updated last month
- [ICLR 2026] InftyThink: Breaking the Length Limits of Long-Context Reasoning in Large Language Models☆57May 5, 2026Updated 2 months ago
- Code for SafeMERGE (ICLR 2025).☆15Apr 1, 2025Updated last year
- Official implementation of Self-Taught Agentic Long Context Understanding (ACL 2025).☆13Sep 22, 2025Updated 10 months ago
- The code implementation of Skill-MoE☆46May 22, 2026Updated 2 months ago
- AdaptiveStep: Automatically Dividing Reasoning Step through Model Confidence☆10Mar 2, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official implementation of "Can Test-Time Scaling Improve World Foundation Model?"☆15Jul 12, 2025Updated last year
- ☆186Dec 5, 2025Updated 7 months ago
- ☆31Sep 12, 2025Updated 10 months ago
- [ICML 2025] Reward-guided Speculative Decoding (RSD) for efficiency and effectiveness.☆56May 2, 2025Updated last year
- [NeurIPS 2025@FoRLM] R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search☆17Jan 24, 2026Updated 5 months ago
- ☆39May 20, 2025Updated last year
- [ICML 2025] Official code of "AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization"☆31Jan 10, 2026Updated 6 months ago