[ACL'26 Findings] Steering LLM Thinking with Budget Guidance
☆34Feb 19, 2026Updated 7 months ago
Alternatives and similar repositories for BudgetGuidance
Users that are interested in BudgetGuidance are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆27May 14, 2026Updated 4 months ago
- Official code for paper "Revisiting Model Interpolation for Efficient Reasoning"☆17Jul 14, 2026Updated 2 months ago
- ☆29Aug 8, 2025Updated last year
- [NeurIPS 2025] The implementation of paper "On Reasoning Strength Planning in Large Reasoning Models"☆33Jul 6, 2025Updated last year
- [NeurIPS 2025] Reasoning Models Better Express Their Confidence"☆23Nov 19, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [NeurIPS 2025] Official Implementation of paper "Sherlock: Self-Correcting Reasoning in Vision-Language Models"☆31Jun 4, 2026Updated 3 months ago
- ☆76Jun 10, 2025Updated last year
- The official repo for “Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem” [EMNLP25]☆33Sep 1, 2025Updated last year
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 5 months ago
- [COLM 2026] An efficient 3D sampling method for long-CoT LLM.☆16May 25, 2025Updated last year
- [COLM 2025] SEAL: Steerable Reasoning Calibration of Large Language Models for Free☆65Apr 6, 2025Updated last year
- [ICLR 2026] RuleReasoner: Reinforced Rule-based Reasoning via Domain-aware Dynamic Sampling☆42Feb 25, 2026Updated 7 months ago
- A model serving framework for various research and production scenarios. Seamlessly built upon the PyTorch and HuggingFace ecosystem.☆23Oct 11, 2024Updated last year
- [COLM 2026] Resa: Transparent Reasoning Models via SAEs☆49Sep 23, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CCS 2026] The official implementation of our CCS 2026 paper "ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathological…☆18Aug 5, 2026Updated last month
- Reinforcement Learning from Text Feedback☆49Feb 17, 2026Updated 7 months ago
- ☆57Jul 7, 2025Updated last year
- Machine Bullshit: Characterizing the Emergent Disregard for Truth in Large Language Models☆26Sep 14, 2025Updated last year
- Measuring Thinking Efficiency in Reasoning Models - Research Repository☆40Dec 2, 2025Updated 10 months ago
- ☆27May 12, 2026Updated 4 months ago
- Revisiting Mid-training in the Era of Reinforcement Learning Scaling☆188Jul 23, 2025Updated last year
- Ongoing research project for code&math LLMs☆32Jul 4, 2025Updated last year
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆276Updated this week
- ☆15Jan 27, 2025Updated last year
- Gemini evaluations☆28Sep 3, 2026Updated last month
- A Recipe for Building LLM Reasoners to Solve Complex Instructions☆32Oct 9, 2025Updated 11 months ago
- ☆74Feb 4, 2026Updated 7 months ago
- ☆47Nov 25, 2024Updated last year
- [NeurIPS'25] Beyond Accuracy: Dissecting Mathematical Reasoning for LLMs Under Reinforcement Learning☆17Dec 12, 2025Updated 9 months ago
- Agent ADA is a comprehensive evaluation and data analytics framework focused on insights generation and skills assessment.☆15Aug 19, 2025Updated last year
- ☆17Sep 11, 2026Updated 3 weeks ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- a phone to app implementation using Asterisk and Node.js☆15Oct 27, 2014Updated 11 years ago
- Efficient non-uniform quantization with GPTQ for GGUF☆66Sep 17, 2025Updated last year
- ☆23Jun 12, 2025Updated last year
- [EMNLP 25] An effective and interpretable weight-editing method for mitigating overly short reasoning in LLMs, and a mechanistic study un…☆20Aug 24, 2026Updated last month
- [Tech Report] Expanded Hyper-Connections☆68Jul 21, 2026Updated 2 months ago
- DCPO: Dynamic Adaptive Clipping for RL☆50Apr 1, 2026Updated 6 months ago
- OnePlus 8T Param Read/Write☆14Dec 4, 2020Updated 5 years ago