[ACL'26 Findings] Steering LLM Thinking with Budget Guidance
☆34Feb 19, 2026Updated 6 months ago
Alternatives and similar repositories for BudgetGuidance
Users that are interested in BudgetGuidance are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code for paper "Revisiting Model Interpolation for Efficient Reasoning"☆17Jul 14, 2026Updated last month
- ☆29Aug 8, 2025Updated last year
- [NeurIPS 2025] The implementation of paper "On Reasoning Strength Planning in Large Reasoning Models"☆31Jul 6, 2025Updated last year
- [NeurIPS 2025] Reasoning Models Better Express Their Confidence"☆23Nov 19, 2025Updated 9 months ago
- ☆75Jun 10, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- The official repo for “Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem” [EMNLP25]☆33Sep 1, 2025Updated 11 months ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 4 months ago
- [COLM 2026] An efficient 3D sampling method for long-CoT LLM.☆16May 25, 2025Updated last year
- [ICLR 2026] RuleReasoner: Reinforced Rule-based Reasoning via Domain-aware Dynamic Sampling☆40Feb 25, 2026Updated 5 months ago
- A model serving framework for various research and production scenarios. Seamlessly built upon the PyTorch and HuggingFace ecosystem.☆23Oct 11, 2024Updated last year
- [COLM 2026] Resa: Transparent Reasoning Models via SAEs☆49Sep 23, 2025Updated 11 months ago
- ☆30Jun 5, 2025Updated last year
- [CCS 2026] The official implementation of our CCS 2026 paper "ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathological…☆16Aug 5, 2026Updated 2 weeks ago
- Reinforcement Learning from Text Feedback☆47Feb 17, 2026Updated 6 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [WACV2023] This is the official PyTorch impelementation of our paper "[Rethinking Rotation in Self-Supervised Contrastive Learning: Adapt…☆12Feb 24, 2023Updated 3 years ago
- ☆56Jul 7, 2025Updated last year
- Machine Bullshit: Characterizing the Emergent Disregard for Truth in Large Language Models☆27Sep 14, 2025Updated 11 months ago
- Measuring Thinking Efficiency in Reasoning Models - Research Repository☆40Dec 2, 2025Updated 8 months ago
- Ongoing research project for code&math LLMs☆31Jul 4, 2025Updated last year
- ☆27May 12, 2026Updated 3 months ago
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated last year
- ☆15Jan 27, 2025Updated last year
- Gemini evaluations☆26Feb 19, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Source Code for our ICLR'26 paper☆17Feb 22, 2026Updated 6 months ago
- ☆70Feb 4, 2026Updated 6 months ago
- A Recipe for Building LLM Reasoners to Solve Complex Instructions☆32Oct 9, 2025Updated 10 months ago
- ☆47Nov 25, 2024Updated last year
- Multichannel Looper/Feedback System for Riffusion☆14May 6, 2023Updated 3 years ago
- [NeurIPS'25] Beyond Accuracy: Dissecting Mathematical Reasoning for LLMs Under Reinforcement Learning☆16Dec 12, 2025Updated 8 months ago
- ☆17Jul 31, 2025Updated last year
- Agent ADA is a comprehensive evaluation and data analytics framework focused on insights generation and skills assessment.☆15Aug 19, 2025Updated last year
- An Advanced Basic Math Reasoning and Overthinking Evaluation Framework for LLMs☆12Apr 20, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆17Apr 25, 2025Updated last year
- a phone to app implementation using Asterisk and Node.js☆15Oct 27, 2014Updated 11 years ago
- Efficient non-uniform quantization with GPTQ for GGUF☆64Sep 17, 2025Updated 11 months ago
- ☆22Jun 12, 2025Updated last year
- [Tech Report] Expanded Hyper-Connections☆61Jul 21, 2026Updated last month
- DCPO: Dynamic Adaptive Clipping for RL☆49Apr 1, 2026Updated 4 months ago
- ☆23Jun 16, 2026Updated 2 months ago