[ICLR'26] "Nabla-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space" by Peihao Wang*, Ruisi Cai*, Zhen Wang, Hongyuan Mei, Qiang Liu, Pan Li, Zhangyang Wang
☆35Mar 10, 2026Updated 5 months ago
Alternatives and similar repositories for Nabla-Reasoner
Users that are interested in Nabla-Reasoner are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of Latent-GRPO: reinforcement learning for vocabulary-space latent reasoning.☆18May 12, 2026Updated 3 months ago
- The offical repo for "Parallel-Probe: Towards Efficient Parallel Thinking via 2D Probing"☆19Feb 3, 2026Updated 6 months ago
- [ICML'26] Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning☆16Jun 1, 2026Updated 2 months ago
- Code for the SofT-GRPO algorithm on the LLM soft-thinking reasoning pattern.☆52Jan 2, 2026Updated 7 months ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2026] Thinking on the Fly: Test-Time Reasoning Enhancement via Latent Thought Policy Optimization☆32Mar 6, 2026Updated 5 months ago
- Official repository for distributing ECG-Reasoning-Benchmark dataset