Code for the SofT-GRPO algorithm on the LLM soft-thinking reasoning pattern.
☆52Jan 2, 2026Updated 8 months ago
Alternatives and similar repositories for SofT-GRPO-master
Users that are interested in SofT-GRPO-master are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Feb 4, 2026Updated 6 months ago
- Code for [ICML2025]``Monte Carlo Tree Search for Comprehensive Exploration in LLM-Based Automatic Heuristic Design``.☆85May 23, 2025Updated last year
- Official implementation of Latent-SFT: teaching LLMs to reason with vocabulary-space latent chains.☆59Updated this week
- Official implementation of Latent-GRPO: reinforcement learning for vocabulary-space latent reasoning.☆19Updated this week
- [ICLR'26] "Nabla-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space" by Peihao Wang*, Ruisi Cai*, Zhen Wang, Hongyuan…☆35Mar 10, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repository contains a regularly updated paper list for LLMs-reasoning-in-latent-space.☆378Jun 20, 2026Updated 2 months ago
- [ICLR 2026] Thinking on the Fly: Test-Time Reasoning Enhancement via Latent Thought Policy Optimization☆32Mar 6, 2026Updated 5 months ago
- Parallel Continuous Chain-of-Thought with Jacobi Iteration. Accepted to EMNLP 2025.☆25Mar 29, 2026Updated 5 months ago
- [ACL 2026 oral] SeLaR: Selective Latent Reasoning in Large Language Models☆22Apr 25, 2026Updated 4 months ago
- Multiplex Thinking: Reasoning via Token-wise Branch-and-Merge☆135May 24, 2026Updated 3 months ago
- [ACL 2026] Dissecting Failure Dynamics in Large Language Model Reasoning☆18Apr 17, 2026Updated 4 months ago
- ACL'2025: SoftCoT: Soft Chain-of-Thought for Efficient Reasoning with LLMs. and preprint: SoftCoT++: Test-Time Scaling with Soft Chain-of…☆95May 30, 2025Updated last year
- [NeurIPS 2025] Official code for paper: Latent Chain-of-Thought for Visual Reasoning☆36Oct 16, 2025Updated 10 months ago
- ☆16Apr 15, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ACL2026 Findings] "Towards Hierarchical Multi-Step Reward Models for Enhanced Reasoning in Large Language Models"☆20Mar 25, 2025Updated last year
- Self-Hinting Language Models Enhance Reinforcement Learning☆28Mar 28, 2026Updated 5 months ago
- [ICLR 2026] SwiReasoning: Switch-Thinking in Latent and Explicit for Pareto-Superior Reasoning LLMs☆139Aug 2, 2026Updated last month
- ☆25Feb 18, 2025Updated last year
- Learning MLPs to replace GNN☆10Jun 3, 2023Updated 3 years ago
- Demonstrates failures of bias mitigation methods under varying types/levels of biases (WACV 2021)☆26Mar 31, 2024Updated 2 years ago
- Three-level Hierarchical Transformer Networks for Long-sequence and Multiple Clinical Documents Classification☆11Apr 7, 2022Updated 4 years ago
- Under construction☆14Jan 15, 2025Updated last year
- PhysGame Benchmark for Physical Commonsense Evaluation in Gameplay Videos☆49Jul 3, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- NJU OS lab 2023☆12Apr 26, 2023Updated 3 years ago
- ☆51Jan 30, 2026Updated 7 months ago
- ☆21May 30, 2025Updated last year
- ☆14Jun 3, 2025Updated last year
- Kaggle AIMO2 solution with token-efficient reasoning LLM recipes☆51Aug 7, 2025Updated last year
- 王者荣耀开悟平台——重返秘境☆15Nov 23, 2025Updated 9 months ago
- Official PyTorch implementation of our CVPR 2025 paper, "LoRA Subtraction for Drift-Resistant Space in Exemplar-Free Continual Learning."☆18Mar 28, 2025Updated last year
- brain to speech☆13Mar 17, 2026Updated 5 months ago
- ☆13Mar 26, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A Python reimplementation + extension of "Planning with Large Language Models for Code Generation" (https://arxiv.org/abs/2303.05510)☆17Dec 1, 2023Updated 2 years ago
- Official implementation of "Reasoning by Superposition: A Theoretical Perspective on Chain of Continuous Thought" (NeurIPS 2025)☆44Oct 8, 2025Updated 10 months ago
- [ICLR 2026 Oral] TRACE: Your Diffusion Model Is Secretly an Instance Edge Detector☆18Mar 2, 2026Updated 6 months ago
- ☆37May 29, 2025Updated last year
- EA codes from CIAM Group at SUSTech, Shenzhen, China☆21Apr 18, 2025Updated last year
- The source code for the paper: Yirong Mao, Ruiping Wang, Shiguang Shan, Xilin Chen. COSONet: Compact Second-Order Network for Video Face …☆12Dec 27, 2018Updated 7 years ago
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Maps☆13Mar 26, 2025Updated last year