β50Feb 4, 2026Updated 7 months ago
Alternatives and similar repositories for test_time_recursive_thinking
Users that are interested in test_time_recursive_thinking are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A logical, reasonably standardized, but flexible project structure for conducting ml research πͺβ19Apr 9, 2026Updated 5 months ago
- EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems (ICLR'26)β32Nov 3, 2025Updated 10 months ago
- Official implementation for Text Generation Beyond Discrete Token Samplingβ26Aug 11, 2025Updated last year
- β15Feb 10, 2026Updated 7 months ago
- Official Implementation of Knowledge Flow Promptingβ35Oct 20, 2025Updated 11 months ago
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedbackβ31Mar 30, 2026Updated 5 months ago
- [COLM 2026] Resa: Transparent Reasoning Models via SAEsβ49Sep 23, 2025Updated last year
- Tiny evaluation of leading LLMs on competitive programming problemsβ14Apr 10, 2026Updated 5 months ago
- [ICLR 2026] Thinking on the Fly: Test-Time Reasoning Enhancement via Latent Thought Policy Optimizationβ32Mar 6, 2026Updated 6 months ago
- Official implementation of Self-Taught Agentic Long Context Understanding (ACL 2025).β14Sep 22, 2025Updated last year
- Code for "Language Models Can Learn from Verbal Feedback Without Scalar Rewards"β65Jan 5, 2026Updated 8 months ago
- β28Oct 29, 2025Updated 10 months ago
- π§ Plan-and-Budget: Training-free test-time reasoning framework for adaptive token allocation in large language models (ICLR 2026).β17Mar 2, 2026Updated 6 months ago
- MCOUT: Multimodal Chain of Continuous Thought for Latent Reasoningβ22Oct 4, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Marketplace ML experiment - training without backpropβ28Sep 9, 2025Updated last year
- β17Apr 10, 2025Updated last year
- SRΒ²AM: Efficient Agentic Reasoning Through Self-Regulated Simulative Planningβ21May 22, 2026Updated 4 months ago
- β30May 17, 2026Updated 4 months ago
- ThetaEvolve: Test-time Learning on Open Problems, enabling RL training on AlphaEvolve/OpenEvolve and emphasizing scaling test-time computβ¦β180Feb 27, 2026Updated 7 months ago
- β17Feb 4, 2026Updated 7 months ago
- Sci-Reasoning: A Dataset Decoding AI Innovation Patternsβ19Jan 13, 2026Updated 8 months ago
- Storing the LongCoT-mini results for RLM(GPT-5.2)β20Apr 26, 2026Updated 5 months ago
- [CIKM 2022] Towards Automated Over-Sampling for Imbalanced Classificationβ10Mar 20, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- β18Aug 4, 2025Updated last year
- β11Feb 28, 2024Updated 2 years ago
- [IJCAI'23] Speeding Up Multi-Objective Hyperparameter Optimization by Task Similarity-Based Meta-Learning for the Tree-Structured Parzen β¦β10Apr 24, 2026Updated 5 months ago
- MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Followingβ16Oct 31, 2024Updated last year
- β20Sep 16, 2025Updated last year
- Synthetic Data Generation with Execution-Based Verification and Grounding for LLM Training.β24Feb 7, 2025Updated last year
- The official implementation of "ICDPO: Effectively Borrowing Alignment Capability of Others via In-context Direct Preference Optimizationβ¦β16Feb 15, 2024Updated 2 years ago
- Python Interface of the Scalable Bayesian Rule Listsβ20Feb 2, 2020Updated 6 years ago
- Code and models for EMNLP 2024 paper "WPO: Enhancing RLHF with Weighted Preference Optimization"β41Sep 24, 2024Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradientsβ21Jun 17, 2025Updated last year
- Meta-Reinforcement Learning with Self-Reflectionβ36Mar 26, 2026Updated 6 months ago
- Official Repo for "Multi-objective Differentiable Neural Architecture Search"β13Jul 12, 2024Updated 2 years ago
- β13May 6, 2025Updated last year
- β32Mar 11, 2026Updated 6 months ago
- [ICML 2025] Beyond Bradley-Terry Models: A General Preference Model for Language Model Alignment (https://arxiv.org/abs/2410.02197)β46Jun 15, 2026Updated 3 months ago
- [ICLR 2026] "Co-rewarding: Stable Self-supervised RL for Eliciting Reasoning in Large Language Models"β60Feb 4, 2026Updated 7 months ago