β50Feb 4, 2026Updated 6 months ago
Alternatives and similar repositories for test_time_recursive_thinking
Users that are interested in test_time_recursive_thinking are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A logical, reasonably standardized, but flexible project structure for conducting ml research πͺβ19Apr 9, 2026Updated 4 months ago
- EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems (ICLR'26)β28Nov 3, 2025Updated 9 months ago
- Official implementation for Text Generation Beyond Discrete Token Samplingβ26Aug 11, 2025Updated last year
- Official Implementation of Knowledge Flow Promptingβ35Oct 20, 2025Updated 9 months ago
- Official implementation of Vector-ICL: In-context Learning with Continuous Vector Representations (ICLR 2025)β24Jun 2, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedbackβ28Mar 30, 2026Updated 4 months ago
- Benchmark Test-Time Scaling of General LLM Agentsβ21Apr 14, 2026Updated 4 months ago
- [COLM 2026] Resa: Transparent Reasoning Models via SAEsβ49Sep 23, 2025Updated 10 months ago
- Tiny evaluation of leading LLMs on competitive programming problemsβ14Apr 10, 2026Updated 4 months ago
- [ICLR 2026] Thinking on the Fly: Test-Time Reasoning Enhancement via Latent Thought Policy Optimizationβ32Mar 6, 2026Updated 5 months ago
- Official implementation of Self-Taught Agentic Long Context Understanding (ACL 2025).β14Sep 22, 2025Updated 10 months ago
- Code for "Language Models Can Learn from Verbal Feedback Without Scalar Rewards"β65Jan 5, 2026Updated 7 months ago
- π§ Plan-and-Budget: Training-free test-time reasoning framework for adaptive token allocation in large language models (ICLR 2026).β15Mar 2, 2026Updated 5 months ago
- β16Apr 14, 2026Updated 4 months ago
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Marketplace ML experiment - training without backpropβ28Sep 9, 2025Updated 11 months ago
- β17Apr 10, 2025Updated last year
- SRΒ²AM: Efficient Agentic Reasoning Through Self-Regulated Simulative Planningβ21May 22, 2026Updated 2 months ago
- β24May 17, 2026Updated 2 months ago
- ThetaEvolve: Test-time Learning on Open Problems, enabling RL training on AlphaEvolve/OpenEvolve and emphasizing scaling test-time computβ¦β173Feb 27, 2026Updated 5 months ago
- β17Feb 4, 2026Updated 6 months ago
- Sci-Reasoning: A Dataset Decoding AI Innovation Patternsβ19Jan 13, 2026Updated 7 months ago
- Storing the LongCoT-mini results for RLM(GPT-5.2)β20Apr 26, 2026Updated 3 months ago
- [CIKM 2022] Towards Automated Over-Sampling for Imbalanced Classificationβ10Mar 20, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- β19Aug 4, 2025Updated last year
- [IJCAI'23] Speeding Up Multi-Objective Hyperparameter Optimization by Task Similarity-Based Meta-Learning for the Tree-Structured Parzen β¦β10Apr 24, 2026Updated 3 months ago
- MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Followingβ16Oct 31, 2024Updated last year
- β20Sep 16, 2025Updated 10 months ago
- Synthetic Data Generation with Execution-Based Verification and Grounding for LLM Training.β22Feb 7, 2025Updated last year
- Python Interface of the Scalable Bayesian Rule Listsβ20Feb 2, 2020Updated 6 years ago
- Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoningβ25Jun 25, 2025Updated last year
- Code and models for EMNLP 2024 paper "WPO: Enhancing RLHF with Weighted Preference Optimization"