☆50Feb 4, 2026Updated 7 months ago
Alternatives and similar repositories for test_time_recursive_thinking
Users that are interested in test_time_recursive_thinking are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems (ICLR'26)☆28Nov 3, 2025Updated 10 months ago
- Official implementation for Text Generation Beyond Discrete Token Sampling☆26Aug 11, 2025Updated last year
- ☆15Feb 10, 2026Updated 6 months ago
- Official Implementation of Knowledge Flow Prompting☆35Oct 20, 2025Updated 10 months ago
- Official implementation of Vector-ICL: In-context Learning with Continuous Vector Representations (ICLR 2025)☆25Jun 2, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback☆31Mar 30, 2026Updated 5 months ago
- Benchmark Test-Time Scaling of General LLM Agents☆24Apr 14, 2026Updated 4 months ago
- [COLM 2026] Resa: Transparent Reasoning Models via SAEs☆49Sep 23, 2025Updated 11 months ago
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 4 months ago
- Official implementation of Self-Taught Agentic Long Context Understanding (ACL 2025).☆14Sep 22, 2025Updated 11 months ago
- Code for "Language Models Can Learn from Verbal Feedback Without Scalar Rewards"☆65Jan 5, 2026Updated 8 months ago
- ☆26Oct 29, 2025Updated 10 months ago
- 🧠Plan-and-Budget: Training-free test-time reasoning framework for adaptive token allocation in large language models (ICLR 2026).☆16Mar 2, 2026Updated 6 months ago
- ☆18Aug 15, 2026Updated 3 weeks ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆17Apr 10, 2025Updated last year
- Marketplace ML experiment - training without backprop☆28Sep 9, 2025Updated 11 months ago
- SR²AM: Efficient Agentic Reasoning Through Self-Regulated Simulative Planning☆21May 22, 2026Updated 3 months ago
- ThetaEvolve: Test-time Learning on Open Problems, enabling RL training on AlphaEvolve/OpenEvolve and emphasizing scaling test-time comput…☆176Feb 27, 2026Updated 6 months ago
- ☆16Feb 4, 2026Updated 7 months ago
- Sci-Reasoning: A Dataset Decoding AI Innovation Patterns☆19Jan 13, 2026Updated 7 months ago
- Storing the LongCoT-mini results for RLM(GPT-5.2)☆20Apr 26, 2026Updated 4 months ago
- [CIKM 2022] Towards Automated Over-Sampling for Imbalanced Classification☆10Mar 20, 2023Updated 3 years ago
- ☆18Aug 4, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆20Sep 16, 2025Updated 11 months ago
- Synthetic Data Generation with Execution-Based Verification and Grounding for LLM Training.☆23Feb 7, 2025Updated last year
- Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning☆25Jun 25, 2025Updated last year
- Code and models for EMNLP 2024 paper "WPO: Enhancing RLHF with Weighted Preference Optimization"☆41Sep 24, 2024Updated last year
- Magnetic resonance imaging (MRI) images are known to be sparse. This is an implementation using non-convex penalty function that encourag…☆19Aug 10, 2019Updated 7 years ago
- Meta-Reinforcement Learning with Self-Reflection☆34Mar 26, 2026Updated 5 months ago
- ☆13May 6, 2025Updated last year
- [ICML 2025] Beyond Bradley-Terry Models: A General Preference Model for Language Model Alignment (https://arxiv.org/abs/2410.02197)☆44Jun 15, 2026Updated 2 months ago
- Web app using Pyodide to demo different types of Scikit-learn classifiers☆13Apr 16, 2022Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- This is the official implementation of paper "The Choice of Divergence: A Neglected Key to Mitigating Diversity Collapse in Reinforcement…☆20Feb 10, 2026Updated 6 months ago
- FrontierSmith, a new system that uses AI to synthesize open-ended coding problems at scale☆52May 30, 2026Updated 3 months ago
- [NeurIPS 2022] Supervising the Multi-Fidelity Race of Hyperparameter Configurations☆14Apr 25, 2023Updated 3 years ago
- Reinforcing General Reasoning without Verifiers☆102Jun 24, 2025Updated last year
- A Multi-objective Multi-fidelity acquisition function for Bayesian optimization based on EHVI method.☆14May 18, 2022Updated 4 years ago
- ☆30Jan 31, 2026Updated 7 months ago
- Source codes for "Preference-grounded Token-level Guidance for Language Model Fine-tuning" (NeurIPS 2023).☆17Jan 8, 2025Updated last year