☆48Nov 1, 2025Updated 10 months ago
Alternatives and similar repositories for ParaThinker
Users that are interested in ParaThinker are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The offical repo for "Parallel-R1: Towards Parallel Thinking via Reinforcement Learning"☆266Feb 4, 2026Updated 7 months ago
- Official implementation for paper "How Far Are We from Genuinely Useful Deep Research Agents?"☆66Dec 10, 2025Updated 9 months ago
- Short RL☆19Apr 16, 2026Updated 5 months ago
- ☆20Jun 17, 2024Updated 2 years ago
- ☆15Jan 14, 2026Updated 8 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆47Apr 9, 2025Updated last year
- CIKM 2022: Evaluating Interpolation and Extrapolation Performance of Neural Retrieval Models☆10Aug 4, 2022Updated 4 years ago
- [NeurIPS 2025] A*-Thought: Efficient Reasoning via Bidirectional Compression for Low-Resource Settings☆28Sep 9, 2026Updated last week
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"☆17Jun 30, 2025Updated last year
- ☆29Sep 11, 2025Updated last year
- ☆21May 16, 2024Updated 2 years ago
- Official code release for Delta Activations: A Representation for Finetuned Large Language Models☆21Sep 5, 2025Updated last year
- [ICML 2026] Reasoning in Parallelism via Self-Distilled RL☆113Jun 28, 2026Updated 2 months ago
- ☆88Jun 16, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 这是我的博客《不用框架,使用Python搭建基于numpy的卷积神经网络来进行cifar-10分类的深度学习系统》的代码实现。☆10Jul 1, 2019Updated 7 years ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆17Apr 30, 2025Updated last year
- Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents☆31Apr 16, 2026Updated 5 months ago
- Codebase for the work “Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs?”☆77Apr 14, 2026Updated 5 months ago
- Official code for the paper: "Multi-User Large Language Model Agents"☆35Sep 1, 2026Updated 3 weeks ago
- ☆12Feb 27, 2025Updated last year
- [COLM 2025] Code for Paper: Learning Adaptive Parallel Reasoning with Language Models☆145Dec 17, 2025Updated 9 months ago
- [EMNLP 2026 Main] Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policies☆60Feb 6, 2026Updated 7 months ago
- The official github repo for "Training Optimal Large Diffusion Language Models", the first-ever large-scale diffusion language models sca…☆46Nov 6, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Fork of Flame repo for training of some new stuff in development☆20Aug 27, 2026Updated 3 weeks ago
- The official repository of "Document Image Machine Translation with Dynamic Multi-pre-trained Models Assembling"☆14Nov 26, 2025Updated 9 months ago
- Official implementation of PolySkill, a framework that enables web agents to learn generalizable and compositional skills through polymor…☆18Jul 6, 2026Updated 2 months ago
- This is the official implementation of the paper "S²R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning"☆78Apr 22, 2025Updated last year
- [ACL'26] Official Repository for for paper "Data-Efficient RLVR via Off-Policy Influence Guidance"☆25Jul 26, 2026Updated last month
- ☆24Jun 16, 2026Updated 3 months ago
- Code for paper "The Markovian Thinker: Architecture-Agnostic Linear Scaling of Reasoning"☆350Mar 16, 2026Updated 6 months ago
- [ICLR'26] RM-R1: Unleashing the Reasoning Potential of Reward Models☆171Jun 26, 2025Updated last year
- ☆26Dec 16, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Lightweight Non-Parametric Embedding Fine-Tuning☆43Sep 13, 2025Updated last year
- Optimizing Anytime Reasoning via Budget Relative Policy Optimization☆54Jul 15, 2025Updated last year
- ☆44Mar 31, 2026Updated 5 months ago
- [NeurIPS 2024] Goldfish Loss: Mitigating Memorization in Generative LLMs☆98Nov 17, 2024Updated last year
- Code for "DynaGuard: A Dynamic Guardrail Model With User-Defined Policies."☆25Nov 3, 2025Updated 10 months ago
- Source code for our paper: "ARIA: Training Language Agents with Intention-Driven Reward Aggregation".☆30Aug 9, 2025Updated last year
- Synthetic data generation for TODs☆23Jul 17, 2024Updated 2 years ago