☆105Dec 6, 2024Updated last year
Alternatives and similar repositories for o1_Reasoning_Patterns_Study
Users that are interested in o1_Reasoning_Patterns_Study are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Accelerating the development of large multimodal models (LMMs) with lmms-eval☆14Oct 14, 2024Updated last year
- ☆19Nov 12, 2024Updated last year
- ☆42Jul 15, 2025Updated last year
- ☆1,339Nov 21, 2024Updated last year
- ☆130Jun 18, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- O1 Replication Journey☆2,000Jan 14, 2025Updated last year
- ☆967Jan 23, 2025Updated last year
- Large Reasoning Models☆799Dec 3, 2024Updated last year
- ☆553Jan 2, 2025Updated last year
- ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search (NeurIPS 2024)☆713Jan 20, 2025Updated last year
- ☆342Jun 5, 2025Updated last year
- ☆42Jan 14, 2025Updated last year
- HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models☆60Nov 26, 2024Updated last year
- Composable inference algorithms with LLMs and programmable logic☆69Dec 4, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Repository for the paper Stream of Search: Learning to Search in Language☆152Feb 3, 2025Updated last year
- OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models☆1,859Jan 17, 2025Updated last year
- [ACL 2024] Making Long-Context Language Models Better Multi-Hop Reasoners☆20May 28, 2024Updated 2 years ago
- Scalable RL solution for advanced reasoning of language models☆1,871Mar 18, 2025Updated last year
- ☆127Jun 2, 2026Updated 4 months ago
- A series of technical report on Slow Thinking with LLM☆769Aug 13, 2025Updated last year
- ☆52Nov 14, 2024Updated last year
- The original Shared Recurrent Memory Transformer implementation☆39Aug 24, 2026Updated last month
- A scalable automated alignment method for large language models. Resources for "Aligning Large Language Models via Self-Steering Optimiza…☆20Nov 21, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Systematic evaluation framework that automatically rates overthinking behavior in large language models.☆103May 16, 2025Updated last year
- Recipes to train the self-rewarding reasoning LLMs.☆230Mar 2, 2025Updated last year
- ☆15Apr 26, 2025Updated last year
- An automated data pipeline scaling RL to pretraining levels☆76Jun 2, 2026Updated 4 months ago
- Benchmarking Complex Instruction-Following with Multiple Constraints Composition (NeurIPS 2024 Datasets and Benchmarks Track)☆104Feb 20, 2025Updated last year
- Understanding R1-Zero-Like Training: A Critical Perspective☆1,280Aug 27, 2025Updated last year
- Codes and Data for Scaling Relationship on Learning Mathematical Reasoning with Large Language Models☆269Sep 12, 2024Updated 2 years ago
- This is the official repo for Towards Uncertainty-Aware Language Agent.☆31Aug 15, 2024Updated 2 years ago
- SimpleDeepSearcher: Deep Information Seeking via Web-Powered Reasoning Trajectory Synthesis☆121Jun 3, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- official repo for the paper "Learning From Mistakes Makes LLM Better Reasoner"☆62Dec 20, 2023Updated 2 years ago
- Official Repo for Open-Reasoner-Zero☆2,099Jun 2, 2025Updated last year
- ☆22Sep 19, 2024Updated 2 years ago
- The official repository of the Omni-MATH benchmark.☆92Dec 22, 2024Updated last year
- ☆1,193Jan 10, 2026Updated 9 months ago
- ☆55Oct 10, 2024Updated 2 years ago
- Code for "Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate" [COLM 2025]☆180Jul 8, 2025Updated last year