OpenCoconut implements a latent reasoning paradigm where we generate thoughts before decoding.
โ173Jan 16, 2025Updated last year
Alternatives and similar repositories for OpenCoconut
Users that are interested in OpenCoconut are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- โ29Jan 14, 2025Updated last year
- Implementation of ๐ฅฅ Coconut, Chain of Continuous Thought, in Pytorchโ185Jun 20, 2025Updated last year
- Training Large Language Model to Reason in a Continuous Latent Spaceโ1,711Jul 2, 2026Updated 2 months ago
- โ15Apr 26, 2025Updated last year
- โ208Apr 19, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer โข AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- โ16Mar 22, 2025Updated last year
- Recipes to scale inference-time compute of open modelsโ1,132Updated this week
- โ13Mar 25, 2026Updated 5 months ago
- This repository includes a benchmark and code for the paper "Evaluating LLMs at Detecting Errors in LLM Responses".โ32Aug 18, 2024Updated 2 years ago
- โ161Apr 21, 2026Updated 5 months ago
- Archon provides a modular framework for combining different inference-time techniques and LMs with just a JSON config file.โ219Mar 7, 2025Updated last year
- Fine-tunes a student LLM using teacher feedback for improved reasoning and answer quality. Implements GRPO with teacher-provided evaluatiโฆโ54May 7, 2025Updated last year
- โ139Mar 20, 2025Updated last year
- [ICML 2026] Heimaโ76May 20, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean โข AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- MMLU-Pro eval resultsโ15Aug 21, 2025Updated last year
- Scalable RL solution for advanced reasoning of language modelsโ1,872Mar 18, 2025Updated last year
- implementation of https://arxiv.org/pdf/2312.09299โ21Jul 3, 2024Updated 2 years ago
- โ12Jul 8, 2024Updated 2 years ago
- A comprehensive repository of reasoning tasks for LLMs (and beyond)โ506Sep 27, 2024Updated last year
- [ACL 2025 Findings] Official implementation of the paper "Unveiling the Key Factors for Distilling Chain-of-Thought Reasoning".โ23Feb 26, 2025Updated last year
- Training framework with a goal to explore the frontier of sample efficiency of small language modelsโ101Jan 25, 2026Updated 7 months ago
- An introduction to LLM Samplingโ80Dec 15, 2024Updated last year
- Entropy Based Sampling and Parallel CoT Decodingโ17Oct 9, 2024Updated last year
- Managed Database hosting by DigitalOcean โข AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Repo for the IDESSAI 2024 course on modeling audio with discrete tokens.โ13Sep 13, 2024Updated 2 years ago
- Synthetic data curation for post-training and structured data extractionโ1,733Updated this week
- Post-trained LLaDA model with dynamic step schedulingโ36Mar 4, 2025Updated last year
- [๐๐๐๐๐ ๐ ๐ข๐ง๐๐ข๐ง๐ ๐ฌ ๐๐๐๐ & ๐๐๐ ๐๐๐๐ ๐๐๐๐๐ ๐๐ซ๐๐ฅ] ๐๐ฏ๐ฉ๐ข๐ฏ๐ค๐ช๐ฏ๐จ ๐๐ข๐ต๐ฉ๐ฆ๐ฎ๐ข๐ต๐ช๐ค๐ข๐ญ ๐๐ฆ๐ข๐ด๐ฐ๐ฏ๐ช๐ฏโฆโ52May 4, 2024Updated 2 years ago
- โ25Oct 10, 2025Updated 11 months ago
- skill.md specialized skill registry for AI agentsโ15Apr 24, 2026Updated 4 months ago
- โ149Nov 11, 2024Updated last year
- โ127Jun 2, 2026Updated 3 months ago
- [EMNLP-2025] R1-Zero on ANY TASKโ32Nov 9, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI โข AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Plotting (entropy, varentropy) for small LMsโ99May 20, 2025Updated last year
- Code for "Reasoning to Learn from Latent Thoughts"โ135Mar 28, 2025Updated last year
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.โ108Jul 19, 2025Updated last year
- Simple RL training for reasoningโ3,873Dec 23, 2025Updated 8 months ago
- Entropy Based Sampling and Parallel CoT Decodingโ3,430Nov 13, 2024Updated last year
- MLGym A New Framework and Benchmark for Advancing AI Research Agentsโ623Aug 10, 2025Updated last year
- Our library for RL environments + evalsโ4,639Updated this week