OpenCoconut implements a latent reasoning paradigm where we generate thoughts before decoding.
โ173Jan 16, 2025Updated last year
Alternatives and similar repositories for OpenCoconut
Users that are interested in OpenCoconut are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- โ27Jan 14, 2025Updated last year
- Implementation of ๐ฅฅ Coconut, Chain of Continuous Thought, in Pytorchโ184Jun 20, 2025Updated last year
- Training Large Language Model to Reason in a Continuous Latent Spaceโ1,690Jul 2, 2026Updated last month
- โ15Apr 26, 2025Updated last year
- โ208Apr 19, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient โข AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- โ16Mar 22, 2025Updated last year
- A curated list of resources on Reinforcement Learning with Verifiable Rewards (RLVR) and the reasoning capability boundary of Large Languโฆโ92Dec 12, 2025Updated 8 months ago
- Recipes to scale inference-time compute of open modelsโ1,131May 26, 2026Updated 3 months ago
- โ13Mar 25, 2026Updated 5 months ago
- โ158Apr 21, 2026Updated 4 months ago
- This repository includes a benchmark and code for the paper "Evaluating LLMs at Detecting Errors in LLM Responses".โ32Aug 18, 2024Updated 2 years ago
- Archon provides a modular framework for combining different inference-time techniques and LMs with just a JSON config file.โ215Mar 7, 2025Updated last year
- Fine-tunes a student LLM using teacher feedback for improved reasoning and answer quality. Implements GRPO with teacher-provided evaluatiโฆโ54May 7, 2025Updated last year
- โ139Mar 20, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer โข AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2026] Heimaโ76May 20, 2026Updated 3 months ago
- MMLU-Pro eval resultsโ15Aug 21, 2025Updated last year
- Scalable RL solution for advanced reasoning of language modelsโ1,871Mar 18, 2025Updated last year
- implementation of https://arxiv.org/pdf/2312.09299โ21Jul 3, 2024Updated 2 years ago
- A comprehensive repository of reasoning tasks for LLMs (and beyond)โ504Sep 27, 2024Updated last year
- [ACL 2025 Findings] Official implementation of the paper "Unveiling the Key Factors for Distilling Chain-of-Thought Reasoning".โ23Feb 26, 2025Updated last year
- Training framework with a goal to explore the frontier of sample efficiency of small language modelsโ101Jan 25, 2026Updated 7 months ago
- An introduction to LLM Samplingโ80Dec 15, 2024Updated last year
- Entropy Based Sampling and Parallel CoT Decodingโ17Oct 9, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer โข AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Repo for the IDESSAI 2024 course on modeling audio with discrete tokens.โ13Sep 13, 2024Updated last year
- Synthetic data curation for post-training and structured data extractionโ1,721Aug 7, 2026Updated 3 weeks ago
- Post-trained LLaDA model with dynamic step schedulingโ36Mar 4, 2025Updated last year
- [๐๐๐๐๐ ๐ ๐ข๐ง๐๐ข๐ง๐ ๐ฌ ๐๐๐๐ & ๐๐๐ ๐๐๐๐ ๐๐๐๐๐ ๐๐ซ๐๐ฅ] ๐๐ฏ๐ฉ๐ข๐ฏ๐ค๐ช๐ฏ๐จ ๐๐ข๐ต๐ฉ๐ฆ๐ฎ๐ข๐ต๐ช๐ค๐ข๐ญ ๐๐ฆ๐ข๐ด๐ฐ๐ฏ๐ช๐ฏโฆโ52May 4, 2024Updated 2 years ago
- โ25Oct 10, 2025Updated 10 months ago
- skill.md specialized skill registry for AI agentsโ15Apr 24, 2026Updated 4 months ago
- โ149Nov 11, 2024Updated last year
- Repo of paper "Free Process Rewards without Process Labels"โ172Mar 14, 2025Updated last year
- โ127Jun 2, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer โข AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [EMNLP-2025] R1-Zero on ANY TASKโ32Nov 9, 2025Updated 9 months ago
- Plotting (entropy, varentropy) for small LMsโ99May 20, 2025Updated last year
- Code for "Reasoning to Learn from Latent Thoughts"โ134Mar 28, 2025Updated last year
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.โ106Jul 19, 2025Updated last year
- MLGym A New Framework and Benchmark for Advancing AI Research Agentsโ620Aug 10, 2025Updated last year
- Simple RL training for reasoningโ3,874Dec 23, 2025Updated 8 months ago
- Entropy Based Sampling and Parallel CoT Decodingโ3,432Nov 13, 2024Updated last year