OpenCoconut implements a latent reasoning paradigm where we generate thoughts before decoding.
โ173Jan 16, 2025Updated last year
Alternatives and similar repositories for OpenCoconut
Users that are interested in OpenCoconut are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- โ27Jan 14, 2025Updated last year
- Implementation of ๐ฅฅ Coconut, Chain of Continuous Thought, in Pytorchโ184Jun 20, 2025Updated last year
- Training Large Language Model to Reason in a Continuous Latent Spaceโ1,683Jul 2, 2026Updated last month
- โ15Apr 26, 2025Updated last year
- โ209Apr 19, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer โข AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- โ16Mar 22, 2025Updated last year
- A curated list of resources on Reinforcement Learning with Verifiable Rewards (RLVR) and the reasoning capability boundary of Large Languโฆโ91Dec 12, 2025Updated 8 months ago
- Recipes to scale inference-time compute of open modelsโ1,132May 26, 2026Updated 2 months ago
- โ157Apr 21, 2026Updated 3 months ago
- This repository includes a benchmark and code for the paper "Evaluating LLMs at Detecting Errors in LLM Responses".โ32Aug 18, 2024Updated last year
- Archon provides a modular framework for combining different inference-time techniques and LMs with just a JSON config file.โ209Mar 7, 2025Updated last year
- Fine-tunes a student LLM using teacher feedback for improved reasoning and answer quality. Implements GRPO with teacher-provided evaluatiโฆโ54May 7, 2025Updated last year
- โ139Mar 20, 2025Updated last year
- MMLU-Pro eval resultsโ15Aug 21, 2025Updated 11 months ago
- AI Agents on DigitalOcean Gradient AI Platform โข AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Scalable RL solution for advanced reasoning of language modelsโ1,869Mar 18, 2025Updated last year
- implementation of https://arxiv.org/pdf/2312.09299โ21Jul 3, 2024Updated 2 years ago
- โ12Jul 8, 2024Updated 2 years ago
- A comprehensive repository of reasoning tasks for LLMs (and beyond)โ499Sep 27, 2024Updated last year
- Training framework with a goal to explore the frontier of sample efficiency of small language modelsโ101Jan 25, 2026Updated 6 months ago
- [ACL 2025 Findings] Official implementation of the paper "Unveiling the Key Factors for Distilling Chain-of-Thought Reasoning".โ23Feb 26, 2025Updated last year
- Entropy Based Sampling and Parallel CoT Decodingโ17Oct 9, 2024Updated last year
- Repo for the IDESSAI 2024 course on modeling audio with discrete tokens.โ13Sep 13, 2024Updated last year
- Synthetic data curation for post-training and structured data extractionโ1,713Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer โข AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Post-trained LLaDA model with dynamic step schedulingโ36Mar 4, 2025Updated last year
- [๐๐๐๐๐ ๐ ๐ข๐ง๐๐ข๐ง๐ ๐ฌ ๐๐๐๐ & ๐๐๐ ๐๐๐๐ ๐๐๐๐๐ ๐๐ซ๐๐ฅ] ๐๐ฏ๐ฉ๐ข๐ฏ๐ค๐ช๐ฏ๐จ ๐๐ข๐ต๐ฉ๐ฆ๐ฎ๐ข๐ต๐ช๐ค๐ข๐ญ ๐๐ฆ๐ข๐ด๐ฐ๐ฏ๐ช๐ฏโฆโ52May 4, 2024Updated 2 years ago
- โ25Oct 10, 2025Updated 10 months ago
- skill.md specialized skill registry for AI agentsโ15Apr 24, 2026Updated 3 months ago
- โ149Nov 11, 2024Updated last year
- Allows two LLMs to communicate and run code in the terminalโ28Dec 8, 2024Updated last year
- Repo of paper "Free Process Rewards without Process Labels"โ172Mar 14, 2025Updated last year
- โ127Jun 2, 2026Updated 2 months ago
- [EMNLP-2025] R1-Zero on ANY TASKโ32Nov 9, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer โข AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Plotting (entropy, varentropy) for small LMsโ99May 20, 2025Updated last year
- Code for "Reasoning to Learn from Latent Thoughts"โ134Mar 28, 2025Updated last year
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.โ106Jul 19, 2025Updated last year
- MLGym A New Framework and Benchmark for Advancing AI Research Agentsโ616Aug 10, 2025Updated last year
- Simple RL training for reasoningโ3,871Dec 23, 2025Updated 7 months ago
- Entropy Based Sampling and Parallel CoT Decodingโ3,431Nov 13, 2024Updated last year
- Our library for RL environments + evalsโ4,496Updated this week