Archon provides a modular framework for combining different inference-time techniques and LMs with just a JSON config file.
☆209Mar 7, 2025Updated last year
Alternatives and similar repositories for Archon
Users that are interested in Archon are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆123Sep 25, 2024Updated last year
- Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments (Zhou et al., EMNLP 2024)☆14Oct 3, 2024Updated last year
- [NeurIPS 2024] Fast Best-of-N Decoding via Speculative Rejection☆56Oct 29, 2024Updated last year
- ☆59Jan 28, 2025Updated last year
- ☆47Feb 8, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆15Apr 26, 2025Updated last year
- ☆28Sep 4, 2025Updated 11 months ago
- Tinker ↔ KernelBench Integration enabling RL for GPU Kernel Generation☆29Mar 5, 2026Updated 5 months ago
- Self-Supervised Alignment with Mutual Information☆20May 24, 2024Updated 2 years ago
- ☆127Jun 2, 2026Updated 2 months ago
- Evals meant to evaluate language models' ability to reason over long contexts.☆10Sep 12, 2024Updated last year
- OpenCoconut implements a latent reasoning paradigm where we generate thoughts before decoding.☆173Jan 16, 2025Updated last year
- Code release for the paper "Style Vectors for Steering Generative Large Language Models", accepted to the Findings of the EACL 2024.☆37Sep 26, 2024Updated last year
- Repository for the paper Stream of Search: Learning to Search in Language☆154Feb 3, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The original Shared Recurrent Memory Transformer implementation☆36Jul 11, 2025Updated last year
- 🌾 OAT: A research-friendly framework for LLM online alignment, including reinforcement learning, preference learning, etc.☆669Jan 29, 2026Updated 6 months ago
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.☆433Updated this week
- ☆40Jun 19, 2024Updated 2 years ago
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,144Updated this week
- Experimental Code for StructuredRAG: JSON Response Formatting with Large Language Models☆121Apr 27, 2026Updated 3 months ago
- Small, simple agent task environments for training and evaluation☆20Nov 1, 2024Updated last year
- Source codes for "Preference-grounded Token-level Guidance for Language Model Fine-tuning" (NeurIPS 2023).☆17Jan 8, 2025Updated last year
- AllenAI's post-training codebase☆3,827Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [𝐄𝐌𝐍𝐋𝐏 𝐅𝐢𝐧𝐝𝐢𝐧𝐠𝐬 𝟐𝟎𝟐𝟒 & 𝐀𝐂𝐋 𝟐𝟎𝟐𝟒 𝐍𝐋𝐑𝐒𝐄 𝐎𝐫𝐚𝐥] 𝘌𝘯𝘩𝘢𝘯𝘤𝘪𝘯𝘨 𝘔𝘢𝘵𝘩𝘦𝘮𝘢𝘵𝘪𝘤𝘢𝘭 𝘙𝘦𝘢𝘴𝘰𝘯𝘪𝘯…☆52May 4, 2024Updated 2 years ago
- [NeurIPS'25] Beyond Accuracy: Dissecting Mathematical Reasoning for LLMs Under Reinforcement Learning☆16Dec 12, 2025Updated 7 months ago
- Optimizing inference proxy for LLMs☆4,236Jul 18, 2026Updated 3 weeks ago
- ☆27Sep 11, 2024Updated last year
- GRadient-INformed MoE☆265Sep 25, 2024Updated last year
- Reproducing R1 for Code with Reliable Rewards☆13Apr 9, 2025Updated last year
- ☆17Dec 16, 2025Updated 7 months ago
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.☆106Jul 19, 2025Updated last year
- Code for "Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate" [COLM 2025]☆182Jul 8, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Your Python AI Coder!☆36May 21, 2025Updated last year
- Evaluating LLMs with fewer examples☆182Jul 4, 2026Updated last month
- ☆1,033Dec 17, 2024Updated last year
- Survival of the Most Influential Prompts: Efficient Black-Box Prompt Search via Clustering and Pruning (Zhou et al.; EMNLP 2023 Findings)☆17Feb 17, 2024Updated 2 years ago
- Code for the paper "VinePPO: Unlocking RL Potential For LLM Reasoning Through Refined Credit Assignment"☆192May 25, 2025Updated last year
- AWM: Agent Workflow Memory☆456Dec 22, 2025Updated 7 months ago
- KernelBench: Can LLMs Write GPU Kernels? - Benchmark + Toolkit with Torch -> CUDA (+ more DSLs)☆1,194Mar 24, 2026Updated 4 months ago