Reference code for the Meta-Harness paper.
☆1,314Jul 11, 2026Updated last week
Alternatives and similar repositories for meta-harness
Users that are interested in meta-harness are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Meta-Harness: 76.4% on Terminal-Bench 2.0 (Claude Opus 4.6)☆1,150Mar 26, 2026Updated 3 months ago
- Meta Harness Implementation☆147Jun 13, 2026Updated last month
- Hierarchal Agent Loop Optimizer☆1,114Updated this week
- Production-grade DSPy 3.2.x agent skills + validated end-to-end examples for Claude Code and Codex CLI — fundamentals, evaluation, GEPA, …☆264Jun 20, 2026Updated last month
- Optimize prompts, code, and more with AI-powered Reflective Optimization☆5,719Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Method for Long Context RLMs using verifiable Lambda Calculus☆304Apr 24, 2026Updated 2 months ago
- The open-source agent-serving project☆481Updated this week
- Production focused Self-harnessed LM runtime (RLM) that allows the LM to call its sub-lm with DSPy signatures. Define your inputs, output…☆412Updated this week
- General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.☆5,296Jun 26, 2026Updated 3 weeks ago
- A plugin for your agentic framework that optimizes code using the GEPA algorithm (Genetic-Pareto LLM-driven search).☆96Apr 28, 2026Updated 2 months ago
- The official repository of "Position: Agentic Evolution is the Path to Evolving LLMs".☆700Jun 29, 2026Updated 3 weeks ago
- autonomous harness engineering☆4,550Apr 3, 2026Updated 3 months ago
- context-efficient terminal agent powered by an RLM☆60Feb 7, 2026Updated 5 months ago
- An implementation of a Meta Harness for Hermes.☆102Jul 11, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A feature rich implementation of Recursive Language Models, with ACP integration, REPL tool support, structured IO, advanced visualizatio…☆450Jul 7, 2026Updated 2 weeks ago
- ⚒ Evolutionary self-improvement for Hermes Agent — optimize skills, prompts, and code using DSPy + GEPA☆4,767Jun 17, 2026Updated last month
- Autonomous experiment loop extension for pi☆7,227Jul 15, 2026Updated last week
- turns your codebase into an autoresearch loop — discovers what to measure, instruments the benchmark, then runs tree search with parallel…☆1,346Updated this week
- ☆111Jun 10, 2026Updated last month
- a recursive self-improving harness designed to help your agents (and future iterations of those agents) succeed on any task☆1,257Updated this week
- Official AHE code — Agentic Harness Engineering: observability-driven automatic evolution of coding-agent harnesses (concurrent w/ meta-h…☆760Jun 14, 2026Updated last month
- Evolve your language agent with Agentic Context Engineering (ACE)☆1,223May 19, 2026Updated 2 months ago
- Bring your own agent and build a self-improving agentic system. Automatically mine failures, optimize the agent harness, and gate against…☆526Jul 8, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 🔥🔥COLM 2026🔥🔥 CORAL is a robust, lightweight infrastructure for multi-agent autonomous self-evolution, built for autoresearch. W…☆834Updated this week
- LLM-as-a-Verifier is a general-purpose framework that provides fine-grained feedback for any agent without requiring additional training.…☆586Jul 7, 2026Updated 2 weeks ago
- Self-referential self-improving agents that can optimize for any computable task☆2,647May 9, 2026Updated 2 months ago
- A benchmark for evaluating AI agents on frontier ultra long-horizon auto research tasks.☆157Jun 17, 2026Updated last month
- Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement…☆10,505Updated this week
- DSPy's Recursive Language Model (RLM) with Daytona Sandbox for secure cloud-based code execution☆49Updated this week
- Symphony turns project work into isolated, autonomous implementation runs, allowing teams to manage work instead of supervising coding ag…☆26,141Updated this week
- A sandboxed Python runtime for AI agents, written in Rust.☆145Apr 21, 2026Updated 3 months ago
- Browser Harness | Self-healing harness that enables LLMs to complete any task.☆16,162Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A self-learning data agent built with systems engineering principles. It grounds answers in 6 layers of context and improves with every q…☆2,110Jul 10, 2026Updated last week
- DSPy: The framework for programming—not prompting—language models☆36,306Updated this week
- AI agents running research on single-GPU nanochat training automatically☆91,712Mar 26, 2026Updated 3 months ago
- A recursive coding agent inpired by RLMs☆375Jun 22, 2026Updated last month
- The World's First Unified Virtual Filesystem For AI Agents☆3,341Updated this week
- Automated harness evolution for AI agents. A Claude Code plugin that iteratively optimizes system prompts, routing, retrieval, and orches…☆40Apr 18, 2026Updated 3 months ago
- ☆39Apr 15, 2026Updated 3 months ago