Reference code for the Meta-Harness paper.
☆1,402Jul 11, 2026Updated last month
Alternatives and similar repositories for meta-harness
Users that are interested in meta-harness are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Meta-Harness: 76.4% on Terminal-Bench 2.0 (Claude Opus 4.6)☆1,170Mar 26, 2026Updated 4 months ago
- Meta Harness Implementation☆153Jun 13, 2026Updated 2 months ago
- Hierarchal Agent Loop Optimizer☆1,139Jul 30, 2026Updated 2 weeks ago
- Production-grade DSPy 3.2.x agent skills + validated end-to-end examples for Claude Code and Codex CLI — fundamentals, evaluation, GEPA, …☆271Jun 20, 2026Updated last month
- Optimize prompts, code, and more with AI-powered Reflective Optimization☆6,093Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Method for Long Context RLMs using verifiable Lambda Calculus☆305Apr 24, 2026Updated 3 months ago
- The open-source agent-serving project☆483Jul 16, 2026Updated 3 weeks ago
- Production focused Self-harnessed LM runtime (RLM) that allows the LM to call its sub-lm with DSPy signatures. Define your inputs, output…☆425Aug 5, 2026Updated last week
- General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.☆5,458Updated this week
- A plugin for your agentic framework that optimizes code using the GEPA algorithm (Genetic-Pareto LLM-driven search).☆99Apr 28, 2026Updated 3 months ago
- The official repository of "Position: Agentic Evolution is the Path to Evolving LLMs".☆724Jun 29, 2026Updated last month
- autonomous harness engineering☆4,563Apr 3, 2026Updated 4 months ago
- context-efficient terminal agent powered by an RLM☆60Feb 7, 2026Updated 6 months ago
- An implementation of a Meta Harness for Hermes.☆106Jul 11, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A feature rich implementation of Recursive Language Models, with ACP integration, REPL tool support, structured IO, advanced visualizatio…☆466Jul 7, 2026Updated last month
- ⚒ Evolutionary self-improvement for Hermes Agent — optimize skills, prompts, and code using DSPy + GEPA☆5,005Jun 17, 2026Updated last month
- Autonomous experiment loop extension for pi☆7,578Jul 15, 2026Updated 3 weeks ago
- turns your codebase into an autoresearch loop — discovers what to measure, instruments the benchmark, then runs tree search with parallel…☆1,367Jul 17, 2026Updated 3 weeks ago
- ☆111Jun 10, 2026Updated 2 months ago
- a recursive self-improving harness designed to help your agents (and future iterations of those agents) succeed on any task☆1,270Updated this week
- Official AHE code — Agentic Harness Engineering: observability-driven automatic evolution of coding-agent harnesses (concurrent w/ meta-h…☆823Aug 3, 2026Updated last week
- Evolve your language agent with Agentic Context Engineering (ACE)☆1,260May 19, 2026Updated 2 months ago
- Bring your own agent and build a self-improving agentic system. Automatically mine failures, optimize the agent harness, and gate against…☆531Jul 8, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Open-source autoresearch powered by autonomous coding agents. Run Claude Code, OpenCode, and Codex with grading, shared knowledge, and mu…☆888Updated this week
- LLM-as-a-Verifier is a general-purpose framework that provides fine-grained feedback for any agent without requiring additional training.…☆635Updated this week
- Self-referential self-improving agents that can optimize for any computable task☆2,677Jul 31, 2026Updated last week
- A benchmark for evaluating AI agents on frontier ultra long-horizon auto research tasks.☆159Jun 17, 2026Updated last month
- Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement…☆10,580Updated this week
- DSPy's Recursive Language Model (RLM) with Daytona Sandbox for secure cloud-based code execution☆51Updated this week
- Symphony turns project work into isolated, autonomous implementation runs, allowing teams to manage work instead of supervising coding ag…☆26,558Updated this week
- A sandboxed Python runtime for AI agents, written in Rust.☆144Apr 21, 2026Updated 3 months ago
- Browser Harness | Self-healing harness that enables LLMs to complete any task.☆16,675Aug 3, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A self-learning data agent built with systems engineering principles. It grounds answers in 6 layers of context and improves with every q…☆2,249Jul 10, 2026Updated last month
- DSPy: The framework for programming—not prompting—language models☆37,168Updated this week
- AI agents running research on single-GPU nanochat training automatically☆93,794Mar 26, 2026Updated 4 months ago
- The World's First Unified Virtual Filesystem For AI Agents☆3,403Updated this week
- A recursive coding agent inpired by RLMs☆381Jun 22, 2026Updated last month
- Automated harness evolution for AI agents. A Claude Code plugin that iteratively optimizes system prompts, routing, retrieval, and orches…☆43Apr 18, 2026Updated 3 months ago
- ☆40Apr 15, 2026Updated 3 months ago