agent-harness

Installation
SKILL.md

Agent harness — what the agent is told, and how to audit what someone else told theirs

agent-orchestrator wires the loop. agent-evals proves it behaves. agent-interop gets it talking to other processes. This skill is the layer between them and the model: the prompt, the tools, and the choice of technique. The outside term for this ground is harness engineering — OpenAI's article of that name (openai.com/index/harness-engineering, read 2026-08-30) and Anthropic's harness-design guidance (anthropic.com/engineering/harness-design-long-running-apps, read 2026-08-30) both name this same layer, and its leverage is measured: on ARC-AGI-3, harness-level changes alone moved a fixed model from 13.3% to 38.3% while spending a sixth of the tokens (as reported 2026-08-30).

It runs in both directions. Building one and auditing one are the same checklist read forwards and backwards, which is why they live together here.


Rule zero — check the prompt first (a diagnostic heuristic, with exceptions)

Installs
109
GitHub Stars
4
First Seen
Aug 14, 2026
agent-harness — ssheleg/agent-stack