omh-agent-evaluation

Installation
SKILL.md

Agent Evaluation

This is a Hermes-native agent-evaluation workflow skill.

Why This Exists

agent-evaluation gives OMH a way to improve executor choice empirically, not by vibes, while preserving executor-neutral product language across Codex, Claude Code, Hermes, and generic runtimes.

Do Not Use When

  • The user needs current runtime readiness only; use executor-runtime-readiness.
  • The user already selected an executor and wants implementation; use the coding handoff or delivery workflow.
  • The user asks for workflow learning from a single failed route; use workflow-learning.
  • The ask is to find and fix runtime, memory, cost, or rendering hotspots rather than score executor or model output quality; use ultraperf.

Examples

Good example:

Installs
15
GitHub Stars
3.0K
First Seen
Aug 27, 2026
omh-agent-evaluation — rlaope/oh-my-hermes