agenthog-add-eval
Installation
SKILL.md
Agenthog add eval — PLANNED
This skill is planned but not yet written. The intended scope:
- Pick an evaluator template (hallucination, helpfulness, format-correctness, custom LLM-as-judge).
- Wire the eval to run on a sample of incoming traces (e.g. 10%) or on a dataset on demand.
- Surface the eval score as a
agent.evalevent on the trace. - Verify the score shows up in the AgentHog Evaluators tab.
Track progress at https://github.com/TheAgentOS/agentos-skills/issues