monitor-experiments
Installation
SKILL.md
Experiment Monitor & Report Generator
Scan active and recently completed experiments, surface what needs attention, and report on the ones that matter.
This is a monitoring skill — keep output accessible to non-experts. Avoid statistical jargon (no p-values, no power analysis). For deep-dive analysis of a specific experiment, use the analyze-experiments skill instead.
CRITICAL: Managing Response Sizes
get_experiments: 3-5 IDs max per call. Filter usingsearchresults BEFORE fetching.query_experimentresponses are large. Extract onlysummaryobjects and validity flags. Ignoretimeseries,xValues, bulk arrays.- Metric name resolution:
searchdoes NOT match metric IDs inqueries. Search withentityTypes: ["METRIC"], emptyqueries,limitPerQuery: 50, scoped to project. Match IDs from results.
Report Structure
The report has two parts: