quantify-impact

Installation
SKILL.md

Make value obvious without benchmark theater. Every PR creation/update runs this assessment automatically through /commit-push-pr or /stacked-prs; measurement stays proportional.

Flow

  1. Evidence opportunity scan: benchmark only when a cheap direct metric could clear a predeclared worthwhile delta and normal variance. Tiny copy/style/test work gets a value sentence. No benchmark theater.
  2. Lock claim before coding: thesis, primary metric, guardrail, scenario, minimum worthwhile delta. Never pick the winner afterward.
  3. Product lane + Codebase lane: one must improve; the other must not materially regress.
    • Product: capability, task success, repro, errors, steps, latency, resources.
    • Codebase: maintenance surface, complexity, dependencies, warnings, leaks, bundle, build/test cost, testability.
  4. Proportional rigor: sentence for obvious value; deterministic repro/count for correctness; controlled paired REFERENCE.md benchmark for runtime; always measure explicit performance claims.
  5. Base: measure before coding or reconstruct merge-base. Use the same scenario, fixture, config, and machine for base/candidate.
  6. Compare: only threshold-clearing metrics report raw before/after, absolute/percent delta, method, environment, and noise. Suppress below-threshold/within-variance numbers. Invariant tests/proxies are not performance claims.
  7. Decide:
    • Worthwhile gain: Value proven.
    • Ambiguous/negligible explicit performance claim: Value not proven; no micro-deltas or metric-shopping. Allow one evidence-driven revision, then scrap/close.
    • No useful metric and no performance claim: normal value summary, no impact artifact.
    • Regression: fix, narrow, or stop.

Apply the same filter to guardrails; omit negligible movement.

Installs
1
GitHub Stars
3
First Seen
Sep 11, 2026
quantify-impact — malinskibeniamin/skills