benchmark-praxis

Installation
SKILL.md

Benchmark Praxis

Structured benchmarking of Praxis methodology against external frameworks/tools. 5 modes: full, quick, gap, track, inventory.

Benchmarking: $ARGUMENTS

Mode Detection

Argument starts with Mode Set Depth
quick ... Quick Set 1 (4-5 dims) Lightweight, no deep research
gap ... or gap (no arg) Gap Set 2 (4 dims) Self-assessment, inward-looking
track ... Track Feature delta Changelog diff vs features/{subject}.yaml, triage new entries
inventory ... Inventory Full feature sweep Crawl docs tree, append every missing feature as new for batch triage
anything else Full Set 1 (all 8 dims) Deep research + verdict + actions

When to use inventory vs track: Inventory is a one-shot baseline build (or rare refresh). Track is recurring delta maintenance against an existing baseline. Run inventory first for any new subject; track keeps it fresh afterward.

AskUserQuestion Guard

Installs
1
GitHub Stars
20
First Seen
Jul 17, 2026
benchmark-praxis — digital-stoic-org/agent-skills