agent-skills-evaluator
Installation
SKILL.md
Description
This skill run agent rate command line to evaluate
The Agtm Skills CLI manages local skill bundles for supported agents (for example claude-code, codex, openclaw). It can download skills from GitHub, install them into the correct agent folders, list what is installed, record run logs, and apply rating benchmarks.
It also serves as a benchmarking tool to evaluate skill outputs:
Benchmark your AI agent against real-world standards — from Google-level engineering to Apple-caliber product launches.
Rate performance of each run with structured scores and levels, helping agents like Claude Code choose the right skills more effectively.
Tutorial
Setup
To use the rate command, have to setup the benchmark levels configuration. save to ./agtm/levels/*.json files
agtm setup --levels