agent-skills-evaluator

Installation
SKILL.md

Description

This skill run agent rate command line to evaluate

The Agtm Skills CLI manages local skill bundles for supported agents (for example claude-code, codex, openclaw). It can download skills from GitHub, install them into the correct agent folders, list what is installed, record run logs, and apply rating benchmarks.

It also serves as a benchmarking tool to evaluate skill outputs:
Benchmark your AI agent against real-world standards — from Google-level engineering to Apple-caliber product launches.
Rate performance of each run with structured scores and levels, helping agents like Claude Code choose the right skills more effectively.

Tutorial

Setup

To use the rate command, have to setup the benchmark levels configuration. save to ./agtm/levels/*.json files

agtm setup --levels
Installs
1
Repository
aiagenta2z/agtm
GitHub Stars
2
First Seen
Mar 20, 2026
agent-skills-evaluator — aiagenta2z/agtm