run-llm-evals
Installation
SKILL.md
Run LLM evals
Action playbook from twenty-five AI Engineer eval/benchmark talks. Do not summarize talks — pick a workflow and execute it.
Supporting files (read when needed):
- workflows.md — workflows A–N (steps, deliverables, stop conditions)
- source-index.md — src-NNN → talk learnings in ingest-into-skills
Optional: {SKILL_OUTPUT_DIR}/run-llm-evals/
Step 0 — Pick workflow
Use the decision tree below. Open the matching section in workflows.md.