mllm-eval
Pass
Audited by Gen Agent Trust Hub on Jul 14, 2026
Risk Level: SAFE
Full Analysis
- [SAFE]: The skill operates as an advisory tool for clinical evaluation design. Technical review of
scripts/check_mllm_eval_completeness.pyshows it is a task-aware linter using standard library regex for presence checking, with no external dependencies or remote execution paths.\n- [SAFE]: No network exfiltration, remote code execution, or credential management issues were detected. Auditing logic is performed locally on user-provided protocol drafts without dangerous interpolation of untrusted content.
Audit Metadata