model-risk-manager

Pass

Audited by Gen Agent Trust Hub on Jun 15, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill's primary purpose is to provide a framework for analyzing AI model risks, failure modes, and mitigation strategies. This is a purely instructional task without executable code.
  • [SAFE]: No network operations, file system access, or credential harvesting patterns were detected in the instructions or metadata.
  • [SAFE]: The instructions do not contain any prompt injection attempts or bypass techniques; rather, they explicitly include 'prompt injection surfaces' as a category of risk to be analyzed and mitigated.
  • [SAFE]: No external dependencies, package installations, or remote script executions are present.
  • [SAFE]: The metadata includes restrictive configurations such as 'sandbox_mode: read-only', which limits the potential impact of the skill.
Audit Metadata
Risk Level
SAFE
Analyzed
Jun 15, 2026, 01:40 AM
Security Audit — agent-trust-hub — model-risk-manager