compute-validation

Pass

Audited by Gen Agent Trust Hub on Jun 28, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill provides a comprehensive validation methodology (Verification, Orchestration Safety, Smoke Loop, and Production) designed to improve scientific reproducibility and reduce wasted compute resources.
  • [COMMAND_EXECUTION]: The skill uses Bash and Python to extract metrics from simulation logs (e.g., energy drift, box dimensions, and throughput). These operations are standard for computational research and are limited to specific data analysis tasks.
  • [PROMPT_INJECTION]: The skill ingests data from external simulation logs and scheduler outputs, creating a theoretical indirect prompt injection surface. However, the risk is negligible as the instructions guide the agent to perform deterministic data extraction (e.g., numeric grep and awk) rather than executing logic based on free-form text.
  • [EXTERNAL_DOWNLOADS]: The skill utilizes WebSearch and WebFetch for 'External Research' to investigate unfamiliar physical regimes or tool-specific documentation. All referenced domains (e.g., UIUC for NAMD and SchedMD for SLURM) are reputable sources for the respective scientific tools.
Audit Metadata
Risk Level
SAFE
Analyzed
Jun 28, 2026, 09:24 PM
Security Audit — agent-trust-hub — compute-validation