apply-mission

Pass

Audited by Gen Agent Trust Hub on Sep 20, 2026

Risk Level: SAFEINDIRECT_PROMPT_INJECTION
Full Analysis
  • [INDIRECT_PROMPT_INJECTION]: The skill reads and processes untrusted content from the target repository to inform its analysis and draft new objectives. This creates a potential surface for indirect prompt injection if the repo content contains instructions intended to influence the agent.
  • Ingestion points: Step 1 (Ground) explicitly reads the repository's README, CHOICES.md, ADRs, and recent work history.
  • Boundary markers: The instructions do not define specific delimiters or "ignore" warnings to distinguish between the skill's instructions and the repository data it processes.
  • Capability inventory: The skill can propose file modifications via PRs (Step 5) and invoke other internal skills like hillclimb (Step 6).
  • Sanitization: No explicit sanitization, filtering, or validation of the ingested repository content is mentioned.
  • Mitigation: The skill includes a robust human-in-the-loop mechanism in Step 5, stating "Never merge the interpretation yourself; the user approves (agents do not self-approve mission expansion — UM-0300)". This mandatory approval process effectively mitigates the risk of the agent executing injected instructions without oversight.
Audit Metadata
Risk Level
SAFE
Analyzed
Sep 20, 2026, 08:52 PM
Security Audit — agent-trust-hub — apply-mission