long-horizon-prompting

Pass

Audited by Gen Agent Trust Hub on Jul 13, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill is primarily instructional documentation and does not contain any executable scripts or configuration files that perform system-level operations.
  • [SAFE]: External resource references (URLs) target trusted domains such as openai.com, anthropic.com, and metr.org. These are used neutrally to provide academic and technical context for the described prompting techniques.
  • [SAFE]: The skill implements defensive design patterns against agentic failure modes. For example, it advocates for 'adversarial audit' and 'loophole closure' in prompts to prevent reward-hacking and unreliable agent behavior.
  • [SAFE]: No instances of obfuscation, credential exposure, or persistence mechanisms were detected in the skill body or reference files.
  • [SAFE]: The guidance regarding multi-agent orchestration follows established vendor best practices from OpenAI and Anthropic, emphasizing context isolation and precise delegation.
Audit Metadata
Risk Level
SAFE
Analyzed
Jul 13, 2026, 08:09 AM
Security Audit — agent-trust-hub — long-horizon-prompting