analysis

Pass

Audited by Gen Agent Trust Hub on Sep 15, 2026

Risk Level: SAFE
Full Analysis
  • [SAFE]: The skill's primary purpose is to provide analytical and ethical frameworks for reasoning about complex user-provided scenarios.
  • [SAFE]: No evidence of prompt injection, data exfiltration, or unauthorized command execution was found across the skill's markdown files.
  • [SAFE]: The skill does not download external code or depend on third-party packages, operating entirely through instruction-based reasoning.
  • [INDIRECT_PROMPT_INJECTION]: The skill ingests untrusted data from users (e.g., news stories or conflict details) as part of its analytical workflow. However, the skill lacks capabilities to perform high-risk actions like network communication or system modification, mitigating the risk of exploitation via indirect injection.
  • Ingestion points: Ingests user-provided material in SKILL.md, conflict-mediation.md, and ethics-review.md to produce judgment briefs.
  • Boundary markers: Not explicitly defined in the instructions.
  • Capability inventory: Limited to text generation; no subprocess, network, or file-write capabilities detected.
  • Sanitization: None specified for input data.
Audit Metadata
Risk Level
SAFE
Analyzed
Sep 15, 2026, 12:18 AM
Security Audit — agent-trust-hub — analysis