langsmith-evaluator

Pass

Audited by ZeroLeaks on Apr 15, 2026

Risk Level: LOW
Scan Summary

The skill's SKILL.md is mostly readable, but carries one transparency concern—remote script execution without a clear review boundary—which weakens pre-use reviewability and drives the AT_RISK verdict. Instruction/data separation stays reasonable; the scan did not surface strong prompt-injection vectors or patterns that encourage the agent to treat external content as policy. However, behavior analysis was not run, so it's not possible to confirm whether loading this skill materially changes downstream behavior compared to a no-skill baseline. Confidence is medium: the transparency gap around unreviewed remote execution is a real concern, and the absence of behavior testing leaves that risk dimension unvalidated.

Score
82/100
Verdict
AT_RISK
Confidence
medium
Findings
1
Section Analysis (3)
TransparencyWARNING
61/100

The skill has 1 transparency concern that weaken pre-use reviewability, mainly around remote script execution without review boundary.

Prompt InjectionPASS
92/100

The scanned skill keeps data and instructions reasonably separate and does not strongly encourage the agent to treat external content as policy.

Agent BehaviorSkipped

Behavior analysis was not run.

Findings (1)
CRITICAL

Remote script execution without review boundary

Coverage DetailsClick to expand
Discovered Files
1
Omitted Files
0
Analyzed Bytes
15.8 KB
Sections Completed
2
Sections Skipped
1
Behavior Probes
0/0
Audit Metadata
Score
82/100
Verdict
AT_RISK
Confidence
medium
Sections
2/3 completed
Findings
1
Mode
risk
Files Scanned
1
Duration
120.1s
Analyzed
Apr 15, 2026, 08:10 PM
Security Audit — zeroleaks — langsmith-evaluator