Versioned security assessment

Report ID: SA-B7E95074

7/9/2026, 12:24:20 PM

send-experiment-designer security assessment v3

Skill Security Certification Report

Audit History
Audit model: codex Historical report
Skill name
send-experiment-designer
Version
v3
Maintainer
aaron-he-zhu
Coverage
1 Files scanned · 141 Lines analyzed
Policy version
Unavailable

Confirmed finding summary

No confirmed security findings

The completed audit recorded no confirmed security findings. This is not proof that the Skill has no side effects.

Installation context

Historical evidence

This report may not describe the currently installable artifact. Open the current Skill page for install guidance.

Open current Skill page

This report does not block or authorize the manifest or ZIP.

Most static detections are false positives from Markdown backticks, fenced prompt examples, connector placeholders, and fixed repo-relative documentation links. I confirmed only the optional local experiment.py and python3 statistical helper references as external command surfaces; no prompt injection, network exfiltration, or dynamic path traversal evidence was found.

Report position

Historical report

Open audit history before using this report to install.

Audit attestation

Not attestable

The required immutable binding is incomplete.

Human verification

Not verified

No human verification is recorded for this report.

Coverage

1 Files scanned · 141 Lines analyzed

3 items shown for review

Limitations

This report does not claim runtime or sandbox execution and does not prove the absence of side effects.

Evidence chain

Follow the evidence from source binding to the install contract. Available evidence supports verification; it is not a safety guarantee.

  1. Source

    Binding unavailable

  2. Artifact

    Identity incomplete

  3. Audit

    Complete

  4. Install contract

    Open manifest to verify

    Open manifest

Capabilities observed

Observed means this report recorded supporting evidence. Not recorded does not prove that a capability is absent.

Contains scripts

May execute code included with the Skill.

Not recorded by this audit

Network access

May connect to external services.

Observed in 2 evidence locations

Filesystem access

May read or write local files.

Observed in 12 evidence locations

Env variables

May read values from the process environment.

Not recorded by this audit

External commands

May invoke commands or programs outside the Skill.

Observed in 36 evidence locations

Capability review items (3)
Medium
Ruby/shell backtick execution
> **Significance (keyless — closes the design→measure loop):** once the send results are in, `python
Line 66 instructs use of a python3 command against a plugin-root script for statistical read-outs. It appears intended for local numeric analysis, but it is still an external command surface that should be reviewed.
Medium
Ruby/shell backtick execution
5. **Sample size, MDE, duration, power — from the baseline.** Size each cell for **power 1−β ≥ 0.80
Line 94 tells the agent to use an experiment.py samplesize helper when available. The context is legitimate sample-size math, but it still delegates work to an external executable helper.
Medium
Ruby/shell backtick execution
- Apply **p<0.05 AND ≥ the minimum practical lift set at design time** — statistical significance al
Line 115 prefers experiment.py for significance reads on user ESP exports. The command is used for analysis rather than exfiltration, but invoking an external helper remains a real execution surface.

Risk findings

Confirmed security concerns are separated from items that still need review.

No confirmed security findings were recorded for this completed audit.

Remediation

Suggested fixes recorded by this audit. Applying them is the maintainer’s responsibility.

  1. FIX-001
    Medium
    Optional statistical helper uses external command execution.
    Document the expected script path, require numeric arguments, and provide a no-command formula fallback for hosts that block execution.
  2. FIX-002
    Low
    Parent-directory documentation links trigger path traversal scanners.
    Package referenced materials with the skill or replace parent-directory links with stable skill-local references.
  3. FIX-003
    Low
    The skill can save experiment summaries to memory after consent.
    Keep the explicit consent step and avoid storing raw ESP exports, customer data, or revenue records in memory.

Expert evidence

Immutable subject identity, scanner metadata, dismissed matches, and source-level evidence.

Artifact subject

Marketplace commit
Unavailable
Content hash
Unavailable
Tree hash
Unavailable
Skill path
Unavailable
Audit payload hash
Unavailable

Analysis metadata

Audit model: codex

Analysis state: Complete

Scope is limited to the recorded files, lines, methods, and evidence. No runtime or sandbox execution is claimed.

Verify and export

The manifest and lockfile bind install artifacts to cryptographic hashes. This integrity claim is separate from the security assessment.

Audit attestation: not_attestable