AgentStack
SKILL verified MIT Self-run

Agent Run Evidence Reviewer

skill-45ck-skill-harness-agent-run-evidence-reviewer · by 45ck

Review agent run traces, eval summaries, tool logs, and handoff evidence before accepting a workflow result or self-improvement proposal.

No reviews yet
0 installs
14 views
0.0% view→install

Install

$ agentstack add skill-45ck-skill-harness-agent-run-evidence-reviewer

✓ scanned · ✓ verified — works with Claude Code, Cursor, and more.

Security review

✓ Passed

No issues found. Passed automated security review. · v0.1.0 How review works →

  • Prompt-injection patterns
  • Secret / credential exfiltration
  • Dangerous shell & filesystem operations
  • Untrusted network calls
  • Known-malicious package signatures

What it can access

  • Network access No
  • Filesystem access No
  • Shell / process execution No
  • Environment & secrets No
  • Dynamic code execution No

From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.

Are you the author of Agent Run Evidence Reviewer? Claim this listing to set pricing, connect Stripe payouts, and keep 70% of every sale.
Sign up to claim

About

Agent Run Evidence Reviewer

Use this skill before accepting that an agent run worked, that a workflow improved, or that a learning should be made durable.

Evidence Checks

  • The run has a task or issue id.
  • The agent role and tool boundaries are clear.
  • Changed files, commands, tests, and validators are named.
  • Failures and retries are visible, not edited out.
  • Claims are tied to concrete evidence.
  • Generated artifacts point back to source.
  • Sensitive traces are excluded or redacted.
  • Follow-up work is filed instead of hidden in prose.

Verdicts

  • accept: evidence supports the result
  • needs-gate: result may be correct but validation is missing
  • needs-redaction: evidence is useful but unsafe to store or share
  • needs-follow-up: result is partial and requires tracked work
  • reject: evidence contradicts the result or is too weak

Output

Verdict

Evidence Present

Missing Gates

Redaction Needs

Follow-Up Issues

Durable Learning Decision

Source & license

This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.

  • Author: 45ck
  • Source: 45ck/skill-harness
  • License: MIT
  • Homepage: https://github.com/45ck/skill-harness/releases/tag/v0.1.0

Install and usage instructions live in the source repository linked above.

Reviews

No reviews yet — be the first.

Versions

  • v0.1.0 Imported from the upstream source.