AgentStack
Browse Sign in
Browse Why AgentStack Sell Docs
Sign in
SKILL verified MIT Self-run

Answerability Tester

skill-npbuilds-skill-library-answerability-tester · by npbuilds

>

No reviews yet
0 installs
22 views
0.0% view→install

Install

$ agentstack add skill-npbuilds-skill-library-answerability-tester

✓ scanned · ✓ verified, works with Claude Code, Cursor, and more.

Security review

✓ Passed

No issues found. Passed automated security review. · v0.1.0 How review works →

  • Prompt-injection patterns
  • Secret / credential exfiltration
  • Dangerous shell & filesystem operations
  • Untrusted network calls
  • Known-malicious package signatures

What it can access

  • Network access No
  • Filesystem access No
  • Shell / process execution No
  • Environment & secrets No
  • Dynamic code execution No

From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.

View the full security report →

Verified badge

Passed review? Show it. Paste this badge into your README, it links to the public security report.

AgentStack Verified badge Links to your public security report.
[![AgentStack Verified](https://agentstack.voostack.com/badges/verified.svg)](https://agentstack.voostack.com/security/report/skill-npbuilds-skill-library-answerability-tester)

Reliability & compatibility

Security review passed
0 installs to date
no reviews yet
17d ago

Declared compatibility

Claude CodeClaude Desktop

Compatibility is declared by the source manifest. End-to-end runtime verification is coming, see below.

Preview Execution monitoring

We're building live execution health for every listing: tool-call success rate, median latency, uptime, and last-checked timestamps, measured, not self-reported. It isn't live yet, so we don't show numbers we can't stand behind.

How agent discovery & health will work →
Are you the author of Answerability Tester? Claim this listing to set pricing, connect Stripe payouts, and keep 70% of every sale.
Sign up to claim

About

Answerability Tester — Could Any Evidence Resolve This?

A statement that no evidence could resolve isn't a problem statement; it's a values commitment, a definitional choice, or a circular claim wearing a problem-statement's clothes. Answerability is the precondition for everything downstream — there is no point auditing falsifiability or scope on a statement that fundamentally cannot be answered.

This skill coordinates with evidence-evaluator (philosophy/epistemology) for the evidence-quality assessment. The binding-vow contribution is the framing — translating problem-statement context into the evidence-type question that evidence-evaluator can answer.

The Three Diagnostic Questions

Q1 — What evidence would resolve this if observed?

Name the specific kind of observation, data, study, or demonstration that would settle the question. Be concrete: not "research," but "a randomized trial of N≥500 with X measured at Y intervals."

If you cannot name the evidence type, the statement is provisionally unanswerable. Try Q2 to see whether that's a framing issue or a deeper one.

Q2 — Could that evidence exist in principle?

Sometimes the evidence type is namable but logically impossible to obtain (counterfactual histories, predictions about a singular event before it occurs, claims about subjective experiences of others).

| Outcome | Verdict | |---|---| | Evidence is namable AND obtainable in principle | Provisionally answerable; continue to Q3 | | Evidence is namable but logically impossible | Unanswerable in principle — needs reformulation | | No evidence type satisfies the question even in principle | Values-question masquerading as a fact-question |

Q3 — Is that evidence accessible to the user?

Even when evidence is namable and possible in principle, practical access matters. A statement that requires a 10-year longitudinal study the user cannot fund is technically answerable but functionally not.

| Outcome | Verdict | |---|---| | Evidence is accessible (data exists, can be collected, can be queried) | Answerable | | Evidence exists but requires resources beyond the user's reach | Hard-but-answerable (downgrade scoring; consider reformulation to a tractable proxy) | | Evidence is in principle obtainable but practically not in this user's context | Hard-but-answerable with note |

Coordination with evidence-evaluator

For the Q1 evidence-type assessment, call evidence-evaluator with:

  • Pass: the candidate evidence type (from Q1), the kind of claim being made, the field/domain
  • Receive: judgment on the evidence's epistemic weight if obtained; common confounders; comparable bodies of evidence

evidence-evaluator evaluates strength of evidence; this skill evaluates answerability. They are complementary: a question can be answerable (some evidence resolves it) but the available evidence weak (Q3 returns "Hard-but-answerable").

For the Q2 logical-possibility check, the determination is largely linguistic:

| Linguistic pattern | Likely Q2 verdict | |---|---| | "What if X had happened?" (counterfactual past) | Logically impossible — no evidence of an unrun history | | "Will X happen?" (prediction about singular future event) | Possible only after the fact; provisional answers via base rates and reference classes | | "Does Y truly feel Z?" (other minds, qualia) | Logically constrained — proxies only | | "Is X the right thing to do?" (normative) | Values-question; not a fact-question | | Moral/aesthetic adjectives applied to people without behavioral specifics ("CEOs are crazy", "founders are visionaries", "managers are toxic") | Values-question — the adjective is doing values work, not descriptive work. Push for behavioral specifics; if they're not forthcoming, route to values-excavator | | "Is X effective for purpose Y?" (instrumental claim) | Possible — evidence of effectiveness is the evidence type |

Process

  1. Read the statement.
  2. Q1: Name the evidence type that would resolve. If you can't, jump to Q2 directly.
  3. Q2: Check logical possibility using the linguistic patterns above.
  4. If Q2 returns "logically impossible" or "values-question," halt — return verdict immediately.
  5. Otherwise proceed to Q3: assess accessibility. Optionally call evidence-evaluator for evidence-strength side-information.
  6. Produce output in the structured format below.

Output Format

ANSWERABILITY — [first 60 chars of statement...]
─────────────────────────────────────────────
Verdict: [Answerable | Hard-but-answerable | Unanswerable as posed | Values-question]
Evidence type (Q1): [the kind of evidence that would resolve, or "none nameable"]
Logical possibility (Q2): [possible | impossible | values-shaped]
Practical accessibility (Q3): [accessible | hard | unreachable | N/A]

If Unanswerable as posed:
  Reformulation suggestion: [a different question that IS answerable and gets at the same goal]

If Values-question:
  Underlying values: [what's being commitment-claimed disguised as fact]
  Reformulation: [the value claim made explicit]

If Hard-but-answerable:
  Tractable proxy: [a simpler question that's answerable and approximates the original]

Verdicts and What to Do With Them

| Verdict | What it means | Re-state level recommendation | |---|---|---| | Answerable | Evidence exists and is accessible; statement is well-formed for empirical resolution | None — passes this axis | | Hard-but-answerable | Possible but expensive; downgrade or use a proxy | L1 (rephrase to use the proxy) | | Unanswerable as posed | Logically impossible to resolve as currently framed | L2 (decompose — there's a sub-question that IS answerable) | | Values-question | Not a fact-question; resolution requires ethical/political commitment, not evidence | L3 (depth upgrade — call values-excavator cross-domain to surface the values explicitly) or L4 (user assist) |

Output Contract for six-eyes

When called from Phase 6:

  • Return verdict + evidence type + accessibility assessment
  • Feed the answerability axis on statement-grader (axes 5)
  • If verdict is "Values-question," set re-state level to L3 and recommend values-excavator (philosophy/ethics) as the next call

Failure Modes

| Failure | Response | |---|---| | Q1 returns multiple plausible evidence types | List all of them; the one with the best Q3 access wins | | User insists on the original framing despite Values-question verdict | Surface the values explicitly via values-excavator; don't force-fit empirical framing | | Evidence-evaluator returns inconclusive | Default to "Hard-but-answerable" with note; that's the honest verdict | | Statement is partially answerable (some sub-claims answerable, others not) | Recommend decomposition via claim-decomposer (research) before continuing |

Connections

  • statement-grader (binding-vow) — feeds the answerability axis
  • xy-detector (binding-vow) — XY patterns often flag as "Hard-but-answerable" until Y is surfaced
  • evidence-evaluator (philosophy/epistemology) — primary cross-domain call
  • values-excavator (philosophy/ethics) — escalation path for Values-question verdicts
  • claim-decomposer (research) — for partial-answerability cases

Sources

  • Popper, K. R. (1959). The Logic of Scientific Discovery. (Background on the empirical/non-empirical distinction.)
  • Hume, D. (1739). A Treatise of Human Nature. (The is/ought distinction underlying values-vs-fact verdicts.)
  • See [[kahneman-framing]] for attribute-substitution patterns where unanswerable questions get silently swapped for answerable ones.

Source & license

This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.

Install and usage instructions live in the source repository linked above.

Reviews

No reviews yet, be the first.

Versions

  • v0.1.0 Imported from the upstream source.