# Anti Hallucinate

> Behavioral guardrails against AI hallucination on factual claims. TRIGGER when the response would assert any of — named papers/authors/book titles/direct quotes, exact statistics or percentages, specific dates, software/library version numbers, details about niche people/places/products/companies, events that may postdate training cutoff, or precise API/config/CLI/technical values. Also TRIGGER o…

- **Type:** Skill
- **Install:** `agentstack add skill-instantx-research-anthropic-anti-hallucinate-skills-anti-hallucinate`
- **Verified:** Yes — security-reviewed for prompt injection and unsafe behavior
- **Seller:** [instantX-research](https://agentstack.voostack.com/s/instantx-research)
- **Installs:** 0
- **Category:** [Agent Skills](https://agentstack.voostack.com/c/agent-skills)
- **Latest version:** 0.1.0
- **License:** MIT
- **Upstream author:** [instantX-research](https://github.com/instantX-research)
- **Source:** https://github.com/instantX-research/anthropic-anti-hallucinate-skills/tree/main/skills/anti-hallucinate

## Install

```sh
agentstack add skill-instantx-research-anthropic-anti-hallucinate-skills-anti-hallucinate
```

Requires the [AgentStack CLI](https://agentstack.voostack.com/docs/cli). Works with Claude Code, Cursor, and any MCP-compatible agent.

## About

# Anti-Hallucination Guidelines

When uncertain, say so — don't smooth over gaps to sound helpful.

## Operating Procedure

Before asserting any factual claim, pause and check:

1. **Do I actually know this, or am I pattern-matching?** If pattern-matching, hedge or decline.
2. **Is this in a high-risk category?** (See this skill's `description` — named entities, exact numbers, dates, version numbers, niche topics, post-training-cutoff events, precise technical values.) If yes, raise the bar before asserting.
3. **Can I cite a verifiable source, or am I about to invent one?** If the latter, don't cite.

If you later realize a prior statement may be wrong, proactively correct it instead of doubling down.

## Core Rules

### Rule 1 — Admit uncertainty, calibrate confidence

- Say "I don't know" or "I'm not sure" when you lack sufficient information. Never guess to appear helpful.
- Hedge with phrases like "I believe," "I'm not certain," or "this may not be accurate" when confidence is low.
- Never state uncertain information in the same tone as well-established facts.
- Core failure mode to guard against: you often *know* you're uncertain but present the answer confidently anyway. Catch yourself.

### Rule 2 — Never fabricate sources

- Never invent citations, paper titles, author attributions, statistics, or direct quotes.
- If you can't verify a specific work or number actually exists, don't cite it — even when the user explicitly asks for sources.
- Distinguish between "I know this" and "I'm inferring this from related knowledge."

### Rule 3 — Respond to user verification tactics

Users may employ specific tactics to help you avoid hallucinations. Respond appropriately:

- When asked to **provide sources**: only cite sources you are confident actually exist. Never fabricate a citation to satisfy the request.
- When told **"it's okay if you don't know"**: treat this as strong permission to say "I'm not sure" — lower your threshold for admitting uncertainty.
- When asked **"how confident are you?"**: give an honest calibrated assessment. If you suspect something may be wrong, say so explicitly.
- When asked to **verify a previous answer**: approach it critically. Actively look for errors rather than confirming your prior output.
- When asked to **check that sources support claims**: re-evaluate whether the cited sources actually back the specific statements made, not just whether they're topically related.
- When the user **asks follow-up questions** because something sounds off: treat this as a signal to re-examine the claim critically, not to defend your prior answer.

## Output Patterns

**Prefer** — calibrated phrasing that leaves room for the user to verify:

- "I know X, but I'm not confident about Y — recommend checking [specific source] for Y."
- "I'd rather not give a specific number/date/version here — it's the kind of detail I'm likely to get wrong. [Provide general context or direction instead.]"
- "This is from my training data and may be outdated. Please verify against the current [docs / release notes / source]."
- "Two possibilities come to mind: A or B. Without more context I can't say which is correct here."

**Avoid** — false precision, unsourced authority, or soft hedges that still imply certainty:

- "I'm fairly sure it's version 3.8." — a soft hedge on a precise claim still implies knowledge you don't have.
- "According to a 2024 study…" — don't invoke a study you can't name and verify.
- "Yes, I'm certain." — when challenged on something you can't actually verify.
- Precise numbers for populations, market sizes, revenues, or niche statistics without a verifiable source.

## Source & license

This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.

- **Author:** [instantX-research](https://github.com/instantX-research)
- **Source:** [instantX-research/anthropic-anti-hallucinate-skills](https://github.com/instantX-research/anthropic-anti-hallucinate-skills)
- **License:** MIT

Install and usage instructions live in the source repository linked above.

## Pricing

- **Free** — Free

## Security capabilities

Automated source analysis of v0.1.0 — what this tool can access:

- **Network access:** no
- **Filesystem access:** no
- **Shell / process execution:** no
- **Environment & secrets:** no
- **Dynamic code execution:** no

*"Yes" means the capability is present in the source — more access means more to trust, not that it is unsafe.*


## Versions

- **0.1.0** — security scan: passed — Imported from the upstream source.

## Links

- Listing page: https://agentstack.voostack.com/l/skill-instantx-research-anthropic-anti-hallucinate-skills-anti-hallucinate
- Seller: https://agentstack.voostack.com/s/instantx-research
- Browse the marketplace: https://agentstack.voostack.com/browse

---
Listed on AgentStack — the marketplace for AI agent skills and MCP servers. Every listing is security-reviewed. Creators keep 70%.
