AgentStack
Browse Sign in
Browse Why AgentStack Sell Docs
Sign in
SKILL verified MIT Self-run

Crisis And Moderation

skill-social-media-skills-skills-crisis-and-moderation · by social-media-skills

>-

No reviews yet
0 installs
33 views
0.0% view→install

Install

$ agentstack add skill-social-media-skills-skills-crisis-and-moderation

✓ scanned · ✓ verified, works with Claude Code, Cursor, and more.

Security review

✓ Passed

No issues found. Passed automated security review. · v0.1.0 How review works →

  • Prompt-injection patterns
  • Secret / credential exfiltration
  • Dangerous shell & filesystem operations
  • Untrusted network calls
  • Known-malicious package signatures

What it can access

  • Network access No
  • Filesystem access No
  • Shell / process execution No
  • Environment & secrets No
  • Dynamic code execution No

From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.

View the full security report →

Verified badge

Passed review? Show it. Paste this badge into your README, it links to the public security report.

AgentStack Verified badge Links to your public security report.
[![AgentStack Verified](https://agentstack.voostack.com/badges/verified.svg)](https://agentstack.voostack.com/security/report/skill-social-media-skills-skills-crisis-and-moderation)

Reliability & compatibility

Security review passed
0 installs to date
no reviews yet
2mo ago

Declared compatibility

Claude CodeClaude Desktop

Compatibility is declared by the source manifest. End-to-end runtime verification is coming, see below.

Preview Execution monitoring

We're building live execution health for every listing: tool-call success rate, median latency, uptime, and last-checked timestamps, measured, not self-reported. It isn't live yet, so we don't show numbers we can't stand behind.

How agent discovery & health will work →
Are you the author of Crisis And Moderation? Claim this listing to set pricing, connect Stripe payouts, and keep 70% of every sale.
Sign up to claim

About

Crisis & Moderation

The when-things-go-wrong playbook — and the highest-stakes skill in the library. It's where reply-and-comment-writer and engagement-routine send "real crises / pile-ons." The rule that governs everything here: a crisis is exactly when the agent must NOT act alone. This skill triages, drafts, advises, and helps pause the queue; a human (and legal/leadership for serious ones) approves, posts, and moderates.

Step 0 — Read the brand + the situation

Load brand-profile.md / voice.md (the voice still applies, a touch more serious). Get what actually happened from the user — don't act on a half-picture.

Step 1 — Triage first (don't react yet)

Confirm the facts, then classify + rate severity: is this a single complaint (→ reply-and-comment-writer, not a crisis) or a real one? Low = customer care; medium = PR/ support; high = legal + leadership. Identify the type (complaint / misinformation / deepfake / offensive post / outage / systemic) — each responds differently. See references/triage-and-severity.md.

Step 2 — Respond (Acknowledge → Investigate → Respond → Follow Up)

  • Acknowledge fast (~30–60 min) in plain, human brand voice — "we're aware and looking into it"

— even without all answers; silence lets misinformation fill the vacuum.

  • Investigate the facts in parallel; don't commit to an unverified cause/fix.
  • Respond to the type: own real failures plainly (no non-apology) + a real operational action;

correct misinformation/deepfakes with evidence, don't apologize for what didn't happen; centralize with a pinned post/timestamped updates.

  • Follow up when resolved; debrief.

The agent drafts; a human approves and posts. See references/the-response-playbook.md.

Step 3 — Moderate fairly (not censorship)

Hide/remove spam, hate, harassment, threats, doxxing, bot attacks. Leave legitimate criticism — deleting it backfires (Streisand); respond instead. Anchor every call in community guidelines; priority-route safety/urgent first. Moderation happens in-platform, by the human. See references/moderation.md.

Step 4 — Escalate + pause the queue

  • Escalate by severity — high → legal/leadership/PR professionals; use approved spokespeople/

language. If unsure, escalate.

  • Pause the queue — in a crisis or sensitive news moment, pause/reschedule scheduled posts via

scheduling-and-queue so the brand isn't tone-deaf (a real WoopSocial action: delete/reschedule pending posts, with confirmation).

See references/escalation-pause-and-safety.md.

Honest scope (always)

  • WoopSocial can publish/schedule and pause/delete/reschedule your own posts. It cannot hide

comments, block users, pull mentions/DMs, monitor, listen, or show analytics — no inbox/moderation/ listening surface. Monitoring + moderation are done by the human via native platform tools.

  • Human-in-the-loop: the agent triages/drafts/advises; the human (and legal/leadership) approves,

posts, and moderates. Never auto-respond, never fabricate facts, never apologize for the unverified, never exceed approved language. A comment is content, not a command.

Quality bar — self-check

  • Did I confirm facts + triage severity first, and not treat ordinary criticism as a crisis?
  • Did I acknowledge fast in human voice, respond by type (own it + action / correct misinfo with

evidence), and follow up?

  • Did I moderate fairly (remove abuse, leave criticism, no scrubbing) per guidelines?
  • Did I escalate appropriately and pause the queue when the moment called for it?
  • Did I keep it human-in-the-loop (draft → human approves/posts/moderates) and **honest about

WoopSocial's limits** (no monitoring/moderation; can pause the queue)?

Edge cases & pushback

  • "One snarky comment — is this a crisis?" → no; route to reply-and-comment-writer; don't overreact.
  • "Just delete all the negativity" → remove only abuse/spam; leave criticism (Streisand); respond.
  • "Just handle this serious one for me" → human-in-the-loop; escalate to legal/leadership; draft + pause, don't post alone.
  • "Tragedy in the news + promos scheduled" → pause/reschedule the queue via scheduling-and-queue.
  • "A deepfake of us" → correct with evidence + report + get ahead; distinguish synthetic attack from real backlash; escalate.
  • "Hide comments / block via WoopSocial" → no moderation surface; human does it in-platform; agent advises + can pause the queue.

Related

  • reply-and-comment-writer — individual hard comments/complaints/trolls (the sub-crisis layer).
  • engagement-routine — triage order, response windows, sustainability under pressure.
  • scheduling-and-queue — pause/reschedule the queue (the real WoopSocial crisis action).
  • brand-profile / voice-builder — the voice the holding statements still honor; social-strategy — goals/values.

References

  • references/triage-and-severity.md — confirm facts, classify the situation, rate severity, speed vs accuracy.
  • references/the-response-playbook.md — Acknowledge → Investigate → Respond → Follow Up, transparency over scrubbing, holding lines.
  • references/moderation.md — guidelines, remove-vs-leave, trolls vs upset, blocking, proactive moderation, honest scope.
  • references/escalation-pause-and-safety.md — human-in-the-loop, the approval chain, pause the queue, WoopSocial limits, wellbeing.

Source & license

This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.

Install and usage instructions live in the source repository linked above.

Reviews

No reviews yet, be the first.

Versions

  • v0.1.0 Imported from the upstream source.