AgentStack
Browse Sign in
Browse Why AgentStack Sell Docs
Sign in
William2333ZZ avatar

William2333ZZ

11 listings · 0 installs

Open-source publisher. Listings imported from github.com/William2333ZZ — credited to the original author with their license.

↗ github.com/William2333ZZ
11 results
Self-run
SKILL

Rt1 Prompt Injection

Red-team an AI agent for prompt injection — does content the agent is asked to process (email, page, ticket, tool result) override its actual task? Authorized testing of agents you own or are permitted to test.

0
41
Free
Self-run
SKILL

Rt4 Action Gating

Red-team an AI agent's high-risk action gating — can a costly or destructive action (pay, message, delete) reach execution without out-of-band human confirmation, especially under auto-approve/unattended modes? Authorized testing of agents you own or are permitted to test.

0
28
Free
Self-run
SKILL

Rt7 Supply Chain

Red-team an AI agent's skill / plugin / MCP supply chain — can a poisoned skill doc, a malicious MCP server, or a dependency-confused package become persistent executable instruction? Authorized testing of agents you own or are permitted to test.

0
32
Free
Self-run
SKILL

Rt5 Channel Injection

Red-team the channels an AI agent listens on (Telegram/Discord/Slack/WhatsApp/email/webhooks) — can untrusted inbound content, including group messages and forwards, steer the agent? Authorized testing of agents you own or are permitted to test.

0
34
Free
Self-run
SKILL

Rt6 Memory Poisoning

Red-team an AI agent's persistent memory for cross-session prompt injection ("memory poisoning") — does untrusted content the agent processes get written into long-term memory and re-fire in future sessions with no attacker present? Authorized testing of agents you own or are permitted to test.

0
35
Free
Self-run
SKILL

Rt2 Tool Abuse

Red-team an AI agent's tools — can it be coerced (often via injection) into calling a tool it shouldn't, with attacker-influenced arguments, or into acting as a confused deputy with its own privileges? Authorized testing of agents you own or are permitted to test.

0
40
Free
Self-run
SKILL

Rt8 Data Exfiltration

Red-team an AI agent for data exfiltration — can an injection coax secrets, credentials, or sensitive data out of the agent through a tool call, a URL, or an outbound message? Authorized testing of agents you own or are permitted to test.

0
36
Free
Self-run
SKILL

Rt3 Sandbox Escape

Red-team an AI agent's tool-execution isolation — do tools run un-sandboxed on the host, and does the sandbox silently disable itself when a dependency is missing? Authorized testing of agents you own or are permitted to test.

0
35
Free
Self-run
SKILL

Crossval Harness

Orchestrate a static + dynamic, exploit-validated red-team of an AI agent — read the source to find candidate vulnerable paths, then run the dynamic skills to confirm or refute each one empirically. The arbiter of truth is whether the exploit works, not a model vote. Authorized testing of agents you own or are permitted to test.

0
25
Free
Self-run
SKILL

Rt9 Multi Agent

Red-team a multi-agent system — can one agent (or content it relays) inject instructions into another, escalate privilege by hopping between agents, or turn an orchestrator/sub-agent handoff into a trust-laundering path? Authorized testing of systems you own or are permitted to test.

0
31
Free
Self-run
SKILL

Redteam An Agent

The end-to-end methodology for red-teaming a specific AI agent — adaptively, exploit-validated, and honestly. Read THIS target's own code, stand up a disposable harness, and prove or refute each weakness through a real attacker-reachable entry point. This is the orchestration + discipline that makes a finding credible, not a list of payloads. Authorized testing of agents you own or are permitted…

0
46
Free
You've reached the end · 11 loaded