AgentStack
SKILL verified MIT Self-run

Build Multimodal Products

skill-hiteshbandhu-skills-i-use-build-multimodal-products · by hiteshbandhu

>

No reviews yet
0 installs
2 views
0.0% view→install

Install

$ agentstack add skill-hiteshbandhu-skills-i-use-build-multimodal-products

✓ scanned · ✓ verified — works with Claude Code, Cursor, and more.

Security review

✓ Passed

No issues found. Passed automated security review. · v0.1.0 How review works →

  • Prompt-injection patterns
  • Secret / credential exfiltration
  • Dangerous shell & filesystem operations
  • Untrusted network calls
  • Known-malicious package signatures

What it can access

  • Network access No
  • Filesystem access No
  • Shell / process execution No
  • Environment & secrets No
  • Dynamic code execution No

From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.

Are you the author of Build Multimodal Products? Claim this listing to set pricing, connect Stripe payouts, and keep 70% of every sale.
Sign up to claim

About

Build multimodal products

Action playbook from six AI Engineer / World's Fair talks. Do not summarize talks — pick a workflow and execute it.

Supporting files (read when needed):

  • [workflows.md](workflows.md) — workflows A–F (steps, deliverables, stop conditions)
  • [source-index.md](source-index.md) — src-NNN → talk learnings in ingest-into-skills

Optional: {SKILL_OUTPUT_DIR}/build-multimodal-products/


Step 0 — Pick workflow

Use the decision tree below. Open the matching section in [workflows.md](workflows.md).

What is the user trying to do?
├─ Ship realtime voice agent                           → A
├─ Deploy edge/tiny vision (Moondream-class)           → B
├─ Long multimodal streams (SSM)                       → C
├─ Design unbounded multimodal UX                      → D
├─ Build training/eval datasets                          → E
└─ EdTech multimodal literacy                            → F

Stop summarizing once a workflow is identified — run its checklist.


Install

cp -r skills/build-multimodal-products ~/.claude/skills/
cp -r skills/build-multimodal-products ~/.cursor/skills/
cp -r skills/build-multimodal-products ~/.codex/skills/

From skills-i-use or ingest-into-skills (playlists/multimodality-aie-world-s-fair-2024/).


Cross-cutting rules

| Rule | Source | |------|--------| | Measure time-to-first-audio | [src-001 @ 9:28] | | Show model limits in UX | [src-004 @ 14:11] | | Dataset legal/diversity before scale | [src-005 @ 1:38] |

Disputed steps: see [source-index.md](source-index.md). Name workflow A–F; save artifacts to ./skill-outputs/build-multimodal-products/ when requested; do not auto-commit.


Invocation examples

@build-multimodal-products design voice agent under 500ms
multimodal UX without form fields

Source & license

This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.

Install and usage instructions live in the source repository linked above.

Reviews

No reviews yet — be the first.

Versions

  • v0.1.0 Imported from the upstream source.