AgentStack
SKILL verified MIT Self-run

Speak Response

skill-tdimino-claude-code-minoan-speak-response · by tdimino

Vocalize Claude's last response using local Qwen3-TTS. Default voice is the Oracle (deep, resonant Dune narrator). Use --preset for emotion-controlled preset speakers.

No reviews yet
0 installs
14 views
0.0% view→install

Install

$ agentstack add skill-tdimino-claude-code-minoan-speak-response

✓ scanned · ✓ verified — works with Claude Code, Cursor, and more.

Security review

✓ Passed

No issues found. Passed automated security review. · v0.1.0 How review works →

  • Prompt-injection patterns
  • Secret / credential exfiltration
  • Dangerous shell & filesystem operations
  • Untrusted network calls
  • Known-malicious package signatures

What it can access

  • Network access Used
  • Filesystem access No
  • Shell / process execution No
  • Environment & secrets No
  • Dynamic code execution No

From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.

Are you the author of Speak Response? Claim this listing to set pricing, connect Stripe payouts, and keep 70% of every sale.
Sign up to claim

About

Speak Response

Vocalize text using local Qwen3-TTS. Default voice is the Oracle (cloned from a Dune narrator with deep, resonant, prophetic quality).

Quick Examples

| Command | Effect | |---------|--------| | /speak | Last 2 sentences with Oracle voice | | /speak 5 | Last 5 sentences with Oracle voice | | /speak "The sleeper must awaken." | Specific text with Oracle voice | | /speak --preset mood:warm | Last 2 sentences with preset speaker + emotion | | /speak --preset "Hello" speaker:Vivian voice:"nurturing" | Preset speaker with custom voice |

Default: Oracle Voice

The oracle voice is a deep, resonant, prophetic voice cloned from a Dune narrator. It speaks all text with a sense of ancient wisdom and gravitas.

# Default usage - Oracle voice
scripts/speak.sh "The spice must flow."
scripts/speak.sh "He who controls the spice controls the universe."

Limitation

The Oracle uses voice cloning (Base model), which does not support per-message instruction control. The voice characteristics are fixed. For emotion/mood control, use --preset.

Preset Speakers (--preset)

For emotion and mood control, use --preset to switch to CustomVoice with adjustable instructions:

scripts/speak.sh --preset "" [speaker] [instruction]

Quick Preset Examples

# Calm therapeutic voice
scripts/speak.sh --preset "Take a deep breath." Vivian "calm, nurturing, gentle pace"

# Excited announcement
scripts/speak.sh --preset "We did it!" Ryan "joyful, excited, enthusiastic"

# Serious explanation
scripts/speak.sh --preset "This is important." Eric "serious, measured, emphatic"

Custom Voice Instructions

The model understands rich natural language descriptions:

| Aspect | Examples | |--------|----------| | Emotion | joyful, melancholic, anxious, calm, excited, contemplative | | Pace | slow and deliberate, rapid and energetic, measured, hesitant | | Intensity | soft and gentle, loud and commanding, whispered, emphatic | | Style | warm and nurturing, professional, playful, dramatic | | Prosody | with dramatic pauses, rising intonation, emphatic on key words |

Mood Presets (Shortcuts)

| Preset | Expands To | |--------|------------| | calm | "calm, soothing, gentle pace" | | warm | "warm, empathetic, nurturing tone" | | excited | "joyful, excited, enthusiastic" | | serious | "serious, measured, authoritative" | | gentle | "soft, gentle, whispered" | | encouraging | "encouraging, uplifting, sincere" | | contemplative | "thoughtful, slow pace, reflective" |

Speakers

| Speaker | Best For | |---------|----------| | Ryan (default) | Professional, serious, authoritative | | Vivian | Warm, nurturing, therapeutic | | Serena | Calm, gentle, contemplative | | Dylan | Friendly, casual, playful | | Eric | Serious, dramatic, commanding | | Aiden | Encouraging, uplifting, energetic | | UncleFu | Wise, measured | | OnoAnna | Soft, gentle | | Sohee | Clear, professional |

Workflow

  1. Parse arguments for text and mode (default oracle vs --preset)
  2. Extract text from last response if not provided
  3. Default mode: Clone with Oracle voice
  4. Preset mode: Generate with CustomVoice + instruction
  5. Audio plays through macOS speakers

Execution

# Oracle voice (default)
scripts/speak.sh ""

# Preset speaker with instruction
scripts/speak.sh --preset "" [speaker] [instruction]

Voice Cloning (Custom Voices)

Clone any voice from a 3+ second audio sample:

# Get transcript first (use Whisper API)
curl -s https://api.openai.com/v1/audio/transcriptions \
  -H "Authorization: Bearer $OPENAI_API_KEY" \
  -F file="@reference.mp3" -F model="whisper-1"

# Clone the voice
scripts/clone.sh "" "" ""

Voice Design (Create New Voices)

Design entirely new voices from natural language descriptions:

scripts/design-voice.sh "" ""

# Example: Create a warm guide voice
scripts/design-voice.sh \
  "Take a deep breath and feel this moment." \
  "warm, nurturing, gentle pace, empathetic, female"

Then clone the designed voice for reuse:

scripts/clone.sh "New text" designed-voice.wav "Original sample text"

See references/moods.md for more instruction examples.

Source & license

This open-source skill is cataloged on AgentStack and links to its original source — we do not rehost the code.

Install and usage instructions live in the source repository linked above.

Reviews

No reviews yet — be the first.

Versions

  • v0.1.0 Imported from the upstream source.