← Files SpeechdubARCHIVED FILE

SKILL.md

1.53 KB · Sep 30, 2026 · 22:55 UTC

↓ Download file

---
name: speechdub-text-to-speech
description: Generate spoken audio with Speechdub preset voices when the user wants text read aloud, a voice preview, or audio output from their content.
---

# Speechdub text-to-speech

Use when the user wants **audio** or **read aloud** via Speechdub voices, not only silent text in chat.

## Tools

- `speechdub_list_voices` — optional `gender` filter (`male` | `female`).
- `speechdub_get_voice` — metadata for one `voice_id` (presets like `F1`, `M3`).
- `speechdub_synthesize_speech` — required `input` (max 5000 chars) and `voice_id`; optional `language`, `audio_format` (`wav` | `mp3`), `include_audio_data`.

## Workflow

1. Always call `speechdub_list_voices` (or `speechdub_get_voice` if they named a preset) **before** synthesis unless a valid `voice_id` is already confirmed.
2. Pick a voice that matches the user’s language or preference; say which voice you chose.
3. Split long text into chunks under 5000 characters if needed; synthesize per chunk.
4. Report billing context briefly: synthesis uses the user’s Speechdub credit wallet (same as the web app). Do not quote API keys.
5. If `include_audio_data` is false or omitted, describe the result (duration, format, request id) per tool output; follow the client’s rules for playing or attaching audio.

## Boundaries

- Do not create or delete library documents unless the user also asked for library changes.
- Prefer synthesis on user-supplied or document text they explicitly want heard; do not synthesize huge documents without narrowing scope.

SHA-256: 242bdbc438df9c670d8cfacfcb49877421aa619d0a68585c609962059ce91ffa