← Files PikaARCHIVED FILE

skills/pika-speech/SKILL.md

2.12 KB · Oct 10, 2026 · 06:12 UTC

↓ Download file

---
name: pika-speech
description: Use when the user wants text spoken aloud — a voiceover, narration or text-to-speech clip — or asks which Pika voices exist.
---

# Speech and voiceover

1. Use the `create_speech` tool. Pick `model` (default ElevenLabs Multilingual v2); each takes different fields:

   | `model` | Script field (max chars) | Voice field |
   | --- | --- | --- |
   | `pika/pika-audio/pika-speech` | `script` (15000) | `voice_preset`, default `calm_documentary_narrator` |
   | `elevenlabs/eleven-multilingual-v2/text-to-speech` | `text` (10000) | `voice_id`, default Rachel; `stability` 0–1, default 0.5 |
   | `elevenlabs/eleven-turbo-v2-5/text-to-speech` | `text` (40000) | same as Multilingual |
   | `minimax/minimax-speech-02-hd/text-to-speech` | `text` (50000) | `voice_id`: any MiniMax voice id, free text, default `Wise_Woman` |
   | `minimax/minimax-speech-02-turbo/text-to-speech` | `text` (50000) | same as HD |
   | `bytedance/seed-audio-1.0/text-to-audio` | `prompt` (3000) | `voice`, optional; omit and the model picks |

2. To show the voices, call `list_presets` for `create_speech` with the model's collection and pass the pick in its field: `pika-voices` → `voice_preset` (76 stock voices, e.g. `deep_trailer_bass`, `polished_news_anchor`, `cozy_grandpa_storyteller`), `elevenlabs-voices` → `voice_id`, `seed-voices` → `voice`. MiniMax has no list; any MiniMax voice id works.

   ElevenLabs ids: `21m00Tcm4TlvDq8ikWAM` Rachel · `pNInz6obpgDQGcFmaJgB` Adam · `EXAVITQu4vr4xnSDxMaL` Sarah · `ErXwobaYiN019PkySvjV` Antoni · `AZnzlk1XvdvUeBnXmlld` Domi · `IKne3meq5aSn9XLyUdCD` Charlie · `JBFqnCBsd6RMkjVDRZzb` George · `TX3LPaxmHKxFdv7VOQHJ` Liam · `pFZP5JQG7iQjIQuC4Bku` Lily · `nPczCjzI2devNBz1zQrb` Brian · `cgSgspJ2msm6clMCkdW9` Jessica

3. Call `create_speech` with a fresh `request_id` and no `max_credits`, then follow the quote's `next` (it applies the user's approval mode). On a yes, call it again with the same input and `request_id` and the agreed `max_credits`.
4. Offer follow-ups: `talking_head_studio` to lip-sync a face to it (see `pika-talking-head`), or `create_audio` for music under it.

SHA-256: 61ba1e2dd6c41321f816c97e3382dd4c65869e968a198a82b9c5335c003f354e