← PixVerseCONTENT HISTORYWHAT CHANGED · RULE-BASED ANALYSIS
Update to PixVerse
Snapshot Sep 30, 2026 · 23:13 UTC · version 1.3.2
Collection source: not recorded for this historical snapshot.
First saved snapshot
No earlier snapshot is available to establish a change.
Compare saved observations
Download comparison JSONFull technical diff · 0 changed fields
Full snapshot data
{
"description": "Music and sound: create music or sound assets, or route audio work to narration and mixing. Use voiceover for a spoken delivery and captions for subtitles.",
"included_files": [
{
"relative_path": "agents/openai.yaml",
"size_in_bytes": 231
}
],
"name": "pixverse-audio",
"skill_md_contents": "---\nname: pixverse-audio\ndescription: \"Music and sound: create music or sound assets, or route audio work to narration and mixing. Use voiceover for a spoken delivery and captions for subtitles.\"\n---\n\n# Voice and Music\n\n## Installed Command\n\nBefore running helper commands, resolve the absolute plugin root from this `./SKILL.md` path: it is two directories above this skill directory and contains `.codex-plugin/plugin.json`. In every new shell session, set:\n\n```bash\nPVX=\"<absolute-plugin-root>/scripts/pvx\"\n```\n\nRun helper commands through `\"${PVX}\"`. Do not assume the user's current working directory is the plugin checkout, and do not bypass the installed wrapper with a module invocation.\n\n## CLI Execution\n\nApply `../../skills-shared/quality-policy.md` to all generated images/videos: Sunburst 2K/high,\nSeedance 2.5 1080p with automatic prompt enhancement, and an explicit upgrade/fallback\nchoice for Free/Basic or model-entitlement rejection.\n\nFor an explicitly selected Canvas project or node, read\n`../../skills-internal/pixverse-agent-canvas/SKILL.md` first and follow its Canvas\npreflight and delivery contract; the ordinary queue/local route below does not replace it.\n\nFor ordinary creation or editing, read `../../skills-shared/cli-workflow.md` before\nexecuting this workflow. It contains the existing account-aware routing, paid-work\nconfirmation, progress, preview and project-handoff rules. Studio and Production are\nnot prerequisites. The shared account, confirmation and direct-medium rules govern the\nexamples and defaults below. Read the gateway only when writing manual CLI commands or queue specs.\nCreation command examples below are queue-task fragments, not permission to submit paid\n`pixverse create` commands directly.\n\nRead `../../skills-shared/audio-craft.md`.\n\nFor standalone SFX/reference-audio capabilities, read\n`../../skills-shared/extended-capabilities.md`. Native UGC voices follow their\nspecialized workflow; do not simplify its audio map into a sparse generic prompt.\n\n## Voice\n\nUse standalone voice when narration needs timing, reuse, subtitles, or download:\n\n```bash\npixverse voice presets --model speech-2.8-hd --language zh --json\npixverse create voice --model speech-2.8-hd --voice-id <preset_voice_id> --text \"...\" --json\n```\n\nUse `pixverse voice models` and `pixverse voice presets --model <id>` when the user wants voice choice.\n\nFor natural or timed narration, use `../pixverse-voiceover/SKILL.md`. Generate natural speech sections, measure their audio and align captions afterward through `../pixverse-captions/SKILL.md`. Keep the final SRT under `projects/<slug>/prompts/` or `projects/<slug>/assets/subtitles/`. The following segmented queue is only for an explicit fixed-window take list:\n\n```bash\n\"${PVX}\" subtitles clean <srt> <clean.srt>\n\"${PVX}\" subtitles inspect <clean.srt>\n\"${PVX}\" subtitles style --format force-style\n\"${PVX}\" subtitles voice-queue <clean.srt> projects/<slug>/voice-queue.json --project <slug> --segments-dir projects/<slug>/prompts/tts-segments --voice-id <preset_voice_id> --language zh\n```\n\nExisting continuous audio can receive accurately aligned captions without another generation. Preserve accepted spoken words; clean display punctuation only when that style is wanted. Start from the plugin subtitle style and verify readability in the actual frame.\n\n## Music\n\nUse standalone music when BGM/soundtrack is requested:\n\n```bash\npixverse music models --json\npixverse create music --model music-2.6 --prompt \"...\" --instrumental --json\n```\n\nUse lyrics only when vocals are part of the concept. `music-3.0` (MiniMax Music 3.0)\nand `music-v2` (ElevenLabs Music V2) support lyrics, auto lyrics and instrumental\ngeneration; the default remains `music-2.6`. Read\n`../../skills-shared/pixverse-cli-1.4.4.md` for the reviewed alternatives.\n\nMusic is generated at auto-duration and edited to the picture; `--duration-seconds` and\n`--no-duration-auto` are refused by the queue because the service rejects fixed targets.\nTreat a requested length as an edit requirement: measure the returned audio and trim/fade\nlocally. See the shared audio craft reference.\n\n## Video Audio\n\nUse in-video generated sound when the brief wants the model to synchronize ambience and sound effects to the generated motion. Do not describe this as \"music none\" shorthand; write the no-music requirement in plain natural language.\n\nBefore using `--audio`, decide and record one of these modes:\n\nThese switches apply only when the selected model exposes them. Seedance 2.5 omits both\n`--audio` and `--no-audio`; keep sound intent in the prompt, inspect returned sound, and mute\nlocally for a silent deliverable. Missing toggle support does not establish missing audio.\n\nMiniMax H3/H3 Max is a different contract: generated audio is unsupported. Its\n`--audios` accepts reference inputs; it does not enable output sound. Plan separate\nvoice/music or local sound assets when sound is required, preserving paid approval.\n\n| Mode | Use when | Command stance |\n|---|---|---|\n| synced SFX/ambience | model should create sound tied to visible motion, with no music | `--audio`; use natural-language no-music instruction |\n| clean picture | sound will be rebuilt from separate stems or user asked for silent picture | `--no-audio` |\n| fused native audio | casual social/UGC clip where baked audio including music may be acceptable | `--audio` where supported; preserve the workflow’s complete sound/voice plan |\n\nFor in-video generated sound, use a plain-language audio paragraph and avoid musical cue words:\n\n```text\nThis video should not generate any music, background music, score, melody, rhythmic bed, or trailer-style musical hit.\nOnly generate synchronized sound effects and environmental ambience for what is visible in the shot.\nThe sound should be natural and quiet: physical Foley from the characters and objects, room tone, air, footsteps, fabric, breath, water, and other diegetic sounds.\nDo not add orchestral emotion, musical chimes, melodic pads, percussion, or any soundtrack-like layer.\n```\n\nIf the result contains unwanted music or unpleasant SFX/ambience, mark that audio rejected and strip/replace it. Do not mix BGM over a bad generated track.\n\nIf `quote queue` adds an audio note, treat it as an advisory reminder. If the prompt already says \"no music\" or the user's intent is clear, proceed to confirmation without changing the route. Only rewrite the sound paragraph when it is quick and obviously improves the prompt; do not switch to `--no-audio` or delay the run just to silence the note.\n\nFor post overlay, use `../pixverse-video-editing/SKILL.md` and keep audio assets separate.\n"
}SHA-256 of public snapshot: 7010d6371b32265826fbea3912a03f0b36c795ce0d7a8a53762c08ba3c3cdcd5