← PixVerseCONTENT HISTORY

Update to PixVerse

Snapshot Sep 30, 2026 · 23:13 UTC · version 1.3.2

Collection source: not recorded for this historical snapshot.

WHAT CHANGED · RULE-BASED ANALYSIS

First saved snapshot

No earlier snapshot is available to establish a change.

Compare saved observations

Download comparison JSON
Full technical diff · 0 changed fields
Full snapshot data
{
  "description": "Video captions: create or correct timed subtitles from actual speech and render readable captions into supplied video, including translated subtitles when requested.",
  "included_files": [
    {
      "relative_path": "agents/openai.yaml",
      "size_in_bytes": 276
    },
    {
      "relative_path": "references/burn-in.md",
      "size_in_bytes": 3213
    }
  ],
  "name": "pixverse-captions",
  "skill_md_contents": "---\nname: pixverse-captions\ndescription: \"Video captions: create or correct timed subtitles from actual speech and render readable captions into supplied video, including translated subtitles when requested.\"\n---\n\n# Captions\n\nWhen this workflow generates images or video, apply `../../skills-shared/quality-policy.md`:\nSunburst 2K/high for images, Seedance 2.5 1080p with automatic prompt enhancement for video,\nand a stopped upgrade/fallback choice for Free/Basic or model-entitlement rejection.\n\n\nMake captions follow what the audience hears. Work from the actual recording, even when a draft script exists.\n\nFor a burned-in deliverable, read `./references/burn-in.md` before transcription or layout.\nIt defines the four ordered gates, verified-word alignment, bold/paper/clean/UGC looks,\nsafe geometry, glyph coverage and voice-tail checks. Sidecar and translation requests\nremain separate; neither requires rendering new video unless requested.\n\n## Establish The Words And Clock\n\nInspect the video duration, audio tracks, frame size, orientation and requested output. Reuse a supplied accurate timed transcript. Otherwise transcribe or align with a tool actually available in this host, then check names, numbers, language changes and uncertain words against the audio. If alignment is unavailable, offer a clearly marked timing draft or manually time a short clip; do not present guessed word timestamps as measured.\n\nA supplied script is a candidate transcript until compared with the recording. Translation may change displayed wording, but retain the original audio time windows and meaning. Keep original and translated tracks separate when both are requested.\n\n## Compose Readable Captions\n\n- Split by meaning and breathing, not by fixed character counts alone. Avoid flashing one-word fragments unless requested.\n- Measure text with the selected font. One or two lines may be appropriate; leave faces, product actions and platform controls visible.\n- Use the plugin subtitle style as a starting point, then check the actual portrait or landscape frame. Preserve user typography and punctuation preferences.\n- Export a timed SRT or ASS. Exact text is a render layer; never ask the video model to draw the captions.\n- For karaoke or word emphasis, use measured word timing. Sentence timing cannot justify precise word highlights.\n- For word-highlighted or word-by-word captions, write the lines as a script (`../../skills-shared/semantic-script.md`),\n  align them to the recording (`../../skills-shared/word-timing.md`) and generate the ASS with\n  `\"${PVX}\" timeline captions <timeline.json> --to captions.ass --style karaoke|bold|clean|ugc|pop`;\n  role colours come from the script's speaker labels. Burn with the FFmpeg `ass` filter as usual.\n\n## Render And Check\n\nFor ordinary local delivery, use FFmpeg and the existing subtitle helpers in `../pixverse-video-editing/SKILL.md`. Test font availability and renderer support before a full export. Inspect representative frames with short, long and mixed-language captions; watch the start and end plus several sentence transitions.\n\nPreserve the selected soundtrack and requested video properties. Verify duration, readable placement, encoding and caption sync after rendering. A subtitle file beside an uncaptioned video does not satisfy burned-in delivery.\n\nChanging color, font or line breaks only rerenders captions. Do not regenerate the video or TTS. Deliver the captioned file and the editable timing file; retain the source video.\n\n## Execution\n\nRead `../../skills-shared/production-brief.md` for reference roles, prompt construction and revisions.\nFor media execution, follow `../../skills-shared/cli-workflow.md`; load the gateway before manual\nCLI commands or queue specs. Resolve the plugin root from this skill's installed path:\nset `PVX=\"<absolute-plugin-root>/scripts/pvx\"` and invoke `\"${PVX}\"` with the documented arguments.\nPrompt-only work ends with the usable plan and prompts; it needs no account or generation calls.\n"
}

SHA-256 of public snapshot: 5187e6a5b65d835bd31293125e384178f001481cca9a127446989bef32534cb4