Video Transcription

PostPlusAI/postplus-skills/skills/40-creative/video-transcription

by PostPlusAI7f28d1494958f69136a7d6fe85fdea3942a300f3No licenseListed Oct 9, 2026Updated Oct 9, 2026

Transcribe speech in local or remote videos with timestamps. Convert existing timed transcripts locally into SRT or ASS; use media-analysis for visual understanding.

Instructions onlyDesign & Creative
AI-generated overview

Transcribes speech in local or remote videos with timestamps and converts existing timed transcripts into SRT or ASS subtitles.

What it does
This skill guides an agent through hosted video transcription using the postplus media transcribe command, accepting a local path, HTTPS URL, existing media reference, or data URI. It requests timestamps by default so results can drive subtitles or edit decisions, and it handles asynchronous runs by preserving the returned handle and following the CLI's reported next action. When a timed transcript already exists, it converts it locally into SRT or ASS using the bundled subtitle conversion reference instead of submitting another hosted request.
When to use it
Use it when the input is a video file and the goal is speech extraction, a timed transcript, caption generation, a multilingual transcript, or edit-prep timestamps. Use a different skill for semantic visual analysis, and skip transcription when a timed transcript already exists and only subtitle output is needed.
Requirements
Requires the postplus CLI with the media transcribe verb and network access for the hosted, asynchronous transcription job. A source video path, HTTPS URL, media reference, or data URI is needed, along with a derived duration value for request validation. It ships no scripts; it includes a subtitle conversion reference document.

Video Transcription

Use When

  • The input is a video file and the goal is speech extraction, timed transcript, caption generation, multilingual transcript, or edit-prep timestamps.
  • Use media-analysis instead when the user needs semantic visual analysis.

If a timed transcript already exists and only subtitle output is requested, skip transcription and read the local subtitle conversion reference.

Do Not Use When

  • The task needs new creative generation or visual analysis rather than speech or subtitles.
  • Required inputs are missing and guessing would change the result.

Execution Boundary

  • Hosted video transcription runs through the public postplus media transcribe verb and is async. The generated example below shows the endpoint key.
  • Pass a local path, HTTPS URL, existing PostPlus media reference, or data URI directly to --video. The CLI validates and prepares local media before the single hosted submit.
  • Request timestamps by default when results drive subtitles or edit decisions.
  • Hosted video transcription is async. Submit records the run handle, current status, and transcript artifacts when available. Inspect the returned format; do not assume a fixed normalized transcript schema.

Source And Path

  • Before submit, derive durationSeconds from the source video or URL and pass it through the endpoint's duration flag for request validation.
  • Start with one source file before larger batches.
  • Keep internal requests, responses, normalized transcripts, and downloaded artifacts under .postplus/video-transcription; keep final user-facing transcript exports outside .postplus.

Handoff

  • If status is pending, preserve the result path and follow the CLI-returned action or resume command for the same operation. Do not submit another job. Stop and report when the CLI wait/recovery boundary is reached.
  • When SRT/ASS is requested, use the actual timed transcript and read local subtitle conversion [blocked]. Convert locally without another hosted request; do not invent a CLI export command.

Stop Conditions

  • Stop when required user intent, source evidence, or owned input artifacts are missing and guessing would change the result.

Public Command Boundary

  • Choose the smallest matching command or workflow from the user input and run it directly.

  • Readiness diagnostics: postplus doctor --skill video-transcription.

  • Use postplus media schema --json only when you need the full endpoint, flag, and enum contract or are repairing an unknown request shape.

  • Run the hosted transcription job with the generated command below; do not use another execution interface.

  • Pass the source directly through --video; do not pre-upload it or construct a manual request object.

  • If the CLI returns a quote-confirmation challenge, obtain user approval for its scope and cost before running postplus quote confirm --json --challenge-file <challenge.json> and retry with the returned token.

<!-- BEGIN GENERATED EXECUTION EXAMPLE -->
bash
postplus media transcribe transcription-video \  --video ./reference.mp4 \  --duration-seconds 1 \  --wait \  --output ./result.json

Follow the CLI's structured result and reported next action; do not infer recovery from free-text messages. Wait for explicit user approval when requested; an action does not authorize spending, publishing, or overwriting. Resume the same operation through its returned checkpoint or action; never resubmit uncertain work, repeat exhausted recovery, or switch providers to bypass failure.

<!-- END GENERATED EXECUTION EXAMPLE -->

Source and attribution

Source:PostPlusAI/postplus-skillsinskills/40-creative/video-transcriptionat commit7f28d14

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal