Audio Generation

PostPlusAI/postplus-skills/skills/10-routing/audio-generation

by PostPlusAI7f28d1494958f69136a7d6fe85fdea3942a300f3No licenseListed Oct 9, 2026Updated Oct 9, 2026

Plan TTS, voice cloning, voice change, translated dub, or lip-sync audio. Resolve voice and reference policy before handing a ready request to voice-batch-runner.

Instructions onlyAI & Agents
AI-generated overview

Plans text-to-speech, voice cloning, dubbing and lip-sync audio requests and hands them to an execution runner.

What it does
This controller skill classifies an audio request into a task class such as tts, change_voice, translate_dub, voice_clone_take, podcast_audio or lip_sync_handoff. It resolves script, voice and reference policy, including which reference audio is binding versus inspiration-only. It then produces a handoff artifact with fields like taskClass, scriptPolicy, voicePolicy, referencePolicy, runnerHandoff, nextVideoHandoff and mustNotDo. It does not submit jobs itself.
When to use it
Use it when the desired final asset is generated audio or audio prepared for a video render, including TTS, voice design, voice cloning, voice change, translated dub, podcast audio or lip-sync handoff. It is not for transcribing existing audio or for requests already normalized for execution.
Requirements
No scripts or tools are shipped; it is instruction-driven. It needs the user's intent, source evidence and owned input artifacts, and it hands off to companion skills such as voice-batch-runner, video-batch-runner or audio-transcription.

Audio Generation

Use When

  • The desired final asset is generated audio or audio prepared for a video render.
  • The request includes TTS, voice design, voice cloning, voice change, translated dub, podcast audio, or lip-sync handoff.
  • The next decision is audio task class, reference policy, and runner handoff.

Do Not Use When

  • The user needs speech-to-text from existing audio. Use audio-transcription; use its local subtitle reference when an existing timed transcript needs subtitle files.
  • The voice request is already normalized for execution. Use voice-batch-runner.
  • The final work is a full video production pipeline. Use video-batch-runner after the audio handoff is clear.

Core Boundary

This is the audio generation controller. It does not submit jobs.

It must classify the task and hand off execution. It must not let a runner invent voice strategy, translation policy, or lip-sync intent.

Task Classes

Task classUse whenHandoff
ttsnew spoken audio from scriptvoice-batch-runner with voice design rules
change_voicepreserve script, alter voice identity or deliveryreference contract, then voice-batch-runner
translate_dubtranslate and dub source audiorequire language, meaning-preservation, and timing policy
voice_clone_takeapproved reference voice should preserve timbrebind reference audio, then voice-batch-runner
podcast_audiospeaker-led or conversational audiocreate voice/script handoff before video assembly
lip_sync_handoffaudio drives talking-head or UGC rendervoice-batch-runner, then video-batch-runner

Reference Rules

  • Approved voice reference audio is binding.
  • Accent, energy, cadence, or genre examples are inspiration-only unless the user explicitly binds them.
  • Source audio used only for translation meaning is not a voice identity binding unless stated.
  • Excluded voices, music, or effects must not enter the runner request.

Routing Table

If not audio-generationSend to
Transcribe existing audioaudio-transcription
Need generated image/video around audiovideo-batch-runner
Need normalized hosted voice executionvoice-batch-runner
Need lip-sync video after audiovideo-batch-runner

Output Shape

Return:

  • taskClass
  • scriptPolicy
  • voicePolicy
  • referencePolicy
  • runnerHandoff
  • nextVideoHandoff when lip-sync or video assembly follows
  • mustNotDo

Stop Conditions

  • Stop when required user intent, source evidence, or owned input artifacts are missing and guessing would change the result.
  • Do not ask voice-batch-runner to decide the creative role of the voice.

Public Command Boundary

  • Choose the smallest matching command or workflow from the user input and run it directly.

  • This public skill is instruction-driven. Produce the controller handoff artifact directly from the available evidence.

  • Do not call private provider/runtime paths or unpublished local tools.

  • If the CLI returns a quote-confirmation challenge, obtain user approval for its scope and cost before running postplus quote confirm --json --challenge-file <challenge.json> and retry with the returned token.

Source and attribution

Source:PostPlusAI/postplus-skillsinskills/10-routing/audio-generationat commit7f28d14

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal