Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
postplusai avatar

Audio Generation

  • 376 installs
  • 19 repo stars
  • Updated July 27, 2026
  • postplusai/postplus-skills

Classify TTS, voice change, dub, podcast, or lip-sync audio requests and route execution to the right PostPlus runner without inventing voice or translation policy.

About

Audio Generation is a PostPlus routing-contract skill for solo builders and small teams automating spoken and dubbed media with agents. Before any TTS, persona voice, voice change, translated dub, cloned take, podcast audio, or lip-sync handoff runs, the skill decides the audio task class, reference policy, and which downstream runner owns execution. It prevents runners from guessing voice design, translation rules, or lip-sync intent. Use it when the desired outcome is generated audio or audio staged for a video render; skip it when the user only needs transcription or subtitles (media-router), when requests are already normalized for voice-batch-runner, or when the primary goal is a full video pipeline without a clear audio class. The skill fits agent-tooling and content workflows where predictable contracts matter more than one-off chat improvisation.

  • Classifies audio work into tts, change_voice, dub, podcast, and lip-sync handoff task classes before any job submission
  • Enforces reference policy and runner handoff so execution cannot invent voice strategy or translation policy
  • Explicit do-not-use boundaries: transcription/analysis → media-router; normalized voice jobs → voice-batch-runner
  • Controller-only: classifies task class and hands off—does not submit generation jobs itself
  • Maps final-asset intent (spoken audio vs video-prepared audio) to voice-batch-runner or video-generation / ugc-flow

Audio Generation by the numbers

  • 376 all-time installs (skills.sh)
  • +43 installs in the week ending Aug 4, 2026 (Skillselion tracking)
  • Ranked #2,066 of 16,546 AI & Agent Building skills by installs in the Skillselion catalog
  • Security screen: LOW risk (skills.sh audit)
  • Data as of Aug 5, 2026 (Skillselion catalog sync)
npx skills add https://github.com/postplusai/postplus-skills --skill audio-generation

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs376
repo stars19
Security audit2 / 3 scanners passed
Last updatedJuly 27, 2026
Repositorypostplusai/postplus-skills

What it does

Classify TTS, voice change, dub, podcast, or lip-sync audio requests and route execution to the right PostPlus runner without inventing voice or translation policy.

Files

SKILL.mdMarkdownGitHub ↗

Audio Generation

Use When

  • The desired final asset is generated audio or audio prepared for a video

render.

  • The request includes TTS, voice design, voice cloning, voice change,

translated dub, podcast audio, or lip-sync handoff.

  • The next decision is audio task class, reference policy, and runner handoff.

Do Not Use When

  • The user only needs transcription, subtitles, or audio analysis. Use

media-router.

  • The voice request is already normalized for execution. Use

voice-batch-runner.

  • The final work is a full video production pipeline. Use video-generation or

ugc-flow after the audio handoff is clear.

Core Boundary

This is the audio generation controller. It does not submit jobs.

It must classify the task and hand off execution. It must not let a runner invent voice strategy, translation policy, or lip-sync intent.

Task Classes

Task classUse whenHandoff
ttsnew spoken audio from scriptvoice-batch-runner with voice design rules
change_voicepreserve script, alter voice identity or deliveryreference contract, then voice-batch-runner
translate_dubtranslate and dub source audiorequire language, meaning-preservation, and timing policy
voice_clone_takeapproved reference voice should preserve timbrebind reference audio, then voice-batch-runner
podcast_audiospeaker-led or conversational audiocreate voice/script handoff before video assembly
lip_sync_handoffaudio drives talking-head or UGC rendervoice-batch-runner, then video-generation

Reference Rules

  • Approved voice reference audio is binding.
  • Accent, energy, cadence, or genre examples are inspiration-only unless the

user explicitly binds them.

  • Source audio used only for translation meaning is not a voice identity

binding unless stated.

  • Excluded voices, music, or effects must not enter the runner request.

Routing Table

If not audio-generationSend to
Transcribe or analyze existing audiomedia-router
Need generated image/video around audiovideo-generation
Need normalized hosted voice executionvoice-batch-runner
Need lip-sync video after audiovideo-generation

Output Shape

Return:

  • taskClass
  • scriptPolicy
  • voicePolicy
  • referencePolicy
  • runnerHandoff
  • nextVideoHandoff when lip-sync or video assembly follows
  • mustNotDo

Stop Conditions

  • Stop when required user intent, source evidence, or owned input artifacts are

missing and guessing would change the result.

  • Do not ask voice-batch-runner to decide the creative role of the voice.
  • If an owned CLI or script command fails, report the exact error and stop. Do

not bypass the failure with metadata-only answers, readiness probing, local payload rewrites, fallback providers, or unpublished tools.

Public Command Boundary

  • Choose the smallest matching command or workflow from the user input and run

it directly.

  • If an owned CLI or script command fails, report the exact error and stop. Do

not bypass the failure with metadata-only answers, readiness probing, local payload rewrites, fallback providers, or unpublished tools.

  • This public skill is instruction-driven. Produce the controller handoff

artifact directly from the available evidence.

  • Do not call private provider/runtime paths or unpublished local tools.
  • If the CLI returns a quote-confirmation challenge, run postplus quote confirm --json --challenge-file <challenge.json> and retry with the returned token.

Related skills

FAQ

Is Audio Generation safe to install?

skills.sh reports 2 of 3 security scanners passed. Review the Security Audits panel on this page before installing in production.

AI & Agent Buildingautomationagents

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.