Generative Media Prompting
SkillMediaThe shared brief-to-prompt convention behind the media point tasks, parse the brief, select the model or tool, construct the prompt, then QA the output against the brief. Use when prompting for image, video, speech, or music generation or editing.
Instructions available. Your AI can read the instructions. Execution depends on the setup they require.
Account requirements not reviewed. Check the skill instructions before use; ahel provides instructions and does not run this skill.
Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.
Then ask your AI: use the Generative Media Prompting skill
What this skill tells your AI
The instructions your AI receives, as published by a5c-ai/babysitter in library/specializations/media/skills/generative-media-prompting/SKILL.md and read by ahel’s review.
All six media point tasks open with the same four-step shape, visible verbatim in their
@description headers. Writing it down once removes six copies of the same tacit
convention.
The four steps
- Parse the brief. Extract what the request actually asks for — the creative intent for a generation task, or the source asset plus the requested operation for an editing task.
- Select the model or tool. Choose the model (generation) or tool (editing) that fits the parsed brief. Each point task names the candidates it selects among; the selection is part of the task, not a caller input.
- Construct the prompt. Turn the parsed brief into the prompt (and, where the task supports them, the structured parameters that accompany it).
- QA the output against the brief. Validate the result before returning it. Every point task ends in a validation step, and what it validates is modality-specific.
Per-modality notes
Limited to what the existing files already state:
image-generation.js— generates variants in parallel; validates technical and creative quality; organises outputs with metadata.image-editing.js— selects among named editing tools; validates edge quality, color consistency, and artifact absence.video-generation.js— the parsed brief includes the request mode (text-to-video, image-to-video, video-to-video); prompt construction carries camera, lighting, and composition parameters; a low-quality output is retried with a fallback model.video-editing.js— selects among named editing tools and runs a per-operation pipeline; validates frame consistency and audio sync.speech-generation.js— the brief includes language, style, emotion, and SSML; validates naturalness, pronunciation, and audio specs.music-generation.js— the brief includes genre, mood, duration, and instruments; mastering and stem separation are applied only if requested; validates musical coherence and technical audio.
Scope
This skill describes prompt construction only. Publication, review, and licensing
decisions are out of scope and belong to
../../media-production-pipeline.js. No model lists
or vendor guidance beyond what the point-task files themselves name.
Signals
- GitHub stars
- 2k
- Forks
- 112
- Last commit
- Sep 2026
Advanced
- Item type
- skill
- Key
generative-media-prompting- Source
- github.com/a5c-ai/babysitter