openrouter-image2video

SkillMedia

Animate a still image into a short video via OpenRouter. Default model google/veo-3.1-fast.

Available today. Use it from your connected AI after setup.

Connect ahel once, and every AI you use reads what you have installed.

Then ask your AI: use the openrouter-image2video skill

What this skill tells your AI

The instructions your AI receives, as published by qinghonglin/data2story-skill in skills/data2story-pro/designer/scripts/openrouter-image2video/SKILL.md and read by ahel’s review.

Image + motion-prompt → video via OpenRouter. Default model: google/veo-3.1-fast.

Use this when you already have a strong still image and want to bring it to life with subtle motion (camera pan, parallax, gentle animation) while preserving the composition. For motion-from-scratch, use openrouter-text2video instead.

The script accepts either a remote image URL or a local image path; local files are base64-encoded and inlined into the request as a data URL.

Usage

Resolve TOOL_DIR = the directory containing this SKILL.md. Commands below use TOOL_DIR as a symbolic placeholder; replace it with the resolved, quoted path before running Bash.

export OPENROUTER_API_KEY=sk-or-v1-...

# From a local image you already generated with text2image
python3 TOOL_DIR/scripts/generate_video_from_image.py \
  --image PROJECT_DIR/assets/teaser.png \
  --prompt "slow parallax push-in, soft drift of ambient particles, no camera shake" \
  --duration 5 \
  --aspect-ratio 16:9 \
  --download PROJECT_DIR/assets/teaser.mp4

# Or from a remote URL
python3 TOOL_DIR/scripts/generate_video_from_image.py \
  --image-url "https://example.com/still.png" \
  --prompt "subtle camera dolly forward, gentle depth-of-field shift" \
  --download PROJECT_DIR/assets/scene.mp4

Flags

FlagDefaultDescription
--promptrequiredMotion prompt — describe what should move and how
--downloadrequiredOutput MP4 path
--imageone of --image / --image-url requiredLocal image path (PNG/JPG); will be base64-encoded
--image-urlone of --image / --image-url requiredRemote image URL
--modelgoogle/veo-3.1-fastAny OpenRouter image-to-video-capable model
--duration5Seconds
--aspect-ratio16:916:9, 9:16, 1:1, 4:3, 3:4, 21:9, 9:21
--resolution720pModel-dependent (e.g. 480p, 720p, 1080p)
--frame-rolefirstfirst or last — anchor frame role for the input image
--generate-audiooffGenerate audio with video (if model supports)
--poll-interval5Seconds between polls
--max-wait600Max total wait time

Flow

  1. POST /api/v1/videos with body:
    {
      "model": "google/veo-3.1-fast",
      "prompt": "...motion prompt...",
      "aspect_ratio": "16:9",
      "duration": 5,
      "resolution": "720p",
      "frame_images": [
        {
          "type": "image_url",
          "frame_type": "first_frame",
          "image_url": {"url": "data:image/png;base64,..." }
        }
      ]
    }
    
    The --frame-role first|last flag maps to frame_type: "first_frame"|"last_frame".
  2. GET /api/v1/videos/{id} every 5s until status == "completed"
  3. GET /api/v1/videos/{id}/content → raw MP4 bytes

Notes

  • Veo 3.1 Fast is optimized for low-latency image-to-video. Typical render ≈ 60-180s for a 5s 720p clip.
  • The motion prompt should describe motion only, not the subject (the subject comes from the image).
  • For subjects with prominent faces, keep motion subtle to avoid uncanny artifacts.
  • If you also want a defined ending state, supply two images via frame_images with roles first and last. The current script wires only one anchor frame; extend body["frame_images"] to add a second.

Signals

GitHub stars
155
Forks
22
Last commit
Jul 2026
Advanced
Catalog kind
skill
Gateway key
openrouter-image2video
Source
github.com/qinghonglin/data2story-skill