agentsclimarketplace

Openrouter image2video

Skill QinghongLin/data2story-skill/skills/data2story-pro/designer/scripts/openrouter-image2video

Data Journalist Agent: Transforming Data into Verifiable Multimodal Story

Install
npx -y skills add QinghongLin/data2story-skill --skill openrouter-image2video

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

What its author says it does

Copied from the file, not written here

Animate a still image into a short video via OpenRouter. Default model google/veo-3.1-fast.

SKILL.md

3.4 KB, as published. Nobody here has run it

openrouter-image2video

Image + motion-prompt → video via OpenRouter. Default model: google/veo-3.1-fast.

Use this when you already have a strong still image and want to bring it to life with subtle motion (camera pan, parallax, gentle animation) while preserving the composition. For motion-from-scratch, use openrouter-text2video instead.

The script accepts either a remote image URL or a local image path; local files are base64-encoded and inlined into the request as a data URL.

Usage

Resolve TOOL_DIR = the directory containing this SKILL.md. Commands below use TOOL_DIR as a symbolic placeholder; replace it with the resolved, quoted path before running Bash.

export OPENROUTER_API_KEY=sk-or-v1-...

# From a local image you already generated with text2image
python3 TOOL_DIR/scripts/generate_video_from_image.py \
  --image PROJECT_DIR/assets/teaser.png \
  --prompt "slow parallax push-in, soft drift of ambient particles, no camera shake" \
  --duration 5 \
  --aspect-ratio 16:9 \
  --download PROJECT_DIR/assets/teaser.mp4

# Or from a remote URL
python3 TOOL_DIR/scripts/generate_video_from_image.py \
  --image-url "https://example.com/still.png" \
  --prompt "subtle camera dolly forward, gentle depth-of-field shift" \
  --download PROJECT_DIR/assets/scene.mp4

Flags

FlagDefaultDescription
--promptrequiredMotion prompt — describe what should move and how
--downloadrequiredOutput MP4 path
--imageone of --image / --image-url requiredLocal image path (PNG/JPG); will be base64-encoded
--image-urlone of --image / --image-url requiredRemote image URL
--modelgoogle/veo-3.1-fastAny OpenRouter image-to-video-capable model
--duration5Seconds
--aspect-ratio16:916:9, 9:16, 1:1, 4:3, 3:4, 21:9, 9:21
--resolution720pModel-dependent (e.g. 480p, 720p, 1080p)
--frame-rolefirstfirst or last — anchor frame role for the input image
--generate-audiooffGenerate audio with video (if model supports)
--poll-interval5Seconds between polls
--max-wait600Max total wait time

Flow

  1. POST /api/v1/videos with body:
    {
      "model": "google/veo-3.1-fast",
      "prompt": "...motion prompt...",
      "aspect_ratio": "16:9",
      "duration": 5,
      "resolution": "720p",
      "frame_images": [
        {
          "type": "image_url",
          "frame_type": "first_frame",
          "image_url": {"url": "data:image/png;base64,..." }
        }
      ]
    }
    
    The --frame-role first|last flag maps to frame_type: "first_frame"|"last_frame".
  2. GET /api/v1/videos/{id} every 5s until status == "completed"
  3. GET /api/v1/videos/{id}/content → raw MP4 bytes

Notes

  • Veo 3.1 Fast is optimized for low-latency image-to-video. Typical render ≈ 60-180s for a 5s 720p clip.
  • The motion prompt should describe motion only, not the subject (the subject comes from the image).
  • For subjects with prominent faces, keep motion subtle to avoid uncanny artifacts.
  • If you also want a defined ending state, supply two images via frame_images with roles first and last. The current script wires only one anchor frame; extend body["frame_images"] to add a second.

Keep looking

Skills are one crate of 328,083. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.