PluginBench
Skill
Review
Audit score 70

image-to-video

101-skills/superpowers

Convert still images to animated videos with AI models—select the right tool for realistic motion, fabric physics, or versatile animation.

What is image-to-video?

Guide to animating static images into videos using inference.sh CLI and models like Wan 2.5 i2v, Seedance, Fabric, and Grok Video. Choose based on your content type (landscapes, portraits, fabric, products) and desired motion style. Best for creating short, subtle-motion clips (2–8 seconds) from photographs or generated stills.

  • Select from six image-to-video models optimized for different scenarios (realistic motion, fabric physics, high resolution, speed)
  • Prompt camera movements (dolly, pan, tilt, orbit, crane) and subject motion (rippling water, flowing fabric, breathing, etc.)
  • Generate 2–8 second video clips with subtle, natural motion using the belt CLI
  • Extend duration by generating multiple clips and stitching them together
  • Create cinemagraph effects (motion in one element only) and product animations
  • Upscale, add foley audio, and merge audio with video in a complete pipeline

How to install image-to-video

npx skills add https://github.com/101-skills/superpowers --skill image-to-video
Prerequisites
  • Install the belt CLI skill: `npx skills add belt-sh/cli`
  • Run `belt login` to authenticate with inference.sh
  • Have a still image (photograph, generated image, or artwork) ready to animate
Claude Code
Cursor
Windsurf
Cline

How to use image-to-video

  1. 1.Log in to inference.sh: `belt login`
  2. 2.Generate or prepare a still image (or use `belt app run falai/flux-dev-lora` to create one)
  3. 3.Select the appropriate model from the table based on your content type and motion needs
  4. 4.Craft a motion prompt using camera movement keywords (e.g., 'slow dolly forward') and subject motion (e.g., 'water rippling')
  5. 5.Run the model with `belt app run [MODEL_ID] --input '{"prompt": "...", "image": "path/to/image.png"}'`
  6. 6.For videos longer than 8 seconds, generate multiple 2–5 second clips and stitch with `belt app run infsh/media-merger`
  7. 7.Optionally upscale, add foley audio, and merge audio back into the video using the full pipeline commands

Use cases

Good for
  • Animate landscape photographs with natural elements (water, clouds, light shifts) for social media or background loops
  • Create product showcase videos with 360° orbits and lighting effects for e-commerce
  • Generate portrait animations with subtle breathing, eye blinks, and hair movement
  • Produce fabric and cloth simulations for fashion or interior design visualization
  • Build architectural walkthroughs with slow camera movement through spaces
Who it's for
  • Content creators and video producers working with still images
  • Product photographers and e-commerce teams
  • Architects and interior designers visualizing spaces
  • Social media managers creating animated posts and backgrounds
  • Motion designers and VFX artists prototyping animations

image-to-video FAQ

Which model should I use for a landscape with water and clouds?

Wan 2.5 i2v is best for realistic, natural motion in landscapes. It excels at subtle water ripples, cloud drifting, and light shifts while maintaining image quality.

Why does my animation look distorted or have artifacts?

AI video models produce better results with subtle, gentle motion. Avoid dramatic action or fast movement; instead use keywords like 'slow,' 'gentle,' and 'subtle' in your prompts.

How long can my video be?

2–3 seconds gives highest quality; 4–5 seconds is good for social media; 6–8 seconds is acceptable. Beyond 10 seconds, quality degrades. For longer videos, generate multiple short clips and stitch them together.

Can I add audio to my video?

Yes. Use `belt app run infsh/hunyuanvideo-foley` to generate ambient audio, then merge it with `belt app run infsh/video-audio-merger`. Seedance 2.0 also has a `generate_audio` option.

What is a cinemagraph and how do I create one?

A cinemagraph is a still photo where only one element moves (e.g., a waterfall in a frozen landscape). Generate your still image, then prompt for motion only in that specific element, keeping duration to 2–4 seconds.

Full instructions (SKILL.md)

Source of truth, from 101-skills/superpowers.


name: image-to-video description: "Still-to-video conversion guide: model selection, motion prompting, and camera movement. Covers Wan 2.5 i2v, Seedance, Fabric, Grok Video with when to use each. Use for: animating images, creating video from stills, adding motion, product animations. Triggers: image to video, i2v, animate image, still to video, add motion to image, image animation, photo to video, animate still, wan i2v, image2video, bring image to life, animate photo, motion from image" allowed-tools: Bash(belt *)

Install the belt CLI skill: npx skills add belt-sh/cli

Image to Video

Convert still images to animated videos via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate a still image
belt app run falai/flux-dev-lora --input '{
  "prompt": "serene mountain lake at sunset, snow-capped peaks reflected in still water, golden hour light, landscape photography",
  "width": 1248,
  "height": 832
}'

# Animate it
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "gentle ripples on the lake surface, clouds slowly drifting, warm light shifting, birds flying in the distance",
  "image": "path/to/lake-image.png"
}'

Model Selection

ModelApp IDBest ForMotion Style
Wan 2.5 i2vfalai/wan-2-5-i2vRealistic motion, natural movementPhotorealistic, subtle
WAN-I2V (Pruna)pruna/wan-i2vEconomical, fast, 480p/720pNatural, efficient
Seedance 2.0bytedance/seedance-2-0Up to 1080p, sync audio, all input typesVersatile, high quality
Seedance 2.0 Fastbytedance/seedance-2-0-fastFast variant, same capabilitiesVersatile, fast
Fabric 1.0falai/fabric-1-0Cloth, fabric, liquid, flowing materialsPhysics-based flow
Grok Imagine Videoxai/grok-imagine-videoGeneral animation, text-guidedVersatile

When to Use Each

ScenarioBest ModelWhy
Landscape with water/cloudsWan 2.5 i2vBest at natural, realistic motion
Portrait with subtle expressionWan 2.5 i2vMaintains face fidelity
Product with fabric/clothFabric 1.0Specialized in material physics
Flag waving, curtain flowingFabric 1.0Cloth simulation
Illustrated/artistic imageSeedance 2.0Matches stylized content
General "bring to life"Seedance 2.0Good all-rounder, up to 1080p
Quick test/iterationSeedance 2.0 FastFaster generation

Motion Types

Camera Movement

MovementPrompt KeywordEffect
Push in / Dolly forward"slow dolly forward", "camera pushes in"Increasing intimacy/focus
Pull out / Dolly back"camera pulls back", "slow zoom out"Reveal, context
Pan left/right"camera pans slowly to the right"Scanning, following
Tilt up/down"camera tilts upward"Revealing height
Orbit"camera orbits around the subject"3D exploration
Crane up"camera rises upward"Grand reveal
Static(no camera movement prompt)Subject motion only

Subject Motion

TypePrompt Examples
Natural elements"water rippling", "clouds drifting", "leaves rustling in breeze"
Hair/clothing"hair blowing gently in wind", "dress fabric flowing"
Atmospheric"fog slowly rolling", "dust particles floating in light beams"
Character"person slowly turns to camera", "subtle breathing motion"
Mechanical"gears turning", "clock hands moving"
Liquid"coffee steam rising", "paint dripping", "water pouring"

Prompting Best Practices

The Golden Rule: Subtle > Dramatic

AI video models produce better results with gentle, subtle motion than dramatic action. Requesting too much movement causes distortion and artifacts.

❌ "person running and jumping over obstacles while the camera spins"
✅ "person slowly walking forward, gentle breeze, camera follows alongside"

❌ "explosion with debris flying everywhere"
✅ "candle flame flickering gently, warm ambient light shifting"

❌ "fast zoom into the eyes with dramatic camera shake"
✅ "slow dolly forward toward the subject, subtle focus shift"

Prompt Structure

[Camera movement] + [Subject motion] + [Atmospheric effects] + [Mood/pace]

Examples by Scenario

# Landscape animation
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "gentle camera pan right, water reflecting moving clouds, trees swaying slightly in breeze, warm golden light, peaceful and slow",
  "image": "landscape.png"
}'

# Portrait animation
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "subtle breathing motion, slight head turn, natural eye blink, hair moving gently, soft ambient lighting shifts",
  "image": "portrait.png"
}'

# Product shot animation
belt app run bytedance/seedance-2-0 --input '{
  "prompt": "slow 360 degree orbit around the product, gentle spotlight movement, subtle reflections shifting, premium product showcase, smooth motion",
  "image": "product.png",
  "generate_audio": true
}'

# Fabric/cloth animation
belt app run falai/fabric-1-0 --input '{
  "prompt": "fabric flowing and rippling in gentle wind, natural cloth physics, soft movement",
  "image": "fabric-scene.png"
}'

# Architectural visualization
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "slow dolly forward through the entrance, slight camera tilt upward, ambient light filtering through windows, dust particles in light beams",
  "image": "building-interior.png"
}'

Duration Guidelines

DurationQualityUse For
2-3 secondsHighest qualityGIFs, looping backgrounds, cinemagraphs
4-5 secondsHigh qualitySocial media posts, product reveals
6-8 secondsGood qualityShort clips, transitions
10+ secondsQuality degradesAvoid unless stitching shorter clips

Extending Duration

For longer videos, generate multiple short clips and stitch:

# Generate 3 clips from the same image with progressive motion
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "slow pan left, gentle water motion",
  "image": "scene.png"
}' --no-wait

belt app run falai/wan-2-5-i2v --input '{
  "prompt": "continuing pan, clouds shifting, light changing",
  "image": "scene.png"
}' --no-wait

# Stitch together
belt app run infsh/media-merger --input '{
  "media": ["clip1.mp4", "clip2.mp4"]
}'

The Full Workflow

Still-to-Final-Video Pipeline

# 1. Generate source image (best quality)
belt app run bytedance/seedream-4-5 --input '{
  "prompt": "cinematic landscape, misty mountains at dawn, lake in foreground, dramatic clouds, golden hour, 4K quality, professional photography",
  "size": "2K"
}'

# 2. Animate the image
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "gentle mist rolling through the valley, lake surface rippling, clouds slowly moving, birds in distance, warm light shifting",
  "image": "landscape.png"
}'

# 3. Upscale video if needed
belt app run falai/topaz-video-upscaler --input '{
  "video": "animated-landscape.mp4"
}'

# 4. Add ambient audio
belt app run infsh/hunyuanvideo-foley --input '{
  "video": "animated-landscape.mp4",
  "prompt": "gentle nature ambience, distant birds, soft wind, water lapping"
}'

# 5. Merge video with audio
belt app run infsh/video-audio-merger --input '{
  "video": "upscaled-landscape.mp4",
  "audio": "ambient-audio.mp3"
}'

Cinemagraph Effect

A cinemagraph is a still photo where only one element moves (e.g., waterfall moving in an otherwise frozen scene). To achieve this:

  1. Generate the still image with the motion element clearly defined
  2. Prompt for motion only in that specific element
  3. Keep to 2-4 seconds for seamless looping
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "only the waterfall is moving, everything else remains perfectly still, water cascading smoothly, rest of scene frozen",
  "image": "waterfall-scene.png"
}'

Common Mistakes

MistakeProblemFix
Too much motion requestedDistortion, artifacts, warpingSubtle > dramatic, always
Wrong model for content typePoor resultsUse selection guide above
Clips too long (10s+)Quality degrades significantlyKeep to 3-5 seconds, stitch if needed
No camera movement specifiedRandom/unpredictable motionAlways specify camera behavior
Conflicting motion directionsChaotic, unnaturalOne primary motion direction
Low-res source imageLow-res video outputStart with highest quality source
Complex action scenesModels can't handleKeep motion simple and natural

Related Skills

npx skills add inference-sh/skills@ai-video-generation
npx skills add inference-sh/skills@ai-image-generation
npx skills add inference-sh/skills@p-video
npx skills add inference-sh/skills@video-prompting-guide
npx skills add inference-sh/skills@prompt-engineering

Browse all apps: belt app list