PluginBench
Skill
Review
Audit score 70

seedance

101-skills/superpowers

Generate AI videos with synchronized audio using ByteDance Seedance 2.0 via inference.sh CLI.

What is seedance?

Seedance 2.0 is a unified video generation model supporting text-to-video, image-to-video, and multimodal reference-to-video with synchronized audio generation. Use it to create social media videos, music videos, product demos, and animated content up to 1080p and 4-15 seconds long, with Pro, Fast, and Studio variants for different quality and consistency needs.

  • Generate videos from text prompts with optional synchronized audio
  • Animate still images or control both first and last frames of videos
  • Use multiple reference images, videos, and audio to guide generation and maintain character consistency
  • Create videos up to 1080p resolution and 4-15 seconds duration
  • Access Studio variants with private asset library for enhanced portrait and character consistency
  • Choose between Pro (best quality), Fast (cheaper, up to 720p), and Studio variants (quality + consistency)

How to install seedance

npx skills add https://github.com/101-skills/superpowers --skill seedance
Prerequisites
  • Install the belt CLI skill: npx skills add belt-sh/cli
  • Run belt login to authenticate with inference.sh
Claude Code
Cursor
Windsurf
Cline

How to use seedance

  1. 1.Install the belt CLI skill and authenticate with belt login
  2. 2.Choose a model variant: bytedance/seedance-2-0 (best quality), bytedance/seedance-2-0-fast (faster/cheaper), or bytedance/seedance-2-0-studio (consistency)
  3. 3.Prepare your inputs: prompt (required), and optionally image, reference_images, reference_videos, reference_audios, or end_image
  4. 4.Run belt app run [model-id] --input '{...}' with your parameters
  5. 5.Reference assets in your prompt using type + index (e.g., 'Image 1', 'Video 1', 'Audio 1') matching their position in the arrays
  6. 6.Adjust parameters like duration (4-15s), ratio (16:9, 9:16, etc.), resolution (480p-1080p), and generate_audio (true/false) as needed

Use cases

Good for
  • Create social media videos and short-form content with AI-generated visuals and sound
  • Generate product demo videos and advertising content with consistent branding
  • Produce music videos by syncing generated visuals with reference audio
  • Extend or edit existing videos by stitching clips or replacing elements while preserving motion
  • Generate animated content for presentations, tutorials, or creative projects with character consistency
Who it's for
  • Content creators and social media managers
  • Product marketers and advertising teams
  • Video editors and motion graphics professionals
  • Musicians and music video producers
  • Developers building video generation into applications

seedance FAQ

What's the difference between the model variants?

Seedance 2.0 offers best quality up to 1080p; Fast is cheaper and faster up to 720p; Studio variants add a private asset library for enhanced character and portrait consistency; Studio Fast combines both benefits.

Can I control both the start and end of a video?

Yes, use image-to-video mode with both image (first frame) and end_image (last frame) parameters, plus a prompt describing the transition.

How do I maintain character consistency across multiple videos?

Use the Studio variants (bytedance/seedance-2-0-studio or studio-fast) which upload reference images to a private asset library, or provide the same reference images in each generation.

Can I generate video without audio?

Yes, set generate_audio to false in your input parameters to skip audio generation.

What's the maximum video length and resolution?

Videos can be 4-15 seconds long; Pro variant supports up to 1080p, Fast variant up to 720p, and resolution can be set to 480p, 720p, or 1080p depending on the model.

Full instructions (SKILL.md)

Source of truth, from 101-skills/superpowers.


name: seedance description: "Generate videos with ByteDance Seedance 2.0 via inference.sh CLI. Unified model for text-to-video, image-to-video, and reference-to-video with synchronized audio, up to 1080p, 4-15s duration. Pro and Fast variants. Studio variants with private asset library for portrait consistency. Use for: social media videos, music videos, product demos, animated content, AI video with sound. Triggers: seedance, seedance 2, bytedance video, seedance t2v, seedance i2v, seedance r2v, video with audio, seedance 2.0, bytedance seedance, seedance studio" allowed-tools: Bash(belt *)

Install the belt CLI skill: npx skills add belt-sh/cli

Seedance 2.0 Video Generation

Generate videos with synchronized audio using ByteDance's Seedance 2.0 via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "a jazz band performing in a dimly lit club",
  "generate_audio": true
}'

Models

ModelApp IDBest For
Seedance 2.0bytedance/seedance-2-0Best quality, up to 1080p
Seedance 2.0 Fastbytedance/seedance-2-0-fastFaster generation, up to 720p
Seedance 2.0 Studiobytedance/seedance-2-0-studioQuality + private asset library for portrait consistency
Seedance 2.0 Studio Fastbytedance/seedance-2-0-studio-fastFast + private asset library for portrait consistency

All models support text-to-video, image-to-video, multimodal reference-to-video, and synchronized audio generation. Studio variants automatically upload reference images to the BytePlus private virtual portrait library for enhanced character consistency - particularly useful for faces and branded characters.

Modes

The model determines the generation mode from your inputs. These modes are mutually exclusive - use either first-frame/last-frame OR reference inputs, not both.

ModeInputsDescription
Text-to-Videoprompt onlyGenerate video from text description
Image-to-Videoprompt + imageAnimate a still image (first frame)
First+Last Frameprompt + image + end_imageControl start and end frames
Multimodal Referenceprompt + reference_images/reference_videos/reference_audiosGuide generation with reference material

Examples

Text-to-Video with Audio

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "ocean waves crashing on rocks during a storm, dramatic cinematic shot",
  "generate_audio": true,
  "duration": 10,
  "ratio": "16:9"
}'

Fast Mode (Cheaper)

belt app run bytedance/seedance-2-0-fast --input '{
  "prompt": "a butterfly landing on a flower in slow motion",
  "generate_audio": true
}'

Image-to-Video

Animate a still image into a video:

belt app run bytedance/seedance-2-0 --input '{
  "image": "https://your-image.jpg",
  "prompt": "gentle camera movement, leaves rustling in the wind",
  "generate_audio": true
}'

Image-to-Video with Start and End Frames

belt app run bytedance/seedance-2-0 --input '{
  "image": "https://start-frame.jpg",
  "end_image": "https://end-frame.jpg",
  "prompt": "smooth transition between scenes",
  "generate_audio": true
}'

Multi-Image Reference

Use multiple reference images to guide character appearance, outfits, and scene elements:

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "The girl from Image 1 wearing the outfit from Image 2 walks through the cafe from Image 3",
  "reference_images": [
    "https://character-portrait.jpg",
    "https://outfit-reference.jpg",
    "https://cafe-scene.jpg"
  ],
  "generate_audio": true,
  "duration": 8
}'

Video Editing (Replace Elements)

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "Replace the perfume in Video 1 with the face cream from Image 1, preserving all original motions and camera work",
  "reference_images": ["https://face-cream.jpg"],
  "reference_videos": ["https://original-video.mp4"],
  "generate_audio": true
}'

Video Extension (Stitch Clips)

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "Video 1 transitions smoothly into Video 2, then the camera enters the painting from Video 3",
  "reference_videos": [
    "https://clip1.mp4",
    "https://clip2.mp4",
    "https://clip3.mp4"
  ],
  "generate_audio": true,
  "duration": 8
}'

Reference with Audio

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "The musician from Image 1 performs the song from Audio 1, voice style referenced from Audio 1",
  "reference_images": ["https://musician.jpg"],
  "reference_audios": ["https://music.mp3"],
  "generate_audio": true
}'

Studio Mode (Portrait Consistency)

Studio variants upload images to BytePlus's private asset library for enhanced face/character consistency:

belt app run bytedance/seedance-2-0-studio --input '{
  "prompt": "The person in Image 1 smiles at the camera, golden hour lighting, cinematic",
  "reference_images": ["https://portrait.jpg"],
  "safety_identifier": "user-abc123",
  "generate_audio": true
}'

Product Ad with Multiple References

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "First-person POV product ad. Opening frame is Image 1, hand picks up the product. Camera pushes into close-up showing details. Use the camera movement style from Video 1. Background music from Audio 1.",
  "reference_images": ["https://product-hero.jpg", "https://product-detail.jpg"],
  "reference_videos": ["https://camera-style.mp4"],
  "reference_audios": ["https://bgm.mp3"],
  "generate_audio": true,
  "ratio": "9:16",
  "duration": 11
}'

Prompt Guide

Reference assets in your prompt using type + index: Image 1, Image 2, Video 1, Audio 1. The index is the position within that type in the arrays you provide. Do NOT use asset IDs in prompts.

Multimodal reference formula:

  • Image reference: "Refer to the [subject] from [Image N] to generate [scene], keeping [subject] consistent"
  • Video reference: "Refer to the [camera movement/action] from [Video N]"
  • Audio reference: "[Character] says: [dialogue], voice style referenced from [Audio N]"

Video editing formula:

  • Add: "At [timing] of [Video N], add [element]"
  • Remove: "Remove [element] from [Video N], keeping the rest unchanged"
  • Modify: "Replace [element] in [Video N] with [new element]"

Video extension formula:

  • Forward: "Generate content after [Video N]: [description]"
  • Backward: "Extend the opening of [Video N]: [description]"
  • Stitch: "[Video 1] + [transition] + followed by [Video 2]"

Parameters

ParameterTypeDefaultDescription
promptstringrequiredText description of the video
generate_audiobooleantrueGenerate synchronized audio
durationinteger5Duration in seconds (4-15), or -1 for auto
ratioenumadaptive21:9, 16:9, 4:3, 1:1, 3:4, 9:16, or adaptive
resolutionenum720p480p, 720p, 1080p (Fast: 480p, 720p only)
seedinteger-1Seed for reproducibility (-1 for random)
watermarkbooleanfalseAdd watermark to output
safety_identifierstring-Unique end-user identifier for safety policy (max 64 chars, hash of user ID recommended)
imagefile-First-frame image (mutually exclusive with reference inputs)
end_imagefile-Last-frame image (requires image)
reference_imagesfile[]-Reference images, up to 9 (mutually exclusive with image/end_image)
reference_videosfile[]-Reference videos, up to 3. Max 15s each, total max 15s. mp4/mov
reference_audiosfile[]-Reference audios, up to 3. Max 15s each, total max 15s. wav/mp3. Requires at least one image or video

Pricing

ModelPricing
Seedance 2.0$4.30-$7.70/M tokens (varies by resolution and input type)
Seedance 2.0 Fast$3.30-$5.60/M tokens

Token formula: (width x height x fps x duration) / 1024

Search Seedance Apps

belt app search "seedance"

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# All video generation models
npx skills add inference-sh/skills@ai-video-generation

# Google Veo
npx skills add inference-sh/skills@google-veo

# Image generation (for image-to-video)
npx skills add inference-sh/skills@ai-image-generation

# AI avatars & lipsync
npx skills add inference-sh/skills@ai-avatar-video

Browse all video apps: belt app list --category video

Documentation