seedance

作者: 101-skills

透過 inference.sh CLI 使用字節跳動 Seedance 2.0 生成影片。統一模型支援文字轉影片、圖片轉影片及參考轉影片,同步音訊,最高 1080p,時長 4-15 秒。提供 Pro 與 Fast 版本。Studio 版本具備私有素材庫,可保持人物一致性。適用於:社群媒體影片、音樂影片、產品展示、動畫內容、附音效的 AI 影片。觸發詞:seedance、seedance 2、bytedance video、seedance t2v、seedance i2v、seedance r2v、video with audio...

npx skills add https://github.com/101-skills/skills --skill seedance

Install the belt CLI skill: npx skills add belt-sh/cli

Seedance 2.0 Video Generation

Generate videos with synchronized audio using ByteDance's Seedance 2.0 via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "a jazz band performing in a dimly lit club",
  "generate_audio": true
}'

Models

ModelApp IDBest For
Seedance 2.0bytedance/seedance-2-0Best quality, up to 1080p
Seedance 2.0 Fastbytedance/seedance-2-0-fastFaster generation, up to 720p
Seedance 2.0 Studiobytedance/seedance-2-0-studioQuality + private asset library for portrait consistency
Seedance 2.0 Studio Fastbytedance/seedance-2-0-studio-fastFast + private asset library for portrait consistency

All models support text-to-video, image-to-video, multimodal reference-to-video, and synchronized audio generation. Studio variants automatically upload reference images to the BytePlus private virtual portrait library for enhanced character consistency - particularly useful for faces and branded characters.

Modes

The model determines the generation mode from your inputs. These modes are mutually exclusive - use either first-frame/last-frame OR reference inputs, not both.

ModeInputsDescription
Text-to-Videoprompt onlyGenerate video from text description
Image-to-Videoprompt + imageAnimate a still image (first frame)
First+Last Frameprompt + image + end_imageControl start and end frames
Multimodal Referenceprompt + reference_images/reference_videos/reference_audiosGuide generation with reference material

Examples

Text-to-Video with Audio

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "ocean waves crashing on rocks during a storm, dramatic cinematic shot",
  "generate_audio": true,
  "duration": 10,
  "ratio": "16:9"
}'

Fast Mode (Cheaper)

belt app run bytedance/seedance-2-0-fast --input '{
  "prompt": "a butterfly landing on a flower in slow motion",
  "generate_audio": true
}'

Image-to-Video

Animate a still image into a video:

belt app run bytedance/seedance-2-0 --input '{
  "image": "https://your-image.jpg",
  "prompt": "gentle camera movement, leaves rustling in the wind",
  "generate_audio": true
}'

Image-to-Video with Start and End Frames

belt app run bytedance/seedance-2-0 --input '{
  "image": "https://start-frame.jpg",
  "end_image": "https://end-frame.jpg",
  "prompt": "smooth transition between scenes",
  "generate_audio": true
}'

Multi-Image Reference

Use multiple reference images to guide character appearance, outfits, and scene elements:

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "The girl from Image 1 wearing the outfit from Image 2 walks through the cafe from Image 3",
  "reference_images": [
    "https://character-portrait.jpg",
    "https://outfit-reference.jpg",
    "https://cafe-scene.jpg"
  ],
  "generate_audio": true,
  "duration": 8
}'

Video Editing (Replace Elements)

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "Replace the perfume in Video 1 with the face cream from Image 1, preserving all original motions and camera work",
  "reference_images": ["https://face-cream.jpg"],
  "reference_videos": ["https://original-video.mp4"],
  "generate_audio": true
}'

Video Extension (Stitch Clips)

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "Video 1 transitions smoothly into Video 2, then the camera enters the painting from Video 3",
  "reference_videos": [
    "https://clip1.mp4",
    "https://clip2.mp4",
    "https://clip3.mp4"
  ],
  "generate_audio": true,
  "duration": 8
}'

Reference with Audio

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "The musician from Image 1 performs the song from Audio 1, voice style referenced from Audio 1",
  "reference_images": ["https://musician.jpg"],
  "reference_audios": ["https://music.mp3"],
  "generate_audio": true
}'

Studio Mode (Portrait Consistency)

Studio variants upload images to BytePlus's private asset library for enhanced face/character consistency:

belt app run bytedance/seedance-2-0-studio --input '{
  "prompt": "The person in Image 1 smiles at the camera, golden hour lighting, cinematic",
  "reference_images": ["https://portrait.jpg"],
  "safety_identifier": "user-abc123",
  "generate_audio": true
}'

Product Ad with Multiple References

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "First-person POV product ad. Opening frame is Image 1, hand picks up the product. Camera pushes into close-up showing details. Use the camera movement style from Video 1. Background music from Audio 1.",
  "reference_images": ["https://product-hero.jpg", "https://product-detail.jpg"],
  "reference_videos": ["https://camera-style.mp4"],
  "reference_audios": ["https://bgm.mp3"],
  "generate_audio": true,
  "ratio": "9:16",
  "duration": 11
}'

Prompt Guide

Reference assets in your prompt using type + index: Image 1, Image 2, Video 1, Audio 1. The index is the position within that type in the arrays you provide. Do NOT use asset IDs in prompts.

Multimodal reference formula:

  • Image reference: "Refer to the [subject] from [Image N] to generate [scene], keeping [subject] consistent"
  • Video reference: "Refer to the [camera movement/action] from [Video N]"
  • Audio reference: "[Character] says: [dialogue], voice style referenced from [Audio N]"

Video editing formula:

  • Add: "At [timing] of [Video N], add [element]"
  • Remove: "Remove [element] from [Video N], keeping the rest unchanged"
  • Modify: "Replace [element] in [Video N] with [new element]"

Video extension formula:

  • Forward: "Generate content after [Video N]: [description]"
  • Backward: "Extend the opening of [Video N]: [description]"
  • Stitch: "[Video 1] + [transition] + followed by [Video 2]"

Parameters

ParameterTypeDefaultDescription
promptstringrequiredText description of the video
generate_audiobooleantrueGenerate synchronized audio
durationinteger5Duration in seconds (4-15), or -1 for auto
ratioenumadaptive21:9, 16:9, 4:3, 1:1, 3:4, 9:16, or adaptive
resolutionenum720p480p, 720p, 1080p (Fast: 480p, 720p only)
seedinteger-1Seed for reproducibility (-1 for random)
watermarkbooleanfalseAdd watermark to output
safety_identifierstring-Unique end-user identifier for safety policy (max 64 chars, hash of user ID recommended)
imagefile-First-frame image (mutually exclusive with reference inputs)
end_imagefile-Last-frame image (requires image)
reference_imagesfile[]-Reference images, up to 9 (mutually exclusive with image/end_image)
reference_videosfile[]-Reference videos, up to 3. Max 15s each, total max 15s. mp4/mov
reference_audiosfile[]-Reference audios, up to 3. Max 15s each, total max 15s. wav/mp3. Requires at least one image or video

Pricing

ModelPricing
Seedance 2.0$4.30-$7.70/M tokens (varies by resolution and input type)
Seedance 2.0 Fast$3.30-$5.60/M tokens

Token formula: (width x height x fps x duration) / 1024

Search Seedance Apps

belt app search "seedance"

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# All video generation models
npx skills add inference-sh/skills@ai-video-generation

# Google Veo
npx skills add inference-sh/skills@google-veo

# Image generation (for image-to-video)
npx skills add inference-sh/skills@ai-image-generation

# AI avatars & lipsync
npx skills add inference-sh/skills@ai-avatar-video

Browse all video apps: belt app list --category video

Documentation

來自 101-skills 的更多技能

ai-avatar-video
101-skills
透過 inference.sh CLI 建立 AI 虛擬角色與說話頭像影片。推薦:P-Video-Avatar(最快、最便宜、內建 TTS)。另可選:OmniHuman、Fabric、PixVerse。音訊:Inworld TTS-2(100 多種語言、角色情感引導)、ElevenLabs、Kokoro。功能:音訊驅動虛擬角色、文字轉虛擬角色、唇形同步影片、說話頭像生成、虛擬主持人、UGC 內容。適用於:AI 主持人、解說影片、虛擬網紅、配音、行銷影片、UGC 廣告、遊戲虛擬角色等。
creativevideomedia
ai-video-generation
101-skills
透過 inference.sh CLI 使用 Google Veo、Seedance 2.0、HappyHorse、Wan、Grok 及 40 多種模型生成 AI 影片。模型包括:Veo 3.1、Veo 3、Seedance 2.0、HappyHorse 1.0、Wan 2.5、Grok Imagine Video、OmniHuman、Fabric、HunyuanVideo。功能涵蓋:文字轉影片、圖片轉影片、參考素材轉影片、影片編輯、唇形同步、虛擬角色動畫、影片畫質提升、擬音音效。適用於:社群媒體影片、行銷內容、解說影片、產品展示、AI 虛擬角色。觸發條件:影片生成、AI 影片...
creativevideomedia
ai-image-generation
101-skills
透過 inference.sh CLI 使用 GPT-Image-2、FLUX、Gemini、Grok、Seedream、Reve 及 50 多種模型生成 AI 圖像。模型包括:GPT-Image-2、FLUX Dev LoRA、FLUX.2 Klein LoRA、Gemini 3 Pro Image、Grok Imagine、Seedream 4.5、Reve、ImagineArt。功能涵蓋:文字轉圖像、圖像轉圖像、修補、LoRA、圖像編輯、放大、文字渲染。適用於:AI 藝術、產品模型、概念藝術、社交媒體圖形、行銷視覺、插圖。觸發詞:flux、圖像生成、AI 圖像、文字轉...
creativemediaimage
remotion-render
101-skills
透過 inference.sh 從 React/Remotion 元件程式碼渲染影片。傳入 TSX 程式碼,取得 MP4。支援所有 Remotion API:useCurrentFrame、useVideoConfig、spring、interpolate、AbsoluteFill、Sequence。可設定解析度、FPS、時長、編碼器。用途:程式化影片生成、動畫圖形、動態設計、資料驅動影片、React 動畫轉影片。觸發詞:remotion、從程式碼渲染影片、tsx 轉影片、react 影片、程式化影片、remotion 渲染、程式碼轉影片、動畫...
videocreativedevelopment
web-search
101-skills
透過 inference.sh CLI 使用 Tavily 和 Exa 進行網路搜尋與內容提取。應用:Tavily 搜尋、Tavily 提取、Exa 搜尋、Exa 回答、Exa 提取。功能:AI 驅動搜尋、內容提取、直接回答、研究。用途:研究、RAG 管線、事實查核、內容彙整、代理。觸發詞:網路搜尋、tavily、exa、搜尋 API、內容提取、研究、網際網路搜尋、AI 搜尋、搜尋助手、網頁抓取、rag、perplexity 替代方案
researchweb-scrapingapi
agent-tools
101-skills
透過 inference.sh CLI 執行 AI 應用程式 - 圖片生成、影片創作、大型語言模型、搜尋、3D、Twitter 自動化。模型:FLUX、Veo、Gemini、Grok、Claude、Seedance、OmniHuman、Tavily、Exa、OpenRouter 等。適用於執行 AI 應用程式、生成圖片/影片、呼叫大型語言模型、網路搜尋或自動化 Twitter 時使用。觸發詞:inference.sh、infsh、ai model、run ai、serverless ai、ai api、flux、veo、claude api、image generation、video generation、openrouter、tavily、exa search、twitter api、grok
infsh-cli
101-skills
透過 inference.sh CLI 執行 AI 應用程式 — 圖片生成、影片創作、大型語言模型、搜尋、3D、Twitter 自動化。模型包括:FLUX、Veo、Gemini、Grok、Claude、Seedance、OmniHuman、Tavily、Exa、OpenRouter 等更多。適用於執行 AI 應用程式、生成圖片/影片、呼叫大型語言模型、網路搜尋或自動化 Twitter 時觸發。觸發關鍵字:inference.sh、infsh、ai model、run ai、serverless ai、ai api、flux、veo、claude api、image generation、video generation、openrouter、tavily、exa search、twitter api、grok
landing-page-design
101-skills
著陸頁轉換優化,包含佈局規則、英雄區塊設計與CTA心理學。涵蓋首屏公式、社會證明放置、行動設計與F型閱讀模式。適用於:新創著陸頁、產品頁面、SaaS行銷、轉換優化。觸發詞:著陸頁、英雄區塊、首屏、轉換優化、著陸頁設計、CTA按鈕、英雄圖片、著陸頁佈局、SaaS著陸頁、產品頁面設計、轉換率、著陸頁...
designmarketingcreative