seedance

Генерация видео с помощью ByteDance Seedance 2.0 через CLI inference.sh. Унифицированная модель для text-to-video, image-to-video и reference-to-video с синхронизированным аудио, до 1080p, длительностью 4–15 секунд. Варианты Pro и Fast. Студийные варианты с приватной библиотекой ассетов для согласованности портретов. Используется для: видео в соцсетях, музыкальных клипов, демонстраций продуктов, анимационного контента, AI-видео со звуком. Триггеры: seedance, seedance 2, bytedance video, seedance t2v, seedance i2v, seedance r2v, video with audio,...

npx skills add https://github.com/qu-skills/skills --skill seedance

Install the belt CLI skill: npx skills add belt-sh/cli

Seedance 2.0 Video Generation

Generate videos with synchronized audio using ByteDance's Seedance 2.0 via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "a jazz band performing in a dimly lit club",
  "generate_audio": true
}'

Models

ModelApp IDBest For
Seedance 2.0bytedance/seedance-2-0Best quality, up to 1080p
Seedance 2.0 Fastbytedance/seedance-2-0-fastFaster generation, up to 720p
Seedance 2.0 Studiobytedance/seedance-2-0-studioQuality + private asset library for portrait consistency
Seedance 2.0 Studio Fastbytedance/seedance-2-0-studio-fastFast + private asset library for portrait consistency

All models support text-to-video, image-to-video, multimodal reference-to-video, and synchronized audio generation. Studio variants automatically upload reference images to the BytePlus private virtual portrait library for enhanced character consistency - particularly useful for faces and branded characters.

Modes

The model determines the generation mode from your inputs. These modes are mutually exclusive - use either first-frame/last-frame OR reference inputs, not both.

ModeInputsDescription
Text-to-Videoprompt onlyGenerate video from text description
Image-to-Videoprompt + imageAnimate a still image (first frame)
First+Last Frameprompt + image + end_imageControl start and end frames
Multimodal Referenceprompt + reference_images/reference_videos/reference_audiosGuide generation with reference material

Examples

Text-to-Video with Audio

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "ocean waves crashing on rocks during a storm, dramatic cinematic shot",
  "generate_audio": true,
  "duration": 10,
  "ratio": "16:9"
}'

Fast Mode (Cheaper)

belt app run bytedance/seedance-2-0-fast --input '{
  "prompt": "a butterfly landing on a flower in slow motion",
  "generate_audio": true
}'

Image-to-Video

Animate a still image into a video:

belt app run bytedance/seedance-2-0 --input '{
  "image": "https://your-image.jpg",
  "prompt": "gentle camera movement, leaves rustling in the wind",
  "generate_audio": true
}'

Image-to-Video with Start and End Frames

belt app run bytedance/seedance-2-0 --input '{
  "image": "https://start-frame.jpg",
  "end_image": "https://end-frame.jpg",
  "prompt": "smooth transition between scenes",
  "generate_audio": true
}'

Multi-Image Reference

Use multiple reference images to guide character appearance, outfits, and scene elements:

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "The girl from Image 1 wearing the outfit from Image 2 walks through the cafe from Image 3",
  "reference_images": [
    "https://character-portrait.jpg",
    "https://outfit-reference.jpg",
    "https://cafe-scene.jpg"
  ],
  "generate_audio": true,
  "duration": 8
}'

Video Editing (Replace Elements)

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "Replace the perfume in Video 1 with the face cream from Image 1, preserving all original motions and camera work",
  "reference_images": ["https://face-cream.jpg"],
  "reference_videos": ["https://original-video.mp4"],
  "generate_audio": true
}'

Video Extension (Stitch Clips)

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "Video 1 transitions smoothly into Video 2, then the camera enters the painting from Video 3",
  "reference_videos": [
    "https://clip1.mp4",
    "https://clip2.mp4",
    "https://clip3.mp4"
  ],
  "generate_audio": true,
  "duration": 8
}'

Reference with Audio

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "The musician from Image 1 performs the song from Audio 1, voice style referenced from Audio 1",
  "reference_images": ["https://musician.jpg"],
  "reference_audios": ["https://music.mp3"],
  "generate_audio": true
}'

Studio Mode (Portrait Consistency)

Studio variants upload images to BytePlus's private asset library for enhanced face/character consistency:

belt app run bytedance/seedance-2-0-studio --input '{
  "prompt": "The person in Image 1 smiles at the camera, golden hour lighting, cinematic",
  "reference_images": ["https://portrait.jpg"],
  "safety_identifier": "user-abc123",
  "generate_audio": true
}'

Product Ad with Multiple References

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "First-person POV product ad. Opening frame is Image 1, hand picks up the product. Camera pushes into close-up showing details. Use the camera movement style from Video 1. Background music from Audio 1.",
  "reference_images": ["https://product-hero.jpg", "https://product-detail.jpg"],
  "reference_videos": ["https://camera-style.mp4"],
  "reference_audios": ["https://bgm.mp3"],
  "generate_audio": true,
  "ratio": "9:16",
  "duration": 11
}'

Prompt Guide

Reference assets in your prompt using type + index: Image 1, Image 2, Video 1, Audio 1. The index is the position within that type in the arrays you provide. Do NOT use asset IDs in prompts.

Multimodal reference formula:

  • Image reference: "Refer to the [subject] from [Image N] to generate [scene], keeping [subject] consistent"
  • Video reference: "Refer to the [camera movement/action] from [Video N]"
  • Audio reference: "[Character] says: [dialogue], voice style referenced from [Audio N]"

Video editing formula:

  • Add: "At [timing] of [Video N], add [element]"
  • Remove: "Remove [element] from [Video N], keeping the rest unchanged"
  • Modify: "Replace [element] in [Video N] with [new element]"

Video extension formula:

  • Forward: "Generate content after [Video N]: [description]"
  • Backward: "Extend the opening of [Video N]: [description]"
  • Stitch: "[Video 1] + [transition] + followed by [Video 2]"

Parameters

ParameterTypeDefaultDescription
promptstringrequiredText description of the video
generate_audiobooleantrueGenerate synchronized audio
durationinteger5Duration in seconds (4-15), or -1 for auto
ratioenumadaptive21:9, 16:9, 4:3, 1:1, 3:4, 9:16, or adaptive
resolutionenum720p480p, 720p, 1080p (Fast: 480p, 720p only)
seedinteger-1Seed for reproducibility (-1 for random)
watermarkbooleanfalseAdd watermark to output
safety_identifierstring-Unique end-user identifier for safety policy (max 64 chars, hash of user ID recommended)
imagefile-First-frame image (mutually exclusive with reference inputs)
end_imagefile-Last-frame image (requires image)
reference_imagesfile[]-Reference images, up to 9 (mutually exclusive with image/end_image)
reference_videosfile[]-Reference videos, up to 3. Max 15s each, total max 15s. mp4/mov
reference_audiosfile[]-Reference audios, up to 3. Max 15s each, total max 15s. wav/mp3. Requires at least one image or video

Pricing

ModelPricing
Seedance 2.0$4.30-$7.70/M tokens (varies by resolution and input type)
Seedance 2.0 Fast$3.30-$5.60/M tokens

Token formula: (width x height x fps x duration) / 1024

Search Seedance Apps

belt app store search "seedance"

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# All video generation models
npx skills add inference-sh/skills@ai-video-generation

# Google Veo
npx skills add inference-sh/skills@google-veo

# Image generation (for image-to-video)
npx skills add inference-sh/skills@ai-image-generation

# AI avatars & lipsync
npx skills add inference-sh/skills@ai-avatar-video

Browse all video apps: belt app store --category video

Documentation

Больше skills от qu-skills

ai-video-generation
qu-skills
Генерируйте AI-видео с помощью Google Veo, Seedance 2.0, HappyHorse, Wan, Grok и 40+ моделей через CLI inference.sh. Модели: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. Возможности: текст-в-видео, изображение-в-видео, референс-в-видео, редактирование видео, липсинк, анимация аватаров, апскейлинг видео, звук фоли. Используйте для: видео для соцсетей, маркетинговый контент, объясняющие видео, демонстрации продуктов, AI-аватары. Триггеры: генерация видео, ai video,...
videocreativemedia
remotion-render
qu-skills
Рендеринг видео из React/Remotion компонентного кода через inference.sh. Передайте TSX-код, получите MP4. Поддерживает все Remotion API: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence. Настраиваемое разрешение, FPS, длительность, кодек. Используется для: программной генерации видео, анимированной графики, моушн-дизайна, видео на основе данных, преобразования React-анимаций в видео. Триггеры: remotion, render video from code, tsx to video, react video, programmatic video, remotion render, code to video, animated...
developmentvideocreative
ai-image-generation
qu-skills
Генерируйте AI-изображения с помощью GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve и более 50 моделей через CLI inference.sh. Модели: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Возможности: текст-в-изображение, изображение-в-изображение, инпейнтинг, LoRA, редактирование изображений, апскейлинг, рендеринг текста. Используется для: AI-арт, макеты продуктов, концепт-арт, графика для соцсетей, маркетинговые визуалы, иллюстрации. Триггеры: flux, image generation, ai image, text to...
creativemediaimage
ai-avatar-video
qu-skills
Создавайте AI-аватары и видео с говорящими головами через CLI inference.sh. Рекомендуется: P-Video-Avatar (самый быстрый, дешёвый, встроенный TTS). Также: OmniHuman, Fabric, PixVerse. Аудио: Inworld TTS-2 (100+ языков, управление эмоциями для персонажей), ElevenLabs, Kokoro. Возможности: аватары, управляемые аудио, текст-в-аватар, видео с синхронизацией губ, генерация говорящих голов, виртуальные ведущие, UGC-контент. Используйте для: AI-ведущие, обучающие видео, виртуальные инфлюенсеры, дубляж, маркетинговые видео, UGC-реклама, игровые аватары,...
videocreativemedia
twitter-automation
qu-skills
Автоматизация Twitter/X с публикацией, взаимодействием и управлением пользователями через CLI inference.sh. Приложения: x/post-tweet, x/post-create (с медиа), x/post-like, x/post-retweet, x/dm-send, x/user-follow. Возможности: публикация твитов, планирование контента, лайки, ретвиты, отправка личных сообщений, подписка на пользователей, получение профилей. Используется для: автоматизации социальных сетей, планирования контента, ботов взаимодействия, роста аудитории, X API. Триггеры: twitter api, x api, tweet automation, post to twitter, twitter bot, social media automation, x...
api
agent-browser
qu-skills
Автоматизация браузера для AI-агентов через inference.sh. Навигация по веб-страницам, взаимодействие с элементами с помощью @e-ссылок, создание скриншотов, запись видео. Возможности: веб-скрапинг, заполнение форм, клики, ввод текста, перетаскивание, загрузка файлов, выполнение JavaScript. Используется для: веб-автоматизации, извлечения данных, тестирования, просмотра страниц агентом, исследований. Триггеры: браузер, веб-автоматизация, скрапинг, навигация, клик, заполнение формы, скриншот, просмотр веб-страниц, playwright, headless-браузер, веб-агент, серфинг в интернете, запись видео
browser-automationweb-scrapingtesting
web-search
qu-skills
Веб-поиск и извлечение контента с помощью Tavily и Exa через CLI inference.sh. Приложения: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract. Возможности: поиск на основе ИИ, извлечение контента, прямые ответы, исследования. Используется для: исследований, RAG-пайплайнов, проверки фактов, агрегации контента, агентов. Триггеры: веб-поиск, tavily, exa, search api, извлечение контента, исследование, интернет-поиск, ИИ-поиск, поисковый ассистент, веб-скрапинг, rag, альтернатива perplexity
researchweb-scrapingapi
agent-tools
qu-skills
Запускайте 250+ AI-приложений через CLI inference.sh — генерация изображений, создание видео, LLM, поиск, 3D, автоматизация Twitter. Модели: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter и многие другие. Используйте при запуске AI-приложений, генерации изображений/видео, вызове LLM, веб-поиске или автоматизации Twitter. Триггеры: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
developmentapicreative