image-to-video

Leitfaden zur Umwandlung von Standbildern in Videos: Modellauswahl, Bewegungs-Prompting und Kamerabewegung. Behandelt Wan 2.5 i2v, Seedance, Fabric, Grok Video mit Angabe, wann welches verwendet werden sollte. Verwendung für: Animieren von Bildern, Erstellen von Videos aus Standbildern, Hinzufügen von Bewegung, Produktanimationen. Auslöser: Bild zu Video, i2v, Bild animieren, Standbild zu Video, Bewegung zu Bild hinzufügen, Bildanimation, Foto zu Video, Standbild animieren, wan i2v, image2video, Bild zum Leben erwecken, Foto animieren, Bewegung aus Bild

npx skills add https://github.com/qu-skills/skills --skill image-to-video

Install the belt CLI skill: npx skills add belt-sh/cli

Image to Video

Convert still images to animated videos via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate a still image
belt app run falai/flux-dev-lora --input '{
  "prompt": "serene mountain lake at sunset, snow-capped peaks reflected in still water, golden hour light, landscape photography",
  "width": 1248,
  "height": 832
}'

# Animate it
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "gentle ripples on the lake surface, clouds slowly drifting, warm light shifting, birds flying in the distance",
  "image": "path/to/lake-image.png"
}'

Model Selection

ModelApp IDBest ForMotion Style
Wan 2.5 i2vfalai/wan-2-5-i2vRealistic motion, natural movementPhotorealistic, subtle
WAN-I2V (Pruna)pruna/wan-i2vEconomical, fast, 480p/720pNatural, efficient
Seedance 2.0bytedance/seedance-2-0Up to 1080p, sync audio, all input typesVersatile, high quality
Seedance 2.0 Fastbytedance/seedance-2-0-fastFast variant, same capabilitiesVersatile, fast
Fabric 1.0falai/fabric-1-0Cloth, fabric, liquid, flowing materialsPhysics-based flow
Grok Imagine Videoxai/grok-imagine-videoGeneral animation, text-guidedVersatile

When to Use Each

ScenarioBest ModelWhy
Landscape with water/cloudsWan 2.5 i2vBest at natural, realistic motion
Portrait with subtle expressionWan 2.5 i2vMaintains face fidelity
Product with fabric/clothFabric 1.0Specialized in material physics
Flag waving, curtain flowingFabric 1.0Cloth simulation
Illustrated/artistic imageSeedance 2.0Matches stylized content
General "bring to life"Seedance 2.0Good all-rounder, up to 1080p
Quick test/iterationSeedance 2.0 FastFaster generation

Motion Types

Camera Movement

MovementPrompt KeywordEffect
Push in / Dolly forward"slow dolly forward", "camera pushes in"Increasing intimacy/focus
Pull out / Dolly back"camera pulls back", "slow zoom out"Reveal, context
Pan left/right"camera pans slowly to the right"Scanning, following
Tilt up/down"camera tilts upward"Revealing height
Orbit"camera orbits around the subject"3D exploration
Crane up"camera rises upward"Grand reveal
Static(no camera movement prompt)Subject motion only

Subject Motion

TypePrompt Examples
Natural elements"water rippling", "clouds drifting", "leaves rustling in breeze"
Hair/clothing"hair blowing gently in wind", "dress fabric flowing"
Atmospheric"fog slowly rolling", "dust particles floating in light beams"
Character"person slowly turns to camera", "subtle breathing motion"
Mechanical"gears turning", "clock hands moving"
Liquid"coffee steam rising", "paint dripping", "water pouring"

Prompting Best Practices

The Golden Rule: Subtle > Dramatic

AI video models produce better results with gentle, subtle motion than dramatic action. Requesting too much movement causes distortion and artifacts.

❌ "person running and jumping over obstacles while the camera spins"
✅ "person slowly walking forward, gentle breeze, camera follows alongside"

❌ "explosion with debris flying everywhere"
✅ "candle flame flickering gently, warm ambient light shifting"

❌ "fast zoom into the eyes with dramatic camera shake"
✅ "slow dolly forward toward the subject, subtle focus shift"

Prompt Structure

[Camera movement] + [Subject motion] + [Atmospheric effects] + [Mood/pace]

Examples by Scenario

# Landscape animation
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "gentle camera pan right, water reflecting moving clouds, trees swaying slightly in breeze, warm golden light, peaceful and slow",
  "image": "landscape.png"
}'

# Portrait animation
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "subtle breathing motion, slight head turn, natural eye blink, hair moving gently, soft ambient lighting shifts",
  "image": "portrait.png"
}'

# Product shot animation
belt app run bytedance/seedance-2-0 --input '{
  "prompt": "slow 360 degree orbit around the product, gentle spotlight movement, subtle reflections shifting, premium product showcase, smooth motion",
  "image": "product.png",
  "generate_audio": true
}'

# Fabric/cloth animation
belt app run falai/fabric-1-0 --input '{
  "prompt": "fabric flowing and rippling in gentle wind, natural cloth physics, soft movement",
  "image": "fabric-scene.png"
}'

# Architectural visualization
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "slow dolly forward through the entrance, slight camera tilt upward, ambient light filtering through windows, dust particles in light beams",
  "image": "building-interior.png"
}'

Duration Guidelines

DurationQualityUse For
2-3 secondsHighest qualityGIFs, looping backgrounds, cinemagraphs
4-5 secondsHigh qualitySocial media posts, product reveals
6-8 secondsGood qualityShort clips, transitions
10+ secondsQuality degradesAvoid unless stitching shorter clips

Extending Duration

For longer videos, generate multiple short clips and stitch:

# Generate 3 clips from the same image with progressive motion
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "slow pan left, gentle water motion",
  "image": "scene.png"
}' --no-wait

belt app run falai/wan-2-5-i2v --input '{
  "prompt": "continuing pan, clouds shifting, light changing",
  "image": "scene.png"
}' --no-wait

# Stitch together
belt app run infsh/media-merger --input '{
  "media": ["clip1.mp4", "clip2.mp4"]
}'

The Full Workflow

Still-to-Final-Video Pipeline

# 1. Generate source image (best quality)
belt app run bytedance/seedream-4-5 --input '{
  "prompt": "cinematic landscape, misty mountains at dawn, lake in foreground, dramatic clouds, golden hour, 4K quality, professional photography",
  "size": "2K"
}'

# 2. Animate the image
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "gentle mist rolling through the valley, lake surface rippling, clouds slowly moving, birds in distance, warm light shifting",
  "image": "landscape.png"
}'

# 3. Upscale video if needed
belt app run falai/topaz-video-upscaler --input '{
  "video": "animated-landscape.mp4"
}'

# 4. Add ambient audio
belt app run infsh/hunyuanvideo-foley --input '{
  "video": "animated-landscape.mp4",
  "prompt": "gentle nature ambience, distant birds, soft wind, water lapping"
}'

# 5. Merge video with audio
belt app run infsh/video-audio-merger --input '{
  "video": "upscaled-landscape.mp4",
  "audio": "ambient-audio.mp3"
}'

Cinemagraph Effect

A cinemagraph is a still photo where only one element moves (e.g., waterfall moving in an otherwise frozen scene). To achieve this:

  1. Generate the still image with the motion element clearly defined
  2. Prompt for motion only in that specific element
  3. Keep to 2-4 seconds for seamless looping
belt app run falai/wan-2-5-i2v --input '{
  "prompt": "only the waterfall is moving, everything else remains perfectly still, water cascading smoothly, rest of scene frozen",
  "image": "waterfall-scene.png"
}'

Common Mistakes

MistakeProblemFix
Too much motion requestedDistortion, artifacts, warpingSubtle > dramatic, always
Wrong model for content typePoor resultsUse selection guide above
Clips too long (10s+)Quality degrades significantlyKeep to 3-5 seconds, stitch if needed
No camera movement specifiedRandom/unpredictable motionAlways specify camera behavior
Conflicting motion directionsChaotic, unnaturalOne primary motion direction
Low-res source imageLow-res video outputStart with highest quality source
Complex action scenesModels can't handleKeep motion simple and natural

Related Skills

npx skills add inference-sh/skills@ai-video-generation
npx skills add inference-sh/skills@ai-image-generation
npx skills add inference-sh/skills@p-video
npx skills add inference-sh/skills@video-prompting-guide
npx skills add inference-sh/skills@prompt-engineering

Browse all apps: belt app store

Mehr Skills von qu-skills

ai-video-generation
qu-skills
Erstelle KI-Videos mit Google Veo, Seedance 2.0, HappyHorse, Wan, Grok und 40+ Modellen über die inference.sh CLI. Modelle: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. Fähigkeiten: Text-zu-Video, Bild-zu-Video, Referenz-zu-Video, Videobearbeitung, Lippen-Sync, Avatar-Animation, Video-Upscaling, Foley-Sound. Verwendung für: Social-Media-Videos, Marketinginhalte, Erklärvideos, Produktdemos, KI-Avatare. Auslöser: Videoerstellung, KI-Video,...
videocreativemedia
remotion-render
qu-skills
Videos aus React/Remotion-Komponentencode über inference.sh rendern. TSX-Code übergeben, MP4 erhalten. Unterstützt alle Remotion-APIs: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence. Konfigurierbare Auflösung, FPS, Dauer, Codec. Verwendung für: programmatische Videogenerierung, animierte Grafiken, Bewegungsdesign, datengesteuerte Videos, React-Animationen zu Video. Auslöser: remotion, Video aus Code rendern, tsx zu Video, React-Video, programmatisches Video, Remotion-Render, Code zu Video, animiert...
developmentvideocreative
ai-image-generation
qu-skills
Generieren Sie KI-Bilder mit GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve und über 50 Modellen über die inference.sh CLI. Modelle: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Fähigkeiten: Text-zu-Bild, Bild-zu-Bild, Inpainting, LoRA, Bildbearbeitung, Hochskalierung, Textrendering. Verwendung für: KI-Kunst, Produkt-Mockups, Konzeptkunst, Social-Media-Grafiken, Marketing-Visuals, Illustrationen. Auslöser: flux, Bildgenerierung, KI-Bild, Text zu...
creativemediaimage
ai-avatar-video
qu-skills
Erstelle KI-Avatar- und Talking-Head-Videos über die inference.sh CLI. Empfohlen: P-Video-Avatar (schnellste, günstigste, integrierte TTS). Auch: OmniHuman, Fabric, PixVerse. Audio: Inworld TTS-2 (100+ Sprachen, Emotionssteuerung für Charaktere), ElevenLabs, Kokoro. Fähigkeiten: audiogesteuerte Avatare, Text-zu-Avatar, Lipsync-Videos, Talking-Head-Generierung, virtuelle Präsentatoren, UGC-Inhalte. Verwenden für: KI-Präsentatoren, Erklärvideos, virtuelle Influencer, Synchronisation, Marketingvideos, UGC-Anzeigen, Gaming-Avatare,...
videocreativemedia
twitter-automation
qu-skills
Automatisiere Twitter/X mit Posting, Engagement und Benutzerverwaltung über die inference.sh CLI. Apps: x/post-tweet, x/post-create (mit Medien), x/post-like, x/post-retweet, x/dm-send, x/user-follow. Fähigkeiten: Tweets posten, Inhalte planen, Beiträge liken, retweeten, DMs senden, Benutzern folgen, Profile abrufen. Verwendung für: Social-Media-Automatisierung, Inhaltsplanung, Engagement-Bots, Zielgruppenwachstum, X API. Auslöser: twitter api, x api, tweet automation, post to twitter, twitter bot, social media automation, x...
api
agent-browser
qu-skills
Browser-Automatisierung für KI-Agenten über inference.sh. Navigieren Sie auf Webseiten, interagieren Sie mit Elementen über @e-Referenzen, machen Sie Screenshots, nehmen Sie Video auf. Fähigkeiten: Web Scraping, Formularausfüllen, Klicken, Tippen, Drag & Drop, Datei-Upload, JavaScript-Ausführung. Verwendung für: Web-Automatisierung, Datenextraktion, Testen, Agenten-Browsing, Recherche. Auslöser: Browser, Web-Automatisierung, Scrapen, Navigieren, Klicken, Formular ausfüllen, Screenshot, Web durchsuchen, Playwright, Headless-Browser, Web-Agent, Internet surfen, Video aufnehmen
browser-automationweb-scrapingtesting
web-search
qu-skills
Websuche und Inhaltsextraktion mit Tavily und Exa über die inference.sh CLI. Apps: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract. Fähigkeiten: KI-gestützte Suche, Inhaltsextraktion, direkte Antworten, Recherche. Verwendung für: Recherche, RAG-Pipelines, Faktenprüfung, Inhaltsaggregation, Agenten. Auslöser: Websuche, tavily, exa, search api, Inhaltsextraktion, Recherche, Internetsuche, KI-Suche, Suchassistent, Web Scraping, rag, Perplexity-Alternative
researchweb-scrapingapi
agent-tools
qu-skills
Führe 250+ KI-Apps über die inference.sh CLI aus – Bildgenerierung, Videocrstellung, LLMs, Suche, 3D, Twitter-Automatisierung. Modelle: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter und viele mehr. Verwenden beim Ausführen von KI-Apps, Generieren von Bildern/Videos, Aufrufen von LLMs, Websuche oder Automatisieren von Twitter. Auslöser: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
developmentapicreative