ai-video-generation

Erstelle KI-Videos mit Google Veo, Seedance 2.0, HappyHorse, Wan, Grok und 40+ Modellen über die inference.sh CLI. Modelle: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. Fähigkeiten: Text-zu-Video, Bild-zu-Video, Referenz-zu-Video, Videobearbeitung, Lippen-Sync, Avatar-Animation, Video-Upscaling, Foley-Sound. Verwendung für: Social-Media-Videos, Marketinginhalte, Erklärvideos, Produktdemos, KI-Avatare. Auslöser: Videoerstellung, KI-Video,...

npx skills add https://github.com/qu-skills/skills --skill ai-video-generation

Install the belt CLI skill: npx skills add belt-sh/cli

AI Video Generation

Generate videos with 40+ AI models via inference.sh CLI.

AI Video Generation

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate a video with Veo
belt app run google/veo-3-1-fast --input '{"prompt": "drone shot flying over a forest"}'

Available Models

Text-to-Video

ModelApp IDBest For
Veo 3.1 Fastgoogle/veo-3-1-fastFast, with optional audio
Veo 3.1google/veo-3-1Best quality, frame interpolation
Veo 3google/veo-3High quality with audio
Veo 3 Fastgoogle/veo-3-fastFast with audio
Veo 2google/veo-2Realistic videos
P-Videopruna/p-videoFast, economical, with audio support
WAN-T2Vpruna/wan-t2vEconomical 480p/720p
Grok Videoxai/grok-imagine-videoxAI, configurable duration
Seedance 2.0bytedance/seedance-2-0Text/image/ref-to-video with sync audio, up to 1080p
Seedance 2.0 Fastbytedance/seedance-2-0-fastFast variant, same capabilities
HappyHorse T2Valibaba/happyhorse-1-0-t2vPhysically realistic, up to 15s

Image-to-Video

ModelApp IDBest For
Wan 2.5falai/wan-2-5Animate any image
Wan 2.5 I2Vfalai/wan-2-5-i2vHigh quality i2v
WAN-I2Vpruna/wan-i2vEconomical 480p/720p
P-Videopruna/p-videoFast i2v with audio
Seedance 2.0bytedance/seedance-2-0Animate images with sync audio, up to 1080p
Seedance 2.0 Fastbytedance/seedance-2-0-fastFast variant, same capabilities
HappyHorse I2Valibaba/happyhorse-1-0-i2vAnimate images, up to 1080P/15s
HappyHorse R2Valibaba/happyhorse-1-0-r2vCharacter-preserving from references

Avatar / Lipsync

ModelApp IDBest For
OmniHuman 1.5bytedance/omnihuman-1-5Multi-character
OmniHuman 1.0bytedance/omnihuman-1-0Single character
Fabric 1.0falai/fabric-1-0Image talks with lipsync
PixVerse Lipsyncfalai/pixverse-lipsyncRealistic lipsync

Video Editing

ModelApp IDBest For
HappyHorse Editalibaba/happyhorse-1-0-video-editNatural language video editing

Utilities

ToolApp IDDescription
HunyuanVideo Foleyinfsh/hunyuanvideo-foleyAdd sound effects to video
Topaz Upscalerfalai/topaz-video-upscalerUpscale video quality
Media Mergerinfsh/media-mergerMerge videos with transitions

Browse All Video Apps

belt app store --category video

Examples

Text-to-Video with Veo

belt app run google/veo-3-1-fast --input '{
  "prompt": "A timelapse of a flower blooming in a garden"
}'

Grok Video

belt app run xai/grok-imagine-video --input '{
  "prompt": "Waves crashing on a beach at sunset",
  "duration": 5
}'

Image-to-Video with Wan 2.5

belt app run falai/wan-2-5 --input '{
  "image_url": "https://your-image.jpg"
}'

AI Avatar / Talking Head

belt app run bytedance/omnihuman-1-5 --input '{
  "image_url": "https://portrait.jpg",
  "audio_url": "https://speech.mp3"
}'

Fabric Lipsync

belt app run falai/fabric-1-0 --input '{
  "image_url": "https://face.jpg",
  "audio_url": "https://audio.mp3"
}'

Seedance 2.0 Text-to-Video with Audio

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "a jazz band performing in a dimly lit club",
  "generate_audio": true,
  "duration": 10
}'

Seedance 2.0 Image-to-Video

belt app run bytedance/seedance-2-0 --input '{
  "image": "https://your-image.jpg",
  "prompt": "gentle camera movement, leaves rustling in the wind",
  "generate_audio": true
}'

Seedance 2.0 Reference-to-Video

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "A person who looks like the reference walking through a garden",
  "reference_image": "https://portrait.jpg",
  "generate_audio": true
}'

HappyHorse Text-to-Video

belt app run alibaba/happyhorse-1-0-t2v --input '{
  "prompt": "a golden retriever running through autumn leaves, slow motion",
  "duration": 10,
  "resolution": "1080P"
}'

HappyHorse Video Editing

belt app run alibaba/happyhorse-1-0-video-edit --input '{
  "video": "https://your-video.mp4",
  "prompt": "change the background to a snowy mountain landscape"
}'

PixVerse Lipsync

belt app run falai/pixverse-lipsync --input '{
  "image_url": "https://portrait.jpg",
  "audio_url": "https://speech.mp3"
}'

Video Upscaling

belt app run falai/topaz-video-upscaler --input '{"video_url": "https://..."}'

Add Sound Effects (Foley)

belt app run infsh/hunyuanvideo-foley --input '{
  "video_url": "https://silent-video.mp4",
  "prompt": "footsteps on gravel, birds chirping"
}'

Merge Videos

belt app run infsh/media-merger --input '{
  "videos": ["https://clip1.mp4", "https://clip2.mp4"],
  "transition": "fade"
}'

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# Pruna P-Video (fast & economical)
npx skills add inference-sh/skills@p-video

# Google Veo specific
npx skills add inference-sh/skills@google-veo

# Seedance 2.0
npx skills add inference-sh/skills@seedance

# HappyHorse 1.0
npx skills add inference-sh/skills@happyhorse

# AI avatars & lipsync
npx skills add inference-sh/skills@ai-avatar-video

# Text-to-speech (for video narration)
npx skills add inference-sh/skills@text-to-speech

# Image generation (for image-to-video)
npx skills add inference-sh/skills@ai-image-generation

# Twitter (post videos)
npx skills add inference-sh/skills@twitter-automation

Browse all apps: belt app store

Documentation

Mehr Skills von qu-skills

remotion-render
qu-skills
Videos aus React/Remotion-Komponentencode über inference.sh rendern. TSX-Code übergeben, MP4 erhalten. Unterstützt alle Remotion-APIs: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence. Konfigurierbare Auflösung, FPS, Dauer, Codec. Verwendung für: programmatische Videogenerierung, animierte Grafiken, Bewegungsdesign, datengesteuerte Videos, React-Animationen zu Video. Auslöser: remotion, Video aus Code rendern, tsx zu Video, React-Video, programmatisches Video, Remotion-Render, Code zu Video, animiert...
developmentvideocreative
ai-image-generation
qu-skills
Generieren Sie KI-Bilder mit GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve und über 50 Modellen über die inference.sh CLI. Modelle: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Fähigkeiten: Text-zu-Bild, Bild-zu-Bild, Inpainting, LoRA, Bildbearbeitung, Hochskalierung, Textrendering. Verwendung für: KI-Kunst, Produkt-Mockups, Konzeptkunst, Social-Media-Grafiken, Marketing-Visuals, Illustrationen. Auslöser: flux, Bildgenerierung, KI-Bild, Text zu...
creativemediaimage
ai-avatar-video
qu-skills
Erstelle KI-Avatar- und Talking-Head-Videos über die inference.sh CLI. Empfohlen: P-Video-Avatar (schnellste, günstigste, integrierte TTS). Auch: OmniHuman, Fabric, PixVerse. Audio: Inworld TTS-2 (100+ Sprachen, Emotionssteuerung für Charaktere), ElevenLabs, Kokoro. Fähigkeiten: audiogesteuerte Avatare, Text-zu-Avatar, Lipsync-Videos, Talking-Head-Generierung, virtuelle Präsentatoren, UGC-Inhalte. Verwenden für: KI-Präsentatoren, Erklärvideos, virtuelle Influencer, Synchronisation, Marketingvideos, UGC-Anzeigen, Gaming-Avatare,...
videocreativemedia
twitter-automation
qu-skills
Automatisiere Twitter/X mit Posting, Engagement und Benutzerverwaltung über die inference.sh CLI. Apps: x/post-tweet, x/post-create (mit Medien), x/post-like, x/post-retweet, x/dm-send, x/user-follow. Fähigkeiten: Tweets posten, Inhalte planen, Beiträge liken, retweeten, DMs senden, Benutzern folgen, Profile abrufen. Verwendung für: Social-Media-Automatisierung, Inhaltsplanung, Engagement-Bots, Zielgruppenwachstum, X API. Auslöser: twitter api, x api, tweet automation, post to twitter, twitter bot, social media automation, x...
api
agent-browser
qu-skills
Browser-Automatisierung für KI-Agenten über inference.sh. Navigieren Sie auf Webseiten, interagieren Sie mit Elementen über @e-Referenzen, machen Sie Screenshots, nehmen Sie Video auf. Fähigkeiten: Web Scraping, Formularausfüllen, Klicken, Tippen, Drag & Drop, Datei-Upload, JavaScript-Ausführung. Verwendung für: Web-Automatisierung, Datenextraktion, Testen, Agenten-Browsing, Recherche. Auslöser: Browser, Web-Automatisierung, Scrapen, Navigieren, Klicken, Formular ausfüllen, Screenshot, Web durchsuchen, Playwright, Headless-Browser, Web-Agent, Internet surfen, Video aufnehmen
browser-automationweb-scrapingtesting
web-search
qu-skills
Websuche und Inhaltsextraktion mit Tavily und Exa über die inference.sh CLI. Apps: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract. Fähigkeiten: KI-gestützte Suche, Inhaltsextraktion, direkte Antworten, Recherche. Verwendung für: Recherche, RAG-Pipelines, Faktenprüfung, Inhaltsaggregation, Agenten. Auslöser: Websuche, tavily, exa, search api, Inhaltsextraktion, Recherche, Internetsuche, KI-Suche, Suchassistent, Web Scraping, rag, Perplexity-Alternative
researchweb-scrapingapi
agent-tools
qu-skills
Führe 250+ KI-Apps über die inference.sh CLI aus – Bildgenerierung, Videocrstellung, LLMs, Suche, 3D, Twitter-Automatisierung. Modelle: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter und viele mehr. Verwenden beim Ausführen von KI-Apps, Generieren von Bildern/Videos, Aufrufen von LLMs, Websuche oder Automatisieren von Twitter. Auslöser: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
developmentapicreative
python-executor
qu-skills
We need to translate the given English text into German, preserving the name "python-executor" only if it appears in the source text. The source text does not contain "python-executor" explicitly; it only appears in the instruction as the name to preserve, but not in the <text> block. So we should not include it. The translation should be accurate, keep technical terms, URLs, numbers, and product names. The text describes executing Python code in a sandboxed environment via inference.sh, lists pre-installed libraries, use cases, and triggers. Translate naturally into German. Key points: "safe sandboxed environment" -> "sicheren Sandkasten-Umgebung" or "sicheren isolierten Umgebung"? "Sandboxed" is often "Sandkasten" or "isoliert". "Pre-installed" -> "Vorinstalliert". "100+ more libraries" -> "über 100 weitere Bibliotheken". "Use for" -> "Verwendung für". "Triggers" -> "Auslöser" or "Trigger
developmentdata-analysisweb-scraping