ai-video-generation

द्वारा skills-101

inference.sh CLI के माध्यम से Google Veo, Seedance 2.0, HappyHorse, Wan, Grok और 40+ मॉडलों के साथ AI वीडियो बनाएं। मॉडल: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo। क्षमताएं: टेक्स्ट-टू-वीडियो, इमेज-टू-वीडियो, रेफरेंस-टू-वीडियो, वीडियो एडिटिंग, लिपसिंक, अवतार एनीमेशन, वीडियो अपस्केलिंग, फोले साउंड। उपयोग के लिए: सोशल मीडिया वीडियो, मार्केटिंग सामग्री, एक्सप्लेनर वीडियो, उत्पाद डेमो, AI अवतार। ट्रिगर: वीडियो जनरेशन, AI वीडियो,...

npx skills add https://github.com/skills-101/superpowers --skill ai-video-generation

Install the belt CLI skill: npx skills add belt-sh/cli

AI Video Generation

Generate videos with 40+ AI models via inference.sh CLI.

AI Video Generation

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate a video with Veo
belt app run google/veo-3-1-fast --input '{"prompt": "drone shot flying over a forest"}'

Available Models

Text-to-Video

ModelApp IDBest For
Veo 3.1 Fastgoogle/veo-3-1-fastFast, with optional audio
Veo 3.1google/veo-3-1Best quality, frame interpolation
Veo 3google/veo-3High quality with audio
Veo 3 Fastgoogle/veo-3-fastFast with audio
Veo 2google/veo-2Realistic videos
P-Videopruna/p-videoFast, economical, with audio support
WAN-T2Vpruna/wan-t2vEconomical 480p/720p
Grok Videoxai/grok-imagine-videoxAI, configurable duration
Seedance 2.0bytedance/seedance-2-0Text/image/ref-to-video with sync audio, up to 1080p
Seedance 2.0 Fastbytedance/seedance-2-0-fastFast variant, same capabilities
HappyHorse T2Valibaba/happyhorse-1-0-t2vPhysically realistic, up to 15s

Image-to-Video

ModelApp IDBest For
Wan 2.5falai/wan-2-5Animate any image
Wan 2.5 I2Vfalai/wan-2-5-i2vHigh quality i2v
WAN-I2Vpruna/wan-i2vEconomical 480p/720p
P-Videopruna/p-videoFast i2v with audio
Seedance 2.0bytedance/seedance-2-0Animate images with sync audio, up to 1080p
Seedance 2.0 Fastbytedance/seedance-2-0-fastFast variant, same capabilities
HappyHorse I2Valibaba/happyhorse-1-0-i2vAnimate images, up to 1080P/15s
HappyHorse R2Valibaba/happyhorse-1-0-r2vCharacter-preserving from references

Avatar / Lipsync

ModelApp IDBest For
OmniHuman 1.5bytedance/omnihuman-1-5Multi-character
OmniHuman 1.0bytedance/omnihuman-1-0Single character
Fabric 1.0falai/fabric-1-0Image talks with lipsync
PixVerse Lipsyncfalai/pixverse-lipsyncRealistic lipsync

Video Editing

ModelApp IDBest For
HappyHorse Editalibaba/happyhorse-1-0-video-editNatural language video editing

Utilities

ToolApp IDDescription
HunyuanVideo Foleyinfsh/hunyuanvideo-foleyAdd sound effects to video
Topaz Upscalerfalai/topaz-video-upscalerUpscale video quality
Media Mergerinfsh/media-mergerMerge videos with transitions

Browse All Video Apps

belt app list --category video

Examples

Text-to-Video with Veo

belt app run google/veo-3-1-fast --input '{
  "prompt": "A timelapse of a flower blooming in a garden"
}'

Grok Video

belt app run xai/grok-imagine-video --input '{
  "prompt": "Waves crashing on a beach at sunset",
  "duration": 5
}'

Image-to-Video with Wan 2.5

belt app run falai/wan-2-5 --input '{
  "image_url": "https://your-image.jpg"
}'

AI Avatar / Talking Head

belt app run bytedance/omnihuman-1-5 --input '{
  "image_url": "https://portrait.jpg",
  "audio_url": "https://speech.mp3"
}'

Fabric Lipsync

belt app run falai/fabric-1-0 --input '{
  "image_url": "https://face.jpg",
  "audio_url": "https://audio.mp3"
}'

Seedance 2.0 Text-to-Video with Audio

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "a jazz band performing in a dimly lit club",
  "generate_audio": true,
  "duration": 10
}'

Seedance 2.0 Image-to-Video

belt app run bytedance/seedance-2-0 --input '{
  "image": "https://your-image.jpg",
  "prompt": "gentle camera movement, leaves rustling in the wind",
  "generate_audio": true
}'

Seedance 2.0 Reference-to-Video

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "A person who looks like the reference walking through a garden",
  "reference_image": "https://portrait.jpg",
  "generate_audio": true
}'

HappyHorse Text-to-Video

belt app run alibaba/happyhorse-1-0-t2v --input '{
  "prompt": "a golden retriever running through autumn leaves, slow motion",
  "duration": 10,
  "resolution": "1080P"
}'

HappyHorse Video Editing

belt app run alibaba/happyhorse-1-0-video-edit --input '{
  "video": "https://your-video.mp4",
  "prompt": "change the background to a snowy mountain landscape"
}'

PixVerse Lipsync

belt app run falai/pixverse-lipsync --input '{
  "image_url": "https://portrait.jpg",
  "audio_url": "https://speech.mp3"
}'

Video Upscaling

belt app run falai/topaz-video-upscaler --input '{"video_url": "https://..."}'

Add Sound Effects (Foley)

belt app run infsh/hunyuanvideo-foley --input '{
  "video_url": "https://silent-video.mp4",
  "prompt": "footsteps on gravel, birds chirping"
}'

Merge Videos

belt app run infsh/media-merger --input '{
  "videos": ["https://clip1.mp4", "https://clip2.mp4"],
  "transition": "fade"
}'

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# Pruna P-Video (fast & economical)
npx skills add inference-sh/skills@p-video

# Google Veo specific
npx skills add inference-sh/skills@google-veo

# Seedance 2.0
npx skills add inference-sh/skills@seedance

# HappyHorse 1.0
npx skills add inference-sh/skills@happyhorse

# AI avatars & lipsync
npx skills add inference-sh/skills@ai-avatar-video

# Text-to-speech (for video narration)
npx skills add inference-sh/skills@text-to-speech

# Image generation (for image-to-video)
npx skills add inference-sh/skills@ai-image-generation

# Twitter (post videos)
npx skills add inference-sh/skills@twitter-automation

Browse all apps: belt app list

Documentation

skills-101 की और Skills

ai-image-generation
skills-101
inference.sh CLI के माध्यम से GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve और 50+ मॉडलों के साथ AI इमेज जनरेट करें। मॉडल: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt। क्षमताएँ: text-to-image, image-to-image, inpainting, LoRA, इमेज एडिटिंग, upscaling, टेक्स्ट रेंडरिंग। उपयोग के लिए: AI आर्ट, प्रोडक्ट मॉकअप, कॉन्सेप्ट आर्ट, सोशल मीडिया ग्राफिक्स, मार्केटिंग विज़ुअल, इलस्ट्रेशन। ट्रिगर: flux, इमेज जनरेशन, ai image, text to...
ai-avatar-video
skills-101
inference.sh CLI के माध्यम से AI अवतार और टॉकिंग हेड वीडियो बनाएं। अनुशंसित: P-Video-Avatar (सबसे तेज़, सबसे सस्ता, बिल्ट-इन TTS)। साथ ही: OmniHuman, Fabric, PixVerse। ऑडियो: Inworld TTS-2 (100+ भाषाएँ, पात्रों के लिए इमोशन स्टीयरिंग), ElevenLabs, Kokoro। क्षमताएँ: ऑडियो-संचालित अवतार, टेक्स्ट-टू-अवतार, लिपसिंक वीडियो, टॉकिंग हेड जनरेशन, वर्चुअल प्रेजेंटर, UGC कंटेंट। उपयोग के लिए: AI प्रेजेंटर, एक्सप्लेनर वीडियो, वर्चुअल इन्फ्लुएंसर, डबिंग, मार्केटिंग वीडियो, UGC विज्ञापन, गेमिंग अवतार,...
agent-browser
skills-101
inference.sh के माध्यम से AI एजेंटों के लिए ब्राउज़र स्वचालन। वेब पेजों पर नेविगेट करें, @e रेफ्स का उपयोग करके तत्वों के साथ इंटरैक्ट करें, स्क्रीनशॉट लें, वीडियो रिकॉर्ड करें। क्षमताएँ: वेब स्क्रैपिंग, फॉर्म भरना, क्लिक करना, टाइपिंग, ड्रैग-ड्रॉप, फ़ाइल अपलोड, जावास्क्रिप्ट निष्पादन। उपयोग करें: वेब स्वचालन, डेटा निष्कर्षण, परीक्षण, एजेंट ब्राउज़िंग, अनुसंधान। ट्रिगर्स: ब्राउज़र, वेब स्वचालन, स्क्रैप, नेविगेट, क्लिक, फॉर्म भरें, स्क्रीनशॉट, वेब ब्राउज़ करें, प्लेराइट, हेडलेस ब्राउज़र, वेब एजेंट, इंटरनेट सर्फ करें, वीडियो रिकॉर्ड करें
agent-tools
skills-101
inference.sh CLI के माध्यम से AI ऐप्स चलाएँ - इमेज जनरेशन, वीडियो निर्माण, LLMs, खोज, 3D, Twitter स्वचालन। मॉडल: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter, और कई और। उपयोग करें जब AI ऐप्स चलाना, इमेज/वीडियो जनरेट करना, LLMs को कॉल करना, वेब खोज, या Twitter स्वचालित करना हो। ट्रिगर्स: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
python-executor
skills-101
Execute Python code in a safe sandboxed environment via [inference.sh](https://inference.sh). Pre-installed: NumPy, Pandas, Matplotlib, requests, BeautifulSoup, Selenium, Playwright, MoviePy, Pillow, OpenCV, trimesh, and 100+ more libraries. Use for: data processing, web scraping, image manipulation, video creation, 3D model processing, PDF generation, API calls, automation scripts. Triggers: python, execute code, run script, web scraping, data analysis, image processing, video editing, 3D...
remotion-render
skills-101
React/Remotion कंपोनेंट कोड से inference.sh के माध्यम से वीडियो रेंडर करें। TSX कोड पास करें, MP4 प्राप्त करें। सभी Remotion APIs का समर्थन करता है: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence। कॉन्फ़िगर करने योग्य रिज़ॉल्यूशन, FPS, अवधि, कोडेक। उपयोग के लिए: प्रोग्रामेटिक वीडियो जनरेशन, एनिमेटेड ग्राफिक्स, मोशन डिज़ाइन, डेटा-संचालित वीडियो, React एनिमेशन से वीडियो। ट्रिगर: remotion, कोड से वीडियो रेंडर करें, tsx से वीडियो, react वीडियो, प्रोग्रामेटिक वीडियो, remotion रेंडर, कोड से वीडियो, एनिमेटेड...
infsh-cli
skills-101
inference.sh CLI के माध्यम से AI ऐप्स चलाएँ - इमेज जनरेशन, वीडियो निर्माण, LLMs, खोज, 3D, Twitter स्वचालन। मॉडल: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter, और कई अन्य। उपयोग करें जब AI ऐप्स चला रहे हों, इमेज/वीडियो जनरेट कर रहे हों, LLMs को कॉल कर रहे हों, वेब खोज कर रहे हों, या Twitter स्वचालित कर रहे हों। ट्रिगर: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
landing-page-design
skills-101
लैंडिंग पेज कन्वर्ज़न ऑप्टिमाइज़ेशन लेआउट नियमों, हीरो सेक्शन डिज़ाइन और CTA मनोविज्ञान के साथ। इसमें अबव-द-फोल्ड फॉर्मूला, सोशल प्रूफ प्लेसमेंट, मोबाइल डिज़ाइन और F-पैटर्न रीडिंग शामिल है। उपयोग के लिए: स्टार्टअप लैंडिंग पेज, प्रोडक्ट पेज, SaaS मार्केटिंग, कन्वर्ज़न ऑप्टिमाइज़ेशन। ट्रिगर: लैंडिंग पेज, हीरो सेक्शन, अबव द फोल्ड, कन्वर्ज़न ऑप्टिमाइज़ेशन, लैंडिंग पेज डिज़ाइन, CTA बटन, हीरो इमेज, लैंडिंग पेज लेआउट, SaaS लैंडिंग पेज, प्रोडक्ट पेज डिज़ाइन, कन्वर्ज़न रेट, लैंडिंग पेज...
designmarketing