ai-video-generation

tarafından halt-catch-fire

Google Veo, Seedance 2.0, HappyHorse, Wan, Grok ve 40'tan fazla model ile inference.sh CLI üzerinden yapay zeka videoları oluşturun. Modeller: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. Yetenekler: metinden videoya, görüntüden videoya, referanstan videoya, video düzenleme, dudak senkronizasyonu, avatar animasyonu, video yükseltme, foley ses. Kullanım alanları: sosyal medya videoları, pazarlama içerikleri, açıklayıcı videolar, ürün tanıtımları, yapay zeka avatarları. Tetikleyiciler: video oluşturma

npx skills add https://github.com/halt-catch-fire/skills --skill ai-video-generation

Install the belt CLI skill: npx skills add belt-sh/cli

AI Video Generation

Generate videos with 40+ AI models via inference.sh CLI.

AI Video Generation

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate a video with Veo
belt app run google/veo-3-1-fast --input '{"prompt": "drone shot flying over a forest"}'

Available Models

Text-to-Video

ModelApp IDBest For
Veo 3.1 Fastgoogle/veo-3-1-fastFast, with optional audio
Veo 3.1google/veo-3-1Best quality, frame interpolation
Veo 3google/veo-3High quality with audio
Veo 3 Fastgoogle/veo-3-fastFast with audio
Veo 2google/veo-2Realistic videos
P-Videopruna/p-videoFast, economical, with audio support
WAN-T2Vpruna/wan-t2vEconomical 480p/720p
Grok Videoxai/grok-imagine-videoxAI, configurable duration
Seedance 2.0bytedance/seedance-2-0Text/image/ref-to-video with sync audio, up to 1080p
Seedance 2.0 Fastbytedance/seedance-2-0-fastFast variant, same capabilities
HappyHorse T2Valibaba/happyhorse-1-0-t2vPhysically realistic, up to 15s

Image-to-Video

ModelApp IDBest For
Wan 2.5falai/wan-2-5Animate any image
Wan 2.5 I2Vfalai/wan-2-5-i2vHigh quality i2v
WAN-I2Vpruna/wan-i2vEconomical 480p/720p
P-Videopruna/p-videoFast i2v with audio
Seedance 2.0bytedance/seedance-2-0Animate images with sync audio, up to 1080p
Seedance 2.0 Fastbytedance/seedance-2-0-fastFast variant, same capabilities
HappyHorse I2Valibaba/happyhorse-1-0-i2vAnimate images, up to 1080P/15s
HappyHorse R2Valibaba/happyhorse-1-0-r2vCharacter-preserving from references

Avatar / Lipsync

ModelApp IDBest For
OmniHuman 1.5bytedance/omnihuman-1-5Multi-character
OmniHuman 1.0bytedance/omnihuman-1-0Single character
Fabric 1.0falai/fabric-1-0Image talks with lipsync
PixVerse Lipsyncfalai/pixverse-lipsyncRealistic lipsync

Video Editing

ModelApp IDBest For
HappyHorse Editalibaba/happyhorse-1-0-video-editNatural language video editing

Utilities

ToolApp IDDescription
HunyuanVideo Foleyinfsh/hunyuanvideo-foleyAdd sound effects to video
Topaz Upscalerfalai/topaz-video-upscalerUpscale video quality
Media Mergerinfsh/media-mergerMerge videos with transitions

Browse All Video Apps

belt app list --category video

Examples

Text-to-Video with Veo

belt app run google/veo-3-1-fast --input '{
  "prompt": "A timelapse of a flower blooming in a garden"
}'

Grok Video

belt app run xai/grok-imagine-video --input '{
  "prompt": "Waves crashing on a beach at sunset",
  "duration": 5
}'

Image-to-Video with Wan 2.5

belt app run falai/wan-2-5 --input '{
  "image_url": "https://your-image.jpg"
}'

AI Avatar / Talking Head

belt app run bytedance/omnihuman-1-5 --input '{
  "image_url": "https://portrait.jpg",
  "audio_url": "https://speech.mp3"
}'

Fabric Lipsync

belt app run falai/fabric-1-0 --input '{
  "image_url": "https://face.jpg",
  "audio_url": "https://audio.mp3"
}'

Seedance 2.0 Text-to-Video with Audio

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "a jazz band performing in a dimly lit club",
  "generate_audio": true,
  "duration": 10
}'

Seedance 2.0 Image-to-Video

belt app run bytedance/seedance-2-0 --input '{
  "image": "https://your-image.jpg",
  "prompt": "gentle camera movement, leaves rustling in the wind",
  "generate_audio": true
}'

Seedance 2.0 Reference-to-Video

belt app run bytedance/seedance-2-0 --input '{
  "prompt": "A person who looks like the reference walking through a garden",
  "reference_image": "https://portrait.jpg",
  "generate_audio": true
}'

HappyHorse Text-to-Video

belt app run alibaba/happyhorse-1-0-t2v --input '{
  "prompt": "a golden retriever running through autumn leaves, slow motion",
  "duration": 10,
  "resolution": "1080P"
}'

HappyHorse Video Editing

belt app run alibaba/happyhorse-1-0-video-edit --input '{
  "video": "https://your-video.mp4",
  "prompt": "change the background to a snowy mountain landscape"
}'

PixVerse Lipsync

belt app run falai/pixverse-lipsync --input '{
  "image_url": "https://portrait.jpg",
  "audio_url": "https://speech.mp3"
}'

Video Upscaling

belt app run falai/topaz-video-upscaler --input '{"video_url": "https://..."}'

Add Sound Effects (Foley)

belt app run infsh/hunyuanvideo-foley --input '{
  "video_url": "https://silent-video.mp4",
  "prompt": "footsteps on gravel, birds chirping"
}'

Merge Videos

belt app run infsh/media-merger --input '{
  "videos": ["https://clip1.mp4", "https://clip2.mp4"],
  "transition": "fade"
}'

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# Pruna P-Video (fast & economical)
npx skills add inference-sh/skills@p-video

# Google Veo specific
npx skills add inference-sh/skills@google-veo

# Seedance 2.0
npx skills add inference-sh/skills@seedance

# HappyHorse 1.0
npx skills add inference-sh/skills@happyhorse

# AI avatars & lipsync
npx skills add inference-sh/skills@ai-avatar-video

# Text-to-speech (for video narration)
npx skills add inference-sh/skills@text-to-speech

# Image generation (for image-to-video)
npx skills add inference-sh/skills@ai-image-generation

# Twitter (post videos)
npx skills add inference-sh/skills@twitter-automation

Browse all apps: belt app list

Documentation

halt-catch-fire tarafından daha fazla skill

ai-image-generation
halt-catch-fire
GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve ve inference.sh CLI üzerinden 50'den fazla model ile AI görselleri oluşturun. Modeller: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Yetenekler: metinden görsele, görselden görsele, iç boyama, LoRA, görsel düzenleme, yükseltme, metin oluşturma. Kullanım alanları: AI sanatı, ürün maketleri, konsept sanatı, sosyal medya grafikleri, pazarlama görselleri, illüstrasyonlar. Tetikleyiciler: flux, görsel oluşturma, ai görsel, metinden...
creativemediaimage
twitter-automation
halt-catch-fire
Twitter/X otomasyonu: inference.sh CLI ile gönderi, etkileşim ve kullanıcı yönetimi. Uygulamalar: x/post-tweet, x/post-create (medya ile), x/post-like, x/post-retweet, x/dm-send, x/user-follow. Yetenekler: tweet gönderme, içerik planlama, gönderi beğenme, retweet yapma, DM gönderme, kullanıcı takip etme, profil alma. Kullanım alanları: sosyal medya otomasyonu, içerik planlama, etkileşim botları, kitle büyütme, X API. Tetikleyiciler: twitter api, x api, tweet otomasyonu, twitter'a gönderi, twitter botu, sosyal medya otomasyonu, x...
marketingapicommunication
ai-avatar-video
halt-catch-fire
We need to translate the given text from English to Turkish. The target language is Türkçe. The directory item type is agent skill, and the name to preserve is "ai-avatar-video". The instruction says: "Translate only the text inside <text>. Do not include the name unless it appears in the source text." The name "ai-avatar-video" does not appear in the source text, so we should not include it. Also, do not include labels like "description", "server name", or "skill name". Just translate the content. The text: "Create AI avatar and talking head videos via inference.sh CLI. Recommended: P-Video-Avatar (fastest, cheapest, built-in TTS). Also: OmniHuman, Fabric, PixVerse. Audio: Inworld TTS-2 (100+ languages, emotion steering for characters), ElevenLabs, Kokoro. Capabilities: audio-driven avatars, text-to-avatar, lipsync videos, talking head generation, virtual presenters, UGC content. Use for: AI
videocreativemedia
agent-browser
halt-catch-fire
AI ajanları için inference.sh üzerinden tarayıcı otomasyonu. Web sayfalarında gezinme, @e referanslarıyla öğelerle etkileşim, ekran görüntüsü alma, video kaydetme. Yetenekler: web kazıma, form doldurma, tıklama, yazma, sürükle-bırak, dosya yükleme, JavaScript çalıştırma. Kullanım alanları: web otomasyonu, veri çıkarma, test, ajan taraması, araştırma. Tetikleyiciler: tarayıcı, web otomasyonu, kazıma, gezinme, tıklama, form doldurma, ekran görüntüsü, web'de gezinme, playwright, başsız tarayıcı, web ajanı, internette gezinme, video kaydetme
browser-automationweb-scrapingtesting
web-search
halt-catch-fire
We need to translate the given text from English to Turkish, preserving specific terms like "web-search", "Tavily", "Exa", "inference.sh CLI", "RAG", etc. The instruction says to preserve product names, protocol names, URLs, numbers, and technical terms. So "Tavily", "Exa", "inference.sh CLI", "RAG", "AI", "API" should remain as is. Also "web-search" is the name to preserve but it's not in the text? Actually the name is "web-search" but the text doesn't contain that exact string. The text has "web search" (two words). The instruction says "Do not include the name unless it appears in the source text." So we translate "web search" as "web araması" or similar? But careful: "web search" appears multiple times. We should translate it as "web araması" but keep technical terms like "Tavily Search" as is? The instruction says preserve product names. "Tavily Search
researchweb-scrapingapi
infsh-cli
halt-catch-fire
inference.sh CLI üzerinden 250'den fazla AI uygulamasını çalıştırın - görüntü oluşturma, video oluşturma, LLM'ler, arama, 3D, Twitter otomasyonu. Modeller: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter ve daha fazlası. AI uygulamalarını çalıştırırken, görüntü/video oluştururken, LLM'leri çağırırken, web araması yaparken veya Twitter'ı otomatikleştirirken kullanın. Tetikleyiciler: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
developmentapicreative
landing-page-design
halt-catch-fire
Açılış sayfası dönüşüm optimizasyonu; düzen kuralları, kahraman bölümü tasarımı ve CTA psikolojisi ile. Ekran üstü formülü, sosyal kanıt yerleşimi, mobil tasarım ve F-deseni okumayı kapsar. Kullanım alanları: startup açılış sayfaları, ürün sayfaları, SaaS pazarlama, dönüşüm optimizasyonu. Tetikleyiciler: açılış sayfası, kahraman bölümü, ekran üstü, dönüşüm optimizasyonu, açılış sayfası tasarımı, cta butonu, kahraman görseli, açılış sayfası düzeni, saas açılış sayfası, ürün sayfası tasarımı, dönüşüm oranı
product-photography
halt-catch-fire
Yapay zeka ile stüdyo aydınlatmalı ürün fotoğrafçılığı, yaşam tarzı çekimleri ve paket görüntüsü kuralları. Açılar, arka planlar, gölge türleri, kahraman çekimleri ve e-ticaret görsel gereksinimlerini kapsar. Kullanım alanları: ürün fotoğrafları, e-ticaret görselleri, Amazon listeleme görselleri, paket görüntüleri, yaşam tarzı fotoğrafçılığı. Tetikleyiciler: ürün fotoğrafçılığı, ürün fotoğrafı, paket görüntüsü, e-ticaret fotoğrafçılığı, ürün çekimi, ürün görseli, stüdyo fotoğrafçılığı, yaşam tarzı
creativeecommerceimage