ai-image-generation

Gere imagens de IA com GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve e mais de 50 modelos via CLI do inference.sh. Modelos: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Capacidades: texto para imagem, imagem para imagem, inpainting, LoRA, edição de imagem, upscaling, renderização de texto. Use para: arte de IA, mockups de produtos, arte conceitual, gráficos para redes sociais, visuais de marketing, ilustrações. Gatilhos: flux, geração de imagem, imagem de IA, texto para...

npx skills add https://github.com/halt-catch-fire/skills --skill ai-image-generation

Install the belt CLI skill: npx skills add belt-sh/cli

AI Image Generation

Generate images with 50+ AI models via inference.sh CLI.

AI Image Generation

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate an image with FLUX
belt app run falai/flux-dev-lora --input '{"prompt": "a cat astronaut in space"}'

Available Models

ModelApp IDBest For
GPT-Image-2openai/gpt-image-2Text-to-image, editing, inpainting
FLUX Dev LoRAfalai/flux-dev-loraHigh quality with custom styles
FLUX.2 Klein LoRAfalai/flux-2-klein-loraFast with LoRA support (4B/9B)
P-Imagepruna/p-imageFast, economical, multiple aspects
P-Image-LoRApruna/p-image-loraFast with preset LoRA styles
P-Image-Editpruna/p-image-editFast image editing
Gemini 3 Progoogle/gemini-3-pro-image-previewGoogle's latest
Gemini 2.5 Flashgoogle/gemini-2-5-flash-imageFast Google model
Grok Imaginexai/grok-imagine-imagexAI's model, multiple aspects
Seedream 4.5bytedance/seedream-4-52K-4K cinematic quality
Seedream 4.0bytedance/seedream-4-0High quality 2K-4K
Seedream 3.0bytedance/seedream-3-0-t2iAccurate text rendering
Revefalai/reveNatural language editing, text rendering
ImagineArt 1.5 Profalai/imagine-art-1-5-pro-previewUltra-high-fidelity 4K
FLUX Klein 4Bpruna/flux-klein-4bUltra-cheap ($0.0001/image)
Topaz Upscalerfalai/topaz-image-upscalerProfessional upscaling

Browse All Image Apps

belt app list --category image

Examples

GPT-Image-2

belt app run openai/gpt-image-2 --input '{
  "prompt": "professional product photo of sneakers, studio lighting",
  "quality": "high"
}'

GPT-Image-2 Editing

belt app run openai/gpt-image-2 --input '{
  "prompt": "change the background to a beach at sunset",
  "images": ["https://your-image.jpg"]
}'

Text-to-Image with FLUX

belt app run falai/flux-dev-lora --input '{
  "prompt": "professional product photo of a coffee mug, studio lighting"
}'

Fast Generation with FLUX Klein

belt app run falai/flux-2-klein-lora --input '{"prompt": "sunset over mountains"}'

Google Gemini 3 Pro

belt app run google/gemini-3-pro-image-preview --input '{
  "prompt": "photorealistic landscape with mountains and lake"
}'

Grok Imagine

belt app run xai/grok-imagine-image --input '{
  "prompt": "cyberpunk city at night",
  "aspect_ratio": "16:9"
}'

Reve (with Text Rendering)

belt app run falai/reve --input '{
  "prompt": "A poster that says HELLO WORLD in bold letters"
}'

Seedream 4.5 (4K Quality)

belt app run bytedance/seedream-4-5 --input '{
  "prompt": "cinematic portrait of a woman, golden hour lighting"
}'

Image Upscaling

belt app run falai/topaz-image-upscaler --input '{"image_url": "https://..."}'

Stitch Multiple Images

belt app run infsh/stitch-images --input '{
  "images": ["https://img1.jpg", "https://img2.jpg"],
  "direction": "horizontal"
}'

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# Pruna P-Image (fast & economical)
npx skills add inference-sh/skills@p-image

# GPT-Image-2 (OpenAI)
npx skills add inference-sh/skills@gpt-image

# FLUX-specific skill
npx skills add inference-sh/skills@flux-image

# Upscaling & enhancement
npx skills add inference-sh/skills@image-upscaling

# Background removal
npx skills add inference-sh/skills@background-removal

# Video generation
npx skills add inference-sh/skills@ai-video-generation

# AI avatars from images
npx skills add inference-sh/skills@ai-avatar-video

Browse all apps: belt app list

Documentation

Mais skills de halt-catch-fire

ai-video-generation
halt-catch-fire
We need to translate the given text from English to Brazilian Portuguese. The text describes an AI video generation skill. We must preserve the name "ai-video-generation" but it's not in the text, so we ignore. Preserve product names, protocol names, URLs, numbers, technical terms. So we keep: Google Veo, Seedance 2.0, HappyHorse, Wan, Grok, inference.sh CLI, Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo, text-to-video, image-to-video, reference-to-video, video editing, lipsync, avatar animation, video upscaling, foley sound, social media videos, marketing content, explainer videos, product demos, AI avatars, video generation, ai video. Also numbers like 40+ models, 1.0, 2.5, etc. Translate the rest naturally. The
creativevideomedia
twitter-automation
halt-catch-fire
Automatize o Twitter/X com postagem, engajamento e gerenciamento de usuários via CLI do inference.sh. Apps: x/post-tweet, x/post-create (com mídia), x/post-like, x/post-retweet, x/dm-send, x/user-follow. Capacidades: postar tweets, agendar conteúdo, curtir posts, retweetar, enviar DMs, seguir usuários, obter perfis. Use para: automação de mídias sociais, agendamento de conteúdo, bots de engajamento, crescimento de audiência, API do X. Gatilhos: twitter api, x api, automação de tweets, postar no twitter, twitter bot, automação de mídias sociais, x...
marketingapicommunication
ai-avatar-video
halt-catch-fire
Crie vídeos de avatar com IA e vídeos de cabeça falante via CLI do inference.sh. Recomendado: P-Video-Avatar (mais rápido, mais barato, TTS integrado). Também: OmniHuman, Fabric, PixVerse. Áudio: Inworld TTS-2 (mais de 100 idiomas, controle de emoção para personagens), ElevenLabs, Kokoro. Capacidades: avatares guiados por áudio, texto para avatar, vídeos com sincronia labial, geração de cabeça falante, apresentadores virtuais, conteúdo UGC. Use para: apresentadores de IA, vídeos explicativos, influenciadores virtuais, dublagem, vídeos de marketing, anúncios UGC, avatares de jogos,...
videocreativemedia
agent-browser
halt-catch-fire
Automação de navegador para agentes de IA via inference.sh. Navegue por páginas da web, interaja com elementos usando referências @e, tire capturas de tela, grave vídeo. Capacidades: raspagem web, preenchimento de formulários, cliques, digitação, arrastar e soltar, upload de arquivos, execução de JavaScript. Use para: automação web, extração de dados, testes, navegação por agente, pesquisa. Gatilhos: navegador, automação web, raspar, navegar, clicar, preencher formulário, captura de tela, navegar na web, playwright, navegador headless, agente web, surfar na internet, gravar vídeo
browser-automationweb-scrapingtesting
web-search
halt-catch-fire
We need to translate the given text from English to Brazilian Portuguese. The instruction says to preserve product names, protocol names, URLs, numbers, and technical terms. Also preserve the name "web-search" but only if it appears in the source text. The source text does not include "web-search" explicitly, so we don't add it. We must not add any labels or extra commentary. Just translate the text inside <text>. Let's translate: "Web search and content extraction with Tavily and Exa via inference.sh CLI. Apps: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract. Capabilities: AI-powered search, content extraction, direct answers, research. Use for: research, RAG pipelines, fact-checking, content aggregation, agents. Triggers: web search, tavily, exa, search api, content extraction, research, internet search, ai search, search assistant, web scraping, rag, perplexity alternative" We need to keep "Tavily", "Exa", "inference.sh CLI",
researchweb-scrapingapi
infsh-cli
halt-catch-fire
We need to translate the given text from English to Brazilian Portuguese. The text describes an agent skill for running AI apps via inference.sh CLI. We must preserve product names, protocol names, URLs, numbers, technical terms. The name "infsh-cli" is not in the text, so we don't include it. We translate only the text inside <text>. No extra commentary, labels, etc. Let's translate: "Run 250+ AI apps via inference.sh CLI - image generation, video creation, LLMs, search, 3D, Twitter automation. Models: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter, and many more. Use when running AI apps, generating images/videos, calling LLMs, web search, or automating Twitter. Triggers: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search,
developmentapicreative
landing-page-design
halt-catch-fire
Otimização de conversão de landing pages com regras de layout, design de seção hero e psicologia de CTA. Abrange fórmula acima da dobra, posicionamento de prova social, design mobile e padrão de leitura em F. Use para: landing pages de startups, páginas de produto, marketing SaaS, otimização de conversão. Gatilhos: landing page, seção hero, acima da dobra, otimização de conversão, design de landing page, botão cta, imagem hero, layout de landing page, landing page saas, design de página de produto, taxa de conversão, landing page...
product-photography
halt-catch-fire
Fotografia de produto com IA com iluminação de estúdio, fotos de estilo de vida e convenções de packshot. Abrange ângulos, fundos, tipos de sombra, hero shots e requisitos de imagem para e-commerce. Use para: fotos de produto, imagens de e-commerce, listagens da Amazon, packshots, fotografia de estilo de vida. Gatilhos: fotografia de produto, foto de produto, packshot, fotografia de e-commerce, foto de produto, imagem de produto, fotografia de estúdio, produto de estilo de vida, foto de produto Amazon, imagem de listagem de produto, hero shot, mockup de produto,...
creativeecommerceimage