ai-image-generation

Generieren Sie KI-Bilder mit GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve und über 50 Modellen über die inference.sh CLI. Modelle: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Fähigkeiten: Text-zu-Bild, Bild-zu-Bild, Inpainting, LoRA, Bildbearbeitung, Hochskalierung, Textrendering. Verwendung für: KI-Kunst, Produkt-Mockups, Konzeptkunst, Social-Media-Grafiken, Marketing-Visuals, Illustrationen. Auslöser: flux, Bildgenerierung, KI-Bild, Text zu...

npx skills add https://github.com/halt-catch-fire/skills --skill ai-image-generation

Install the belt CLI skill: npx skills add belt-sh/cli

AI Image Generation

Generate images with 50+ AI models via inference.sh CLI.

AI Image Generation

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate an image with FLUX
belt app run falai/flux-dev-lora --input '{"prompt": "a cat astronaut in space"}'

Available Models

ModelApp IDBest For
GPT-Image-2openai/gpt-image-2Text-to-image, editing, inpainting
FLUX Dev LoRAfalai/flux-dev-loraHigh quality with custom styles
FLUX.2 Klein LoRAfalai/flux-2-klein-loraFast with LoRA support (4B/9B)
P-Imagepruna/p-imageFast, economical, multiple aspects
P-Image-LoRApruna/p-image-loraFast with preset LoRA styles
P-Image-Editpruna/p-image-editFast image editing
Gemini 3 Progoogle/gemini-3-pro-image-previewGoogle's latest
Gemini 2.5 Flashgoogle/gemini-2-5-flash-imageFast Google model
Grok Imaginexai/grok-imagine-imagexAI's model, multiple aspects
Seedream 4.5bytedance/seedream-4-52K-4K cinematic quality
Seedream 4.0bytedance/seedream-4-0High quality 2K-4K
Seedream 3.0bytedance/seedream-3-0-t2iAccurate text rendering
Revefalai/reveNatural language editing, text rendering
ImagineArt 1.5 Profalai/imagine-art-1-5-pro-previewUltra-high-fidelity 4K
FLUX Klein 4Bpruna/flux-klein-4bUltra-cheap ($0.0001/image)
Topaz Upscalerfalai/topaz-image-upscalerProfessional upscaling

Browse All Image Apps

belt app list --category image

Examples

GPT-Image-2

belt app run openai/gpt-image-2 --input '{
  "prompt": "professional product photo of sneakers, studio lighting",
  "quality": "high"
}'

GPT-Image-2 Editing

belt app run openai/gpt-image-2 --input '{
  "prompt": "change the background to a beach at sunset",
  "images": ["https://your-image.jpg"]
}'

Text-to-Image with FLUX

belt app run falai/flux-dev-lora --input '{
  "prompt": "professional product photo of a coffee mug, studio lighting"
}'

Fast Generation with FLUX Klein

belt app run falai/flux-2-klein-lora --input '{"prompt": "sunset over mountains"}'

Google Gemini 3 Pro

belt app run google/gemini-3-pro-image-preview --input '{
  "prompt": "photorealistic landscape with mountains and lake"
}'

Grok Imagine

belt app run xai/grok-imagine-image --input '{
  "prompt": "cyberpunk city at night",
  "aspect_ratio": "16:9"
}'

Reve (with Text Rendering)

belt app run falai/reve --input '{
  "prompt": "A poster that says HELLO WORLD in bold letters"
}'

Seedream 4.5 (4K Quality)

belt app run bytedance/seedream-4-5 --input '{
  "prompt": "cinematic portrait of a woman, golden hour lighting"
}'

Image Upscaling

belt app run falai/topaz-image-upscaler --input '{"image_url": "https://..."}'

Stitch Multiple Images

belt app run infsh/stitch-images --input '{
  "images": ["https://img1.jpg", "https://img2.jpg"],
  "direction": "horizontal"
}'

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# Pruna P-Image (fast & economical)
npx skills add inference-sh/skills@p-image

# GPT-Image-2 (OpenAI)
npx skills add inference-sh/skills@gpt-image

# FLUX-specific skill
npx skills add inference-sh/skills@flux-image

# Upscaling & enhancement
npx skills add inference-sh/skills@image-upscaling

# Background removal
npx skills add inference-sh/skills@background-removal

# Video generation
npx skills add inference-sh/skills@ai-video-generation

# AI avatars from images
npx skills add inference-sh/skills@ai-avatar-video

Browse all apps: belt app list

Documentation

Mehr Skills von halt-catch-fire

ai-video-generation
halt-catch-fire
Erstelle KI-Videos mit Google Veo, Seedance 2.0, HappyHorse, Wan, Grok und 40+ Modellen über die inference.sh CLI. Modelle: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. Fähigkeiten: Text-zu-Video, Bild-zu-Video, Referenz-zu-Video, Videobearbeitung, Lippen-Sync, Avatar-Animation, Video-Upscaling, Foley-Sound. Verwendung für: Social-Media-Videos, Marketinginhalte, Erklärvideos, Produktdemos, KI-Avatare. Auslöser: Videoerstellung, KI-Video,...
creativevideomedia
twitter-automation
halt-catch-fire
Automatisiere Twitter/X mit Posting, Engagement und Benutzerverwaltung über die inference.sh CLI. Apps: x/post-tweet, x/post-create (mit Medien), x/post-like, x/post-retweet, x/dm-send, x/user-follow. Fähigkeiten: Tweets posten, Inhalte planen, Beiträge liken, retweeten, DMs senden, Benutzern folgen, Profile abrufen. Verwendung für: Social-Media-Automatisierung, Inhaltsplanung, Engagement-Bots, Zielgruppenwachstum, X API. Auslöser: twitter api, x api, tweet automation, post to twitter, twitter bot, social media automation, x...
marketingapicommunication
ai-avatar-video
halt-catch-fire
Erstelle KI-Avatar- und Talking-Head-Videos über die inference.sh CLI. Empfohlen: P-Video-Avatar (schnellste, günstigste, integrierte TTS). Auch: OmniHuman, Fabric, PixVerse. Audio: Inworld TTS-2 (100+ Sprachen, Emotionssteuerung für Charaktere), ElevenLabs, Kokoro. Fähigkeiten: audiogesteuerte Avatare, Text-zu-Avatar, Lipsync-Videos, Talking-Head-Generierung, virtuelle Präsentatoren, UGC-Inhalte. Verwenden für: KI-Präsentatoren, Erklärvideos, virtuelle Influencer, Synchronisation, Marketingvideos, UGC-Anzeigen, Gaming-Avatare,...
videocreativemedia
agent-browser
halt-catch-fire
Browser-Automatisierung für KI-Agenten über inference.sh. Navigieren Sie auf Webseiten, interagieren Sie mit Elementen über @e-Referenzen, machen Sie Screenshots, nehmen Sie Video auf. Fähigkeiten: Web Scraping, Formularausfüllen, Klicken, Tippen, Drag & Drop, Datei-Upload, JavaScript-Ausführung. Verwendung für: Web-Automatisierung, Datenextraktion, Testen, Agenten-Browsing, Recherche. Auslöser: Browser, Web-Automatisierung, Scrapen, Navigieren, Klicken, Formular ausfüllen, Screenshot, Web durchsuchen, Playwright, Headless-Browser, Web-Agent, Internet surfen, Video aufnehmen
browser-automationweb-scrapingtesting
web-search
halt-catch-fire
We need to translate the given English text into German, preserving the specified name "web-search" and other technical terms. The text describes a web search and content extraction skill using Tavily and Exa via inference.sh CLI. It lists apps, capabilities, uses, and triggers. We must not add any extra commentary or labels. The translation should be accurate and natural in German. Let's break down the text: "Web search and content extraction with Tavily and Exa via inference.sh CLI." -> "Websuche und Inhaltsextraktion mit Tavily und Exa über die inference.sh CLI." "Apps: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract." -> "Apps: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract." (preserve names) "Capabilities: AI-powered search, content extraction, direct answers, research." -> "Fähigkeiten: KI-gestützte Suche, Inhaltsextraktion, direkte Antworten, Recherche
researchweb-scrapingapi
infsh-cli
halt-catch-fire
Führe über 250 KI-Apps via inference.sh CLI aus – Bildgenerierung, Videocrstellung, LLMs, Suche, 3D, Twitter-Automatisierung. Modelle: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter und viele mehr. Verwenden beim Ausführen von KI-Apps, Generieren von Bildern/Videos, Aufrufen von LLMs, Websuche oder Automatisieren von Twitter. Auslöser: inference.sh, infsh, KI-Modell, run ai, serverless ai, ai api, flux, veo, claude api, Bildgenerierung, Videogenerierung, openrouter, tavily, exa search, twitter api, grok
developmentapicreative
landing-page-design
halt-catch-fire
Optimierung der Landingpage-Konversion mit Layoutregeln, Hero-Bereich-Design und CTA-Psychologie. Behandelt die Above-the-Fold-Formel, Platzierung sozialer Beweise, mobiles Design und F-Pattern-Leseverhalten. Verwendung für: Startup-Landingpages, Produktseiten, SaaS-Marketing, Konversionsoptimierung. Auslöser: Landingpage, Hero-Bereich, Above the Fold, Konversionsoptimierung, Landingpage-Design, CTA-Button, Hero-Bild, Landingpage-Layout, SaaS-Landingpage, Produktseitendesign, Konversionsrate, Landingpage...
product-photography
halt-catch-fire
KI-Produktfotografie mit Studio-Beleuchtung, Lifestyle-Aufnahmen und Packshot-Konventionen. Behandelt Winkel, Hintergründe, Schattenarten, Hero-Shots und E-Commerce-Bildanforderungen. Verwendung für: Produktfotos, E-Commerce-Bilder, Amazon-Listings, Packshots, Lifestyle-Fotografie. Auslöser: Produktfotografie, Produktfoto, Packshot, E-Commerce-Fotografie, Produktaufnahme, Produktbild, Studiofotografie, Lifestyle-Produkt, Amazon-Produktfoto, Produkt-Listing-Bild, Hero-Shot, Produkt-Mockup,...
creativeecommerceimage