gpt-image

द्वारा qu-skills

OpenAI GPT-Image-2 के माध्यम से inference.sh CLI का उपयोग करके छवियाँ उत्पन्न और संपादित करें। मॉडल: GPT-Image-2। क्षमताएँ: टेक्स्ट-टू-इमेज, छवि संपादन, इनपेंटिंग, मास्क-आधारित संपादन, मल्टी-इमेज रेफरेंस, बैच जनरेशन। उपयोग करें: उत्पाद मॉकअप, मार्केटिंग विज़ुअल्स, छवि संपादन, कॉन्सेप्ट आर्ट, इनपेंटिंग, फोटो मैनिपुलेशन। ट्रिगर: gpt image, gpt-image-2, openai image,

npx skills add https://github.com/qu-skills/skills --skill gpt-image

Install the belt CLI skill: npx skills add belt-sh/cli

GPT-Image

Generate and edit images with OpenAI's GPT-Image-2.5 and GPT-Image-2 via inference.sh CLI.

Models

ModelApp IDPick it for
GPT-Image-2.5 Flareopenai/gpt-image-2-5-flareDefault. Higher quality than GPT-Image-2 at 50% lower latency
GPT-Image-2.5 Sunburstopenai/gpt-image-2-5-sunburstPremium edits: keeps subject and composition, tighter control across multi-turn edits
GPT-Image-2openai/gpt-image-2Previous generation

All three share the same input schema. The 2.5 models add xhigh and max quality tiers.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

belt app run openai/gpt-image-2-5-flare --input '{"prompt": "a cat astronaut floating in space"}'

Capabilities

Every GPT-Image app supports text-to-image generation, image editing with reference images, and mask-based inpainting through a single app.

FeatureDescription
Text-to-ImageGenerate images from text prompts
Image EditingEdit images using reference images
InpaintingMask-based editing of specific regions
Batch GenerationGenerate up to 10 images at once
Multiple FormatsPNG, JPEG, WebP output
Transparent BackgroundAlpha-channel PNG/WebP output for stickers, icons, product cutouts
Flexible ResolutionAny size in 16px increments (256–3840), aspect ratio up to 3:1

Examples

Text-to-Image

belt app run openai/gpt-image-2-5-flare --input '{
  "prompt": "professional product photo of sneakers on a white background, studio lighting",
  "quality": "high"
}'

Multiple Images

belt app run openai/gpt-image-2-5-flare --input '{
  "prompt": "minimalist logo design for a coffee shop",
  "n": 4,
  "quality": "medium"
}'

Image Editing with Reference

belt app run openai/gpt-image-2-5-flare --input '{
  "prompt": "change the background to a beach at sunset",
  "images": ["https://your-image.jpg"]
}'

Precise Edit with Sunburst

Sunburst is built to change only what you ask for and preserve the rest.

belt app run openai/gpt-image-2-5-sunburst --input '{
  "prompt": "change the jacket to red leather, keep the person, pose and background unchanged",
  "images": ["https://your-photo.jpg"],
  "quality": "high"
}'

Maximum Detail

belt app run openai/gpt-image-2-5-flare --input '{
  "prompt": "macro photo of a dragonfly wing, intricate veins, morning dew",
  "quality": "max"
}'

Multi-Image Reference

belt app run openai/gpt-image-2-5-flare --input '{
  "prompt": "combine these two characters into one scene",
  "images": ["https://character1.jpg", "https://character2.jpg"]
}'

Inpainting with Mask

belt app run openai/gpt-image-2-5-flare --input '{
  "prompt": "replace with a red sports car",
  "images": ["https://street-scene.jpg"],
  "mask": "https://car-mask.png"
}'

Transparent Background

Prompt for an isolated subject — describing a scene or backdrop makes the model draw one. Requires png (default) or webp output.

belt app run openai/gpt-image-2-5-flare --input '{
  "prompt": "a single red apple with a green leaf, isolated subject",
  "background": "transparent"
}'

Custom Resolution

belt app run openai/gpt-image-2-5-flare --input '{
  "prompt": "wide cinematic landscape, mountains at golden hour",
  "width": 1920,
  "height": 1080,
  "quality": "high"
}'

Fast Drafts

belt app run openai/gpt-image-2-5-flare --input '{
  "prompt": "quick concept sketch of a robot",
  "quality": "low"
}'

Pricing

Approximate price per 1024x1024 image. Flare and Sunburst cost the same.

QualityGPT-Image-2.5GPT-Image-2
low$0.006$0.006
medium$0.013$0.053
high$0.053$0.21
xhigh$0.094
max$0.21

Larger resolutions cost more. Edits add about $0.008 per reference image. See belt app get openai/gpt-image-2-5-flare for full pricing details.

Parameters

ParameterTypeDefaultDescription
promptstringrequiredText prompt describing the image
imagesarray-Reference image(s) for editing
maskstring-Mask image for inpainting
ninteger1Number of images (1–10)
qualitystringautolow, medium, high, xhigh, max (xhigh/max are 2.5 only)
widthinteger1024Output width (256–3840, multiples of 16, ratio ≤ 3:1)
heightinteger1024Output height (256–3840, multiples of 16, ratio ≤ 3:1)
output_formatstringpngpng, jpeg, or webp
output_compressioninteger-Compression level for jpeg/webp (0–100)
backgroundstringautoauto, transparent, or opaque (transparent needs png/webp)

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# All image generation models
npx skills add inference-sh/skills@ai-image-generation

# FLUX models
npx skills add inference-sh/skills@flux-image

# Pruna P-Image (fast & economical)
npx skills add inference-sh/skills@p-image

Browse all image apps: belt app list --category image

Documentation

qu-skills की और Skills

ai-video-generation
qu-skills
Google Veo, Seedance 2.0, HappyHorse, Wan, Grok और inference.sh CLI के माध्यम से 40+ मॉडलों के साथ AI वीडियो बनाएं। मॉडल: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo। क्षमताएं: टेक्स्ट-टू-वीडियो, इमेज-टू-वीडियो, रेफरेंस-टू-वीडियो, वीडियो एडिटिंग, लिपसिंक, अवतार एनिमेशन, वीडियो अपस्केलिंग, फोले साउंड। उपयोग: सोशल मी
videocreativemedia
remotion-render
qu-skills
React/Remotion कम्पोनेंट कोड से inference.sh के माध्यम से वीडियो रेंडर करें। TSX कोड पास करें, MP4 प्राप्त करें। सभी Remotion API का समर्थन: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence। कॉन्फ़िगरेबल रिज़ॉल्यूशन, FPS, अवधि, कोडेक। उपयोग: प्रोग्रामेटिक वीडियो जनरेशन, एनिमेटेड ग्राफिक्स, मोशन डिज़ाइन, डेटा-संचालित वीडियो, React एनिमेशन से वीडियो। ट्रिगर: remotion, कोड से वीडियो रेंडर करें, tsx से व
developmentvideocreative
ai-image-generation
qu-skills
inference.sh CLI के माध्यम से GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve और 50+ मॉडलों के साथ AI इमेज जनरेट करें। मॉडल: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt। क्षमताएँ: टेक्स्ट-टू-इमेज, इमेज-टू-इमेज, इनपेंटिंग, LoRA, इमेज एडिटिंग, अपस्केलिंग, टेक्स्ट रेंडरिंग। उपयोग: AI कला, उत्पाद मॉकअप, कॉन्सेप्ट आर्ट, सोशल मीडिया ग्राफ
creativemediaimage
ai-avatar-video
qu-skills
inference.sh CLI के माध्यम से AI अवतार और टॉकिंग हेड वीडियो बनाएं। अनुशंसित: P-Video-Avatar (सबसे तेज़, सबसे सस्ता, बिल्ट-इन TTS)। इसके अलावा: OmniHuman, Fabric, PixVerse। ऑडियो: Inworld TTS-2 (100+ भाषाएँ, पात्रों के लिए इमोशन स्टीयरिंग), ElevenLabs, Kokoro। क्षमताएँ: ऑडियो-संचालित अवतार, टेक्स्ट-टू-अवतार, लिपसिंक वीडियो, टॉकिंग हेड जनरेशन, वर्चुअल प्रेजेंटर, UGC कंटेंट
videocreativemedia
twitter-automation
qu-skills
इन्फरेंस.sh CLI के माध्यम से पोस्टिंग, सहभागिता और उपयोगकर्ता प्रबंधन के साथ Twitter/X को स्वचालित करें। ऐप्स: x/post-tweet, x/post-create (मीडिया के साथ), x/post-like, x/post-retweet, x/dm-send, x/user-follow। क्षमताएँ: ट्वीट पोस्ट करना, सामग्री शेड्यूल करना, पोस्ट को लाइक करना, रीट्वीट करना, डीएम भेजना, उपयोगकर्ताओं को फॉलो करना, प्रोफाइल प्राप्त करना। उपयोग करें: सोशल मीडिया ऑटोमेशन, सामग्री शेड्यूल
api
agent-browser
qu-skills
AI एजेंटों के लिए inference.sh के माध्यम से ब्राउज़र ऑटोमेशन। वेब पेजों पर नेविगेट करें, @e रेफरेंस का उपयोग करके तत्वों से इंटरैक्ट करें, स्क्रीनशॉट लें, वीडियो रिकॉर्ड करें। क्षमताएं: वेब स्क्रैपिंग, फॉर्म भरना, क्लिक करना, टाइप करना, ड्रैग-ड्रॉप, फ़ाइल अपलोड, जावास्क्रिप्ट निष्पादन। उपयोग करें: वेब ऑटोमेशन, डेटा निष्कर्षण, परीक्षण, एजेंट ब्राउज़िंग, अनु
browser-automationweb-scrapingtesting
web-search
qu-skills
वेब खोज और सामग्री निष्कर्षण Tavily और Exa के माध्यम से inference.sh CLI के जरिए। ऐप्स: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract। क्षमताएँ: AI-संचालित खोज, सामग्री निष्कर्षण, सीधे उत्तर, शोध। उपयोग: शोध, RAG पाइपलाइन, तथ्य-जांच, सामग्री एकत्रीकरण, एजेंट। ट्रिगर: वेब खोज, tavily, exa, search api, सामग्री निष्कर्षण, शोध, इंटरनेट खोज, ai search, search assistant, वेब स्क्रैपिंग, rag, perplexity विकल्प
researchweb-scrapingapi
agent-tools
qu-skills
इन्फरेंस.शॉ CLI के माध्यम से 250+ AI ऐप्स चलाएँ - इमेज जनरेशन, वीडियो निर्माण, LLMs, सर्च, 3D, ट्विटर ऑटोमेशन। मॉडल: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter, और कई अन्य। AI ऐप्स चलाते समय, इमेज/वीडियो जनरेट करते समय, LLMs कॉल करते समय, वेब सर्च करते समय, या ट्विटर ऑटोमेट करते समय उपयोग करें। ट्रिगर: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux,
developmentapicreative