gpt-image

द्वारा 101-skills

inference.sh CLI के माध्यम से OpenAI GPT-Image-2 के साथ छवियाँ बनाएँ और संपादित करें। मॉडल: GPT-Image-2। क्षमताएँ: टेक्स्ट-टू-इमेज, छवि संपादन, इनपेंटिंग, मास्क-आधारित संपादन, मल्टी-इमेज रेफरेंस, बैच जनरेशन। उपयोग के लिए: प्रोडक्ट मॉकअप, मार्केटिंग विज़ुअल, छवि संपादन, कॉन्सेप्ट आर्ट, इनपेंटिंग, फोटो मैनिपुलेशन। ट्रिगर्स: gpt image, gpt-image-2, openai image, chatgpt image, dall-e, dalle, openai image generation, gpt image edit, gpt inpainting, openai dall-e, gpt 4o image

npx skills add https://github.com/101-skills/skills --skill gpt-image

Install the belt CLI skill: npx skills add belt-sh/cli

GPT-Image-2

Generate and edit images with OpenAI's GPT-Image-2 via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

belt app run openai/gpt-image-2 --input '{"prompt": "a cat astronaut floating in space"}'

Capabilities

GPT-Image-2 supports text-to-image generation, image editing with reference images, and mask-based inpainting — all through a single model.

FeatureDescription
Text-to-ImageGenerate images from text prompts
Image EditingEdit images using reference images
InpaintingMask-based editing of specific regions
Batch GenerationGenerate up to 10 images at once
Multiple FormatsPNG, JPEG, WebP output
Flexible ResolutionAny size in 32px increments (256–4096)

Examples

Text-to-Image

belt app run openai/gpt-image-2 --input '{
  "prompt": "professional product photo of sneakers on a white background, studio lighting",
  "quality": "high"
}'

Multiple Images

belt app run openai/gpt-image-2 --input '{
  "prompt": "minimalist logo design for a coffee shop",
  "n": 4,
  "quality": "medium"
}'

Image Editing with Reference

belt app run openai/gpt-image-2 --input '{
  "prompt": "change the background to a beach at sunset",
  "images": ["https://your-image.jpg"]
}'

Multi-Image Reference

belt app run openai/gpt-image-2 --input '{
  "prompt": "combine these two characters into one scene",
  "images": ["https://character1.jpg", "https://character2.jpg"]
}'

Inpainting with Mask

belt app run openai/gpt-image-2 --input '{
  "prompt": "replace with a red sports car",
  "images": ["https://street-scene.jpg"],
  "mask": "https://car-mask.png"
}'

Custom Resolution

belt app run openai/gpt-image-2 --input '{
  "prompt": "wide cinematic landscape, mountains at golden hour",
  "width": 1920,
  "height": 1080,
  "quality": "high"
}'

Fast Drafts

belt app run openai/gpt-image-2 --input '{
  "prompt": "quick concept sketch of a robot",
  "quality": "low"
}'

Pricing

Quality~Price per Image
Low$0.006
Medium$0.024
High$0.21

Larger resolutions cost more. See belt app get openai/gpt-image-2 for full pricing details.

Parameters

ParameterTypeDefaultDescription
promptstringrequiredText prompt describing the image
imagesarray-Reference image(s) for editing
maskstring-Mask image for inpainting
ninteger1Number of images (1–10)
qualitystring-low, medium, or high
widthinteger-Output width (256–4096, multiples of 32)
heightinteger-Output height (256–4096, multiples of 32)
output_formatstringpngpng, jpeg, or webp
output_compressioninteger-Compression level for jpeg/webp (0–100)

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# All image generation models
npx skills add inference-sh/skills@ai-image-generation

# FLUX models
npx skills add inference-sh/skills@flux-image

# Pruna P-Image (fast & economical)
npx skills add inference-sh/skills@p-image

Browse all image apps: belt app list --category image

Documentation

101-skills की और Skills

ai-avatar-video
101-skills
inference.sh CLI के माध्यम से AI अवतार और टॉकिंग हेड वीडियो बनाएं। अनुशंसित: P-Video-Avatar (सबसे तेज़, सबसे सस्ता, अंतर्निहित TTS)। इसके अलावा: OmniHuman, Fabric, PixVerse। ऑडियो: Inworld TTS-2 (100+ भाषाएँ, पात्रों के लिए भावना नियंत्रण), ElevenLabs, Kokoro। क्षमताएँ: ऑडियो-संचालित अवतार, टेक्स्ट-टू-अवतार, लिपसिंक वीडियो, टॉकिंग हेड जनरेशन, वर्चुअल प्रस्तुतकर्ता, UGC सामग्री। उपयोग के लिए: AI प्रस्तुतकर्ता, व्याख्यात्मक वीडियो, वर्चुअल प्रभावशाली, डबिंग, मार्केटिंग वीडियो, UGC विज्ञापन, गेमिंग अवतार,...
creativevideomedia
ai-video-generation
101-skills
Google Veo, Seedance 2.0, HappyHorse, Wan, Grok और 40+ मॉडल्स के साथ inference.sh CLI के माध्यम से AI वीडियो जनरेट करें। मॉडल्स: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo। क्षमताएँ: text-to-video, image-to-video, reference-to-video, वीडियो एडिटिंग, lipsync, अवतार एनिमेशन, वीडियो अपस्केलिंग, foley sound। उपयोग: सोशल मीडिया वीडियो, मार्केटिंग कंटेंट, एक्सप्लेनर वीडियो, प्रोडक्ट डेमो, AI अवतार। ट्रिगर्स: वीडियो जनरेशन, ai video,...
creativevideomedia
ai-image-generation
101-skills
GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve और 50+ मॉडल्स के साथ inference.sh CLI के माध्यम से AI इमेजेस जनरेट करें। मॉडल्स: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt। क्षमताएँ: टेक्स्ट-टू-इमेज, इमेज-टू-इमेज, इनपेंटिंग, LoRA, इमेज एडिटिंग, अपस्केलिंग, टेक्स्ट रेंडरिंग। उपयोग के लिए: AI कला, उत्पाद मॉकअप, कॉन्सेप्ट आर्ट, सोशल मीडिया ग्राफिक्स, मार्केटिंग विज़ुअल्स, इलस्ट्रेशन। ट्रिगर्स: फ्लक्स, इमेज जनरेशन, एआई इमेज, टेक्स्ट टू...
creativemediaimage
remotion-render
101-skills
inference.sh के माध्यम से React/Remotion कंपोनेंट कोड से वीडियो रेंडर करें। TSX कोड पास करें, MP4 प्राप्त करें। सभी Remotion APIs का समर्थन करता है: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence। कॉन्फ़िगर करने योग्य रिज़ॉल्यूशन, FPS, अवधि, कोडेक। उपयोग के लिए: प्रोग्रामेटिक वीडियो जनरेशन, एनिमेटेड ग्राफिक्स, मोशन डिज़ाइन, डेटा-संचालित वीडियो, React एनिमेशन से वीडियो। ट्रिगर: remotion, render video from code, tsx to video, react video, programmatic video, remotion render, code to video, animated...
videocreativedevelopment
web-search
101-skills
वेब खोज और सामग्री निष्कर्षण Tavily और Exa के माध्यम से inference.sh CLI के जरिए। ऐप्स: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract। क्षमताएँ: AI-संचालित खोज, सामग्री निष्कर्षण, सीधे उत्तर, शोध। उपयोग: शोध, RAG पाइपलाइन, तथ्य-जांच, सामग्री एकत्रीकरण, एजेंट। ट्रिगर: वेब खोज, tavily, exa, search api, सामग्री निष्कर्षण, शोध, इंटरनेट खोज, ai search, search assistant, वेब स्क्रैपिंग, rag, perplexity alternative
researchweb-scrapingapi
agent-tools
101-skills
inference.sh CLI के माध्यम से AI ऐप्स चलाएँ - इमेज जनरेशन, वीडियो निर्माण, LLMs, खोज, 3D, Twitter स्वचालन। मॉडल: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter, और कई अन्य। AI ऐप्स चलाने, इमेज/वीडियो जनरेट करने, LLMs कॉल करने, वेब खोज, या Twitter स्वचालित करने पर उपयोग करें। ट्रिगर: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
infsh-cli
101-skills
inference.sh CLI के माध्यम से AI ऐप्स चलाएँ - छवि निर्माण, वीडियो निर्माण, LLMs, खोज, 3D, Twitter स्वचालन। मॉडल: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter, और कई अन्य। AI ऐप्स चलाने, छवियाँ/वीडियो बनाने, LLMs कॉल करने, वेब खोज, या Twitter स्वचालित करने पर उपयोग करें। ट्रिगर: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
landing-page-design
101-skills
लैंडिंग पृष्ठ रूपांतरण अनुकूलन, लेआउट नियमों, हीरो सेक्शन डिज़ाइन और CTA मनोविज्ञान के साथ। इसमें ऊपर-फोल्ड सूत्र, सामाजिक प्रमाण स्थान, मोबाइल डिज़ाइन और F-पैटर्न पठन शामिल है। उपयोग के लिए: स्टार्टअप लैंडिंग पृष्ठ, उत्पाद पृष्ठ, SaaS मार्केटिंग, रूपांतरण अनुकूलन। ट्रिगर: लैंडिंग पृष्ठ, हीरो सेक्शन, ऊपर-फोल्ड, रूपांतरण अनुकूलन, लैंडिंग पृष्ठ डिज़ाइन, CTA बटन, हीरो छवि, लैंडिंग पृष्ठ लेआउट, SaaS लैंडिंग पृष्ठ, उत्पाद पृष्ठ डिज़ाइन, रूपांतरण दर, लैंडिंग पृष्ठ...
designmarketingcreative