gpt-image

โดย 101-skills

สร้างและแก้ไขรูปภาพด้วย OpenAI GPT-Image-2 ผ่าน CLI inference.sh โมเดล: GPT-Image-2 ความสามารถ: ข้อความเป็นรูปภาพ, แก้ไขรูปภาพ, การเติมเต็มภาพ, การแก้ไขแบบใช้มาสก์, การอ้างอิงหลายรูปภาพ, การสร้างแบบแบตช์ ใช้สำหรับ: ตัวอย่างผลิตภัณฑ์, ภาพการตลาด, การแก้ไขรูปภาพ, งานคอนเซ็ปต์อาร์ต, การเติมเต็มภาพ, การปรับแต่งภาพ ทริกเกอร์: gpt image, gpt-image-2, openai image, chatgpt image, dall-e, dalle, openai image generation, gpt image edit, gpt inpainting, openai dall-e, gpt 4o image

npx skills add https://github.com/101-skills/skills --skill gpt-image

Install the belt CLI skill: npx skills add belt-sh/cli

GPT-Image-2

Generate and edit images with OpenAI's GPT-Image-2 via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

belt app run openai/gpt-image-2 --input '{"prompt": "a cat astronaut floating in space"}'

Capabilities

GPT-Image-2 supports text-to-image generation, image editing with reference images, and mask-based inpainting — all through a single model.

FeatureDescription
Text-to-ImageGenerate images from text prompts
Image EditingEdit images using reference images
InpaintingMask-based editing of specific regions
Batch GenerationGenerate up to 10 images at once
Multiple FormatsPNG, JPEG, WebP output
Flexible ResolutionAny size in 32px increments (256–4096)

Examples

Text-to-Image

belt app run openai/gpt-image-2 --input '{
  "prompt": "professional product photo of sneakers on a white background, studio lighting",
  "quality": "high"
}'

Multiple Images

belt app run openai/gpt-image-2 --input '{
  "prompt": "minimalist logo design for a coffee shop",
  "n": 4,
  "quality": "medium"
}'

Image Editing with Reference

belt app run openai/gpt-image-2 --input '{
  "prompt": "change the background to a beach at sunset",
  "images": ["https://your-image.jpg"]
}'

Multi-Image Reference

belt app run openai/gpt-image-2 --input '{
  "prompt": "combine these two characters into one scene",
  "images": ["https://character1.jpg", "https://character2.jpg"]
}'

Inpainting with Mask

belt app run openai/gpt-image-2 --input '{
  "prompt": "replace with a red sports car",
  "images": ["https://street-scene.jpg"],
  "mask": "https://car-mask.png"
}'

Custom Resolution

belt app run openai/gpt-image-2 --input '{
  "prompt": "wide cinematic landscape, mountains at golden hour",
  "width": 1920,
  "height": 1080,
  "quality": "high"
}'

Fast Drafts

belt app run openai/gpt-image-2 --input '{
  "prompt": "quick concept sketch of a robot",
  "quality": "low"
}'

Pricing

Quality~Price per Image
Low$0.006
Medium$0.024
High$0.21

Larger resolutions cost more. See belt app get openai/gpt-image-2 for full pricing details.

Parameters

ParameterTypeDefaultDescription
promptstringrequiredText prompt describing the image
imagesarray-Reference image(s) for editing
maskstring-Mask image for inpainting
ninteger1Number of images (1–10)
qualitystring-low, medium, or high
widthinteger-Output width (256–4096, multiples of 32)
heightinteger-Output height (256–4096, multiples of 32)
output_formatstringpngpng, jpeg, or webp
output_compressioninteger-Compression level for jpeg/webp (0–100)

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# All image generation models
npx skills add inference-sh/skills@ai-image-generation

# FLUX models
npx skills add inference-sh/skills@flux-image

# Pruna P-Image (fast & economical)
npx skills add inference-sh/skills@p-image

Browse all image apps: belt app list --category image

Documentation

Skills เพิ่มเติมจาก 101-skills

ai-avatar-video
101-skills
สร้างวิดีโอ AI avatar และ talking head ผ่าน CLI ของ inference.sh แนะนำ: P-Video-Avatar (เร็วที่สุด ถูกที่สุด มี TTS ในตัว) นอกจากนี้: OmniHuman, Fabric, PixVerse เสียง: Inworld TTS-2 (100+ ภาษา ปรับอารมณ์ตัวละครได้), ElevenLabs, Kokoro ความสามารถ: avatar ที่ขับเคลื่อนด้วยเสียง, ข้อความเป็น avatar, วิดีโอ lipsync, การสร้าง talking head, ผู้บรรยายเสมือน, เนื้อหา UGC ใช้สำหรับ: ผู้บรรยาย AI, วิดีโออธิบาย, อินฟลูเอนเซอร์เสมือน, การพากย์, วิดีโอการตลาด, โฆษณา UGC, avatar ในเกม,...
creativevideomedia
ai-video-generation
101-skills
สร้างวิดีโอ AI ด้วย Google Veo, Seedance 2.0, HappyHorse, Wan, Grok และโมเดลอีก 40+ รายการผ่าน CLI ของ inference.sh โมเดล: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo ความสามารถ: ข้อความเป็นวิดีโอ, รูปภาพเป็นวิดีโอ, การอ้างอิงเป็นวิดีโอ, การตัดต่อวิดีโอ, การขยับริมฝีปาก, แอนิเมชันอวตาร, การเพิ่มความละเอียดวิดีโอ, เสียงประกอบ ใช้สำหรับ: วิดีโอโซเชียลมีเดีย, เนื้อหาการตลาด, วิดีโออธิบาย, การสาธิตผลิตภัณฑ์, อวตาร AI ทริกเกอร์: การสร้างวิดีโอ, วิดีโอ AI,...
creativevideomedia
ai-image-generation
101-skills
สร้างภาพ AI ด้วย GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve และโมเดลอีก 50+ รุ่นผ่าน CLI inference.sh โมเดล: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt ความสามารถ: ข้อความเป็นภาพ, ภาพเป็นภาพ, การเติมแต่งภาพ, LoRA, การแก้ไขภาพ, การเพิ่มความละเอียด, การเรนเดอร์ข้อความ ใช้สำหรับ: ศิลปะ AI, ต้นแบบผลิตภัณฑ์, คอนเซปต์อาร์ต, กราฟิกโซเชียลมีเดีย, ภาพการตลาด, ภาพประกอบ ทริกเกอร์: flux, การสร้างภาพ, ภาพ ai, ข้อความเป็น...
creativemediaimage
remotion-render
101-skills
เรนเดอร์วิดีโอจากโค้ดคอมโพเนนต์ React/Remotion ผ่าน inference.sh ส่งโค้ด TSX รับไฟล์ MP4 รองรับ Remotion API ทั้งหมด: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence ปรับแต่งความละเอียด FPS ระยะเวลา และโค้ดแยกได้ ใช้สำหรับ: สร้างวิดีโอแบบโปรแกรม, กราฟิกแอนิเมชัน, การออกแบบโมชัน, วิดีโอที่ขับเคลื่อนด้วยข้อมูล, แปลง React แอนิเมชันเป็นวิดีโอ ทริกเกอร์: remotion, เรนเดอร์วิดีโอจากโค้ด, tsx เป็นวิดีโอ, react video, programmatic video, remotion render, code to video, animated...
videocreativedevelopment
web-search
101-skills
การค้นหาเว็บและดึงเนื้อหาด้วย Tavily และ Exa ผ่าน CLI ของ inference.sh แอปพลิเคชัน: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract ความสามารถ: การค้นหาที่ขับเคลื่อนด้วย AI, การดึงเนื้อหา, คำตอบโดยตรง, การวิจัย ใช้สำหรับ: การวิจัย, ไพพ์ไลน์ RAG, การตรวจสอบข้อเท็จจริง, การรวบรวมเนื้อหา, เอเจนต์ ทริกเกอร์: การค้นหาเว็บ, tavily, exa, search api, การดึงเนื้อหา, การวิจัย, การค้นหาทางอินเทอร์เน็ต, การค้นหา AI, ผู้ช่วยค้นหา, การขูดเว็บ, rag, ทางเลือกของ perplexity
researchweb-scrapingapi
agent-tools
101-skills
เรียกใช้แอป AI ผ่าน CLI inference.sh - สร้างภาพ, สร้างวิดีโอ, LLM, ค้นหา, 3D, อัตโนมัติ Twitter โมเดล: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter และอื่นๆ อีกมากมาย ใช้เมื่อเรียกใช้แอป AI, สร้างภาพ/วิดีโอ, เรียก LLM, ค้นหาเว็บ หรืออัตโนมัติ Twitter ทริกเกอร์: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
infsh-cli
101-skills
เรียกใช้แอป AI ผ่าน CLI inference.sh - สร้างภาพ, สร้างวิดีโอ, LLM, ค้นหา, 3D, อัตโนมัติ Twitter โมเดล: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter และอื่นๆ อีกมากมาย ใช้เมื่อเรียกใช้แอป AI, สร้างภาพ/วิดีโอ, เรียก LLM, ค้นหาเว็บ, หรืออัตโนมัติ Twitter ทริกเกอร์: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
landing-page-design
101-skills
การปรับปรุงอัตราการแปลงของแลนดิ้งเพจด้วยกฎการจัดวาง การออกแบบส่วนฮีโร่ และจิตวิทยาของปุ่ม CTA ครอบคลุมสูตรเนื้อหาส่วนบน การวางหลักฐานทางสังคม การออกแบบสำหรับมือถือ และการอ่านแบบ F-pattern ใช้สำหรับ: แลนดิ้งเพจสตาร์ทอัพ หน้าผลิตภัณฑ์ การตลาด SaaS การปรับปรุงอัตราการแปลง ทริกเกอร์: แลนดิ้งเพจ ส่วนฮีโร่ ส่วนบน การปรับปรุงอัตราการแปลง การออกแบบแลนดิ้งเพจ ปุ่ม CTA รูปภาพฮีโร่ การจัดวางแลนดิ้งเพจ แลนดิ้งเพจ SaaS การออกแบบหน้าผลิตภัณฑ์ อัตราการแปลง แลนดิ้งเพจ...
designmarketingcreative