PluginBench
Skill
Review
Audit score 70

gpt-image

101-skills/superpowers

Generate and edit images with OpenAI's GPT-Image-2 via inference.sh CLI.

What is gpt-image?

This skill provides access to OpenAI's GPT-Image-2 model for text-to-image generation, image editing, and mask-based inpainting. Use it when you need to create product mockups, marketing visuals, concept art, or perform photo manipulation tasks.

  • Generate images from text prompts with customizable quality levels
  • Edit existing images using reference images and text prompts
  • Perform mask-based inpainting to replace specific regions in images
  • Generate multiple variations (up to 10) in a single batch
  • Support flexible output resolutions from 256–4096px in 32px increments
  • Output in PNG, JPEG, or WebP formats with compression options

How to install gpt-image

npx skills add https://github.com/101-skills/superpowers --skill gpt-image
Prerequisites
  • Install the belt CLI skill: npx skills add belt-sh/cli
  • Run belt login to authenticate with inference.sh
  • Have an active inference.sh account with API access
Claude Code
Cursor
Windsurf
Cline

How to use gpt-image

  1. 1.Run 'belt login' to authenticate with your inference.sh account
  2. 2.Construct a JSON input object with your prompt and desired parameters
  3. 3.Execute 'belt app run openai/gpt-image-2 --input' with your JSON payload
  4. 4.Retrieve the generated image URL(s) from the response
  5. 5.Optionally adjust quality, resolution, or other parameters for subsequent runs

Use cases

Good for
  • Create product mockups and marketing visuals from text descriptions
  • Generate concept art and design variations for creative projects
  • Edit photos by changing backgrounds, objects, or other elements
  • Perform inpainting to remove or replace specific regions in images
  • Combine multiple reference images into a single cohesive scene
Who it's for
  • Product designers and marketers creating visual assets
  • Creative professionals working on concept art and illustrations
  • Content creators needing image editing and manipulation
  • Teams requiring batch image generation for campaigns
  • Developers building image generation into applications

gpt-image FAQ

What is the difference between low, medium, and high quality?

Quality affects both the visual fidelity and cost: low (~$0.006), medium (~$0.024), and high (~$0.21) per image. Larger resolutions also increase cost.

Can I generate multiple images at once?

Yes, set the 'n' parameter to generate 1–10 images in a single request.

How do I use inpainting to edit specific regions?

Provide the original image URL in the 'images' array, a mask image URL showing the region to edit, and a prompt describing what should replace the masked area.

What image formats are supported for output?

PNG (default), JPEG, and WebP are supported. You can also control compression levels for JPEG and WebP.

What resolution options are available?

Any size from 256–4096 pixels in 32-pixel increments for both width and height, allowing custom aspect ratios.

Full instructions (SKILL.md)

Source of truth, from 101-skills/superpowers.


name: gpt-image description: "Generate and edit images with OpenAI GPT-Image-2 via inference.sh CLI. Models: GPT-Image-2. Capabilities: text-to-image, image editing, inpainting, mask-based editing, multi-image reference, batch generation. Use for: product mockups, marketing visuals, image editing, concept art, inpainting, photo manipulation. Triggers: gpt image, gpt-image-2, openai image, chatgpt image, dall-e, dalle, openai image generation, gpt image edit, gpt inpainting, openai dall-e, gpt 4o image" allowed-tools: Bash(belt *)

Install the belt CLI skill: npx skills add belt-sh/cli

GPT-Image-2

Generate and edit images with OpenAI's GPT-Image-2 via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

belt app run openai/gpt-image-2 --input '{"prompt": "a cat astronaut floating in space"}'

Capabilities

GPT-Image-2 supports text-to-image generation, image editing with reference images, and mask-based inpainting — all through a single model.

FeatureDescription
Text-to-ImageGenerate images from text prompts
Image EditingEdit images using reference images
InpaintingMask-based editing of specific regions
Batch GenerationGenerate up to 10 images at once
Multiple FormatsPNG, JPEG, WebP output
Flexible ResolutionAny size in 32px increments (256–4096)

Examples

Text-to-Image

belt app run openai/gpt-image-2 --input '{
  "prompt": "professional product photo of sneakers on a white background, studio lighting",
  "quality": "high"
}'

Multiple Images

belt app run openai/gpt-image-2 --input '{
  "prompt": "minimalist logo design for a coffee shop",
  "n": 4,
  "quality": "medium"
}'

Image Editing with Reference

belt app run openai/gpt-image-2 --input '{
  "prompt": "change the background to a beach at sunset",
  "images": ["https://your-image.jpg"]
}'

Multi-Image Reference

belt app run openai/gpt-image-2 --input '{
  "prompt": "combine these two characters into one scene",
  "images": ["https://character1.jpg", "https://character2.jpg"]
}'

Inpainting with Mask

belt app run openai/gpt-image-2 --input '{
  "prompt": "replace with a red sports car",
  "images": ["https://street-scene.jpg"],
  "mask": "https://car-mask.png"
}'

Custom Resolution

belt app run openai/gpt-image-2 --input '{
  "prompt": "wide cinematic landscape, mountains at golden hour",
  "width": 1920,
  "height": 1080,
  "quality": "high"
}'

Fast Drafts

belt app run openai/gpt-image-2 --input '{
  "prompt": "quick concept sketch of a robot",
  "quality": "low"
}'

Pricing

Quality~Price per Image
Low$0.006
Medium$0.024
High$0.21

Larger resolutions cost more. See belt app get openai/gpt-image-2 for full pricing details.

Parameters

ParameterTypeDefaultDescription
promptstringrequiredText prompt describing the image
imagesarray-Reference image(s) for editing
maskstring-Mask image for inpainting
ninteger1Number of images (1–10)
qualitystring-low, medium, or high
widthinteger-Output width (256–4096, multiples of 32)
heightinteger-Output height (256–4096, multiples of 32)
output_formatstringpngpng, jpeg, or webp
output_compressioninteger-Compression level for jpeg/webp (0–100)

Related Skills

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# All image generation models
npx skills add inference-sh/skills@ai-image-generation

# FLUX models
npx skills add inference-sh/skills@flux-image

# Pruna P-Image (fast & economical)
npx skills add inference-sh/skills@p-image

Browse all image apps: belt app list --category image

Documentation