Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
intellectronica avatar

Gpt Image 1 5

  • 564 installs
  • 281 repo stars
  • Updated April 25, 2026
  • intellectronica/agent-skills

gpt-image-1-5 is a Claude and Cursor skill that generates and edits PNG images via OpenAI's GPT Image 1.5 model for developers who need text-to-image or inpainting workflows from the terminal.

About

gpt-image-1-5 is an intellectlectronica/agent-skills skill that lets coding agents generate or edit images with OpenAI's GPT Image 1.5 model through scripts/generate_image.py run via uv. Text-to-image generation uses the Responses API image_generation tool, while editing uses the Image API images.edit endpoint with optional --input-image and --mask for inpainting. Quality flags are low, medium, and high; size options are 1024x1024, 1024x1536, 1536x1024, and auto; background can be transparent, opaque, or auto on generation. The script reads OPENAI_API_KEY from the environment or --api-key and saves timestamped PNG files to the developer's current working directory. Developers reach for gpt-image-1-5 when asked to create, modify, change backgrounds, or inpaint regions in an existing image without manually opening an image editor.

  • Text-to-image generation via Responses API with image_generation tool
  • Reliable mask-based inpainting using the Image API for precise edits
  • Supports full-image editing without mask when only prompt is provided
  • Command-line flags for quality, size, background transparency and API key
  • Direct --input-image parameter workflow that skips manual file reading

Gpt Image 1 5 by the numbers

  • 564 all-time installs (skills.sh)
  • +10 installs in the week ending Aug 2, 2026 (Skillselion tracking)
  • Ranked #383 of 1,335 Generative Media skills by installs in the Skillselion catalog
  • Data as of Aug 2, 2026 (Skillselion catalog sync)
npx skills add https://github.com/intellectronica/agent-skills --skill gpt-image-1-5

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs564
repo stars281
Last updatedApril 25, 2026
Repositoryintellectronica/agent-skills

How do you generate images with GPT Image 1.5?

Generate and edit images on demand directly from Claude or Cursor using OpenAI's GPT Image 1.5 model.

Who is it for?

Developers using Claude or Cursor who want terminal-driven GPT Image 1.5 generation and mask-based editing with OPENAI_API_KEY.

Skip if: Developers without an OpenAI API key or those needing vector or SVG output should skip gpt-image-1-5 because it only emits PNG raster images.

When should I use this skill?

A user asks to generate, create, edit, modify, or inpaint an image file using GPT Image 1.5.

What you get

Timestamped PNG image files saved to the working directory with paths printed by the generate_image.py script.

  • PNG image files
  • Inpainted image variants

By the numbers

  • Supports 4 size options: 1024x1024, 1024x1536, 1536x1024, and auto
  • Offers 3 quality levels: low, medium, and high

Files

SKILL.mdMarkdownGitHub ↗

GPT Image 1.5 - Image Generation & Editing

Generate new images or edit existing ones using OpenAI's GPT Image 1.5 model.

  • Generation: Uses the Responses API with image_generation tool
  • Editing: Uses the Image API for reliable mask-based inpainting

Usage

Run the script using absolute path (do NOT cd to skill directory first):

Generate new image:

uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "your image description" --filename "output-name.png" [--quality low|medium|high] [--size 1024x1024|1024x1536|1536x1024|auto] [--background transparent|opaque|auto] [--api-key KEY]

Edit existing image (without mask - full image edit):

uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "editing instructions" --filename "output-name.png" --input-image "path/to/input.png" [--size 1024x1024|1024x1536|1536x1024|auto] [--api-key KEY]

Edit existing image (with mask - precise inpainting):

uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "what to put in masked area" --filename "output-name.png" --input-image "path/to/input.png" --mask "path/to/mask.png" [--size 1024x1024|1024x1536|1536x1024|auto] [--api-key KEY]

Important: Always run from the user's current working directory so images are saved where the user is working, not in the skill directory.

Parameters

Quality Options

  • low - Fastest generation, lower quality
  • medium (default) - Balanced quality and speed
  • high - Best quality, slower generation

Map user requests:

  • No mention of quality -> medium
  • "quick", "fast", "draft" -> low
  • "high quality", "best", "detailed", "high-res" -> high

Size Options

  • 1024x1024 (default) - Square format
  • 1024x1536 - Portrait format
  • 1536x1024 - Landscape format
  • auto - Let the model decide based on prompt

Map user requests:

  • No mention of size -> 1024x1024
  • "square" -> 1024x1024
  • "portrait", "vertical", "tall" -> 1024x1536
  • "landscape", "horizontal", "wide" -> 1536x1024

Background Options (generation only)

  • auto (default) - Model decides
  • transparent - Transparent background (PNG/WebP output)
  • opaque - Solid background

API Key

The script checks for API key in this order: 1. --api-key argument (use if user provided key in chat) 2. OPENAI_API_KEY environment variable

If neither is available, the script exits with an error message.

Filename Generation

Generate filenames with the pattern: yyyy-mm-dd-hh-mm-ss-name.png

Format: {timestamp}-{descriptive-name}.png

  • Timestamp: Current date/time in format yyyy-mm-dd-hh-mm-ss (24-hour format)
  • Name: Descriptive lowercase text with hyphens
  • Keep the descriptive part concise (1-5 words typically)
  • Use context from user's prompt or conversation
  • If unclear, use random identifier (e.g., x9k2, a7b3)

Examples:

  • Prompt "A serene Japanese garden" -> 2025-12-17-14-23-05-japanese-garden.png
  • Prompt "sunset over mountains" -> 2025-12-17-15-30-12-sunset-mountains.png
  • Prompt "create an image of a robot" -> 2025-12-17-16-45-33-robot.png
  • Unclear context -> 2025-12-17-17-12-48-x9k2.png

Image Editing

Both editing modes use the Image API (images.edit endpoint) with gpt-image-1.5 for reliable results.

Without Mask (Full Image Edit)

When the user wants to modify an existing image without specifying exact regions: 1. Use --input-image parameter with the path to the image 2. The prompt should contain editing instructions (e.g., "make the sky more dramatic", "change to cartoon style") 3. A fully transparent mask is auto-generated, allowing the model to edit the entire image

With Mask (Precise Inpainting)

When the user wants to edit specific regions: 1. Use --input-image parameter with the path to the image 2. Use --mask parameter with a PNG mask file 3. The mask should have transparent areas (alpha=0) where edits should occur 4. The prompt describes what should appear in the masked region

Common editing tasks: add/remove elements, change style, adjust colors, replace backgrounds, etc.

Prompt Handling

For generation: Pass user's image description as-is to --prompt. Only rework if clearly insufficient.

For editing: Pass editing instructions in --prompt (e.g., "add a rainbow in the sky", "make it look like a watercolor painting")

Preserve user's creative intent in both cases.

Output

  • Saves PNG to current directory (or specified path if filename includes directory)
  • Script outputs the full path to the generated image
  • Do not read the image back - just inform the user of the saved path

Examples

Generate new image:

uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "A serene Japanese garden with cherry blossoms" --filename "2025-12-17-14-23-05-japanese-garden.png" --quality high --size 1536x1024

Generate with transparent background:

uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "A cute cartoon cat mascot" --filename "2025-12-17-14-25-30-cat-mascot.png" --background transparent --quality high

Edit existing image (full image):

uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "make the sky more dramatic with storm clouds" --filename "2025-12-17-14-27-00-dramatic-sky.png" --input-image "original-photo.jpg"

Edit with mask (inpainting):

uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "a flamingo swimming" --filename "2025-12-17-14-30-00-lounge-flamingo.png" --input-image "lounge.png" --mask "mask.png"

Related skills

How it compares

Use gpt-image-1-5 for OpenAI-native image gen and inpainting from the agent terminal; use design tools for manual pixel editing.

FAQ

What sizes does gpt-image-1-5 support?

gpt-image-1-5 supports 1024x1024 square, 1024x1536 portrait, 1536x1024 landscape, and auto via the --size flag on generate_image.py. Quality can be set to low, medium, or high.

How does gpt-image-1-5 authenticate?

gpt-image-1-5 reads OPENAI_API_KEY from the environment or accepts --api-key on the command line. If neither is set, generate_image.py exits with an error instead of attempting unsigned requests.

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.