Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
ideacco avatar

Baoyu Image Gen

  • 34 installs
  • 12 repo stars
  • Updated March 4, 2026
  • ideacco/baoyu-skills-openclaw

Baoyu Image Gen is an agent skill that configures default image providers and models via EXTEND.md before you generate assets.

About

Baoyu Image Gen (first-time setup flow) is an agent skill for solo builders who want repeatable AI image generation inside an OpenClaw/Baoyu-style skill pack without re-answering provider questions every session. When no EXTEND.md exists, the skill walks through default provider, model, and preferences in one structured interview; when EXTEND.md exists but default_model for the chosen provider is null, it runs a shorter model-selection-only path. That persistence layer matters for indie workflows that batch covers, UI mocks, and social creatives from the same agent. Use it during Build when you integrate multimodal APIs into content or agent tooling, before any batch image job. The skill is configuration and workflow guidance—it does not replace API keys, billing, or your host app's asset pipeline. After setup completes, continuation flows assume EXTEND.md is the contract for downstream baoyu-image-gen generation steps in the same repository.

  • Two-path setup: full provider + model + preferences when EXTEND.md is missing, or model-only when provider is already sa
  • Single AskUserQuestion call batches all first-time questions to reduce setup friction
  • Default provider options include Google Gemini multimodal (recommended) and OpenAI GPT Image with described tradeoffs
  • Writes or updates EXTEND.md so later image-gen runs skip redundant prompts

Baoyu Image Gen by the numbers

  • 34 all-time installs (skills.sh)
  • Ranked #949 of 1,335 Generative Media skills by installs in the Skillselion catalog
  • Security screen: HIGH risk (skills.sh audit)
  • Data as of Aug 2, 2026 (Skillselion catalog sync)
npx skills add https://github.com/ideacco/baoyu-skills-openclaw --skill baoyu-image-gen

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs34
repo stars12
Security audit2 / 3 scanners passed
Last updatedMarch 4, 2026
Repositoryideacco/baoyu-skills-openclaw

What it does

Run first-time setup for Baoyu image generation—pick Google Gemini or OpenAI and persist defaults in EXTEND.md before generating assets in the agent.

Who is it for?

Best when you're standardizing on Gemini or OpenAI image APIs inside a Baoyu/OpenClaw agent repo.

Skip if: Skip if you manage image models only in a central MLOps console with no per-repo EXTEND.md pattern.

When should I use this skill?

Triggered when no EXTEND.md is found (full setup) or EXTEND.md exists but default_model for the provider is null (model selection only).

What you get

EXTEND.md stores your Google or OpenAI defaults so subsequent image generation continues without repeating full onboarding.

  • EXTEND.md created or updated with default provider and model preferences
  • Completed first-time or model-only AskUserQuestion setup flow

By the numbers

  • Two setup flows: full setup without EXTEND.md versus model-only when provider config already exists

Files

SKILL.mdMarkdownGitHub ↗

Image Generation (AI SDK)

Official API-based image generation. Supports OpenAI, Google, DashScope (阿里通义万象) and Replicate providers.

Script Directory

Agent Execution: 1. SKILL_DIR = this SKILL.md file's directory 2. Script path = ${SKILL_DIR}/scripts/main.ts

Step 0: Load Preferences ⛔ BLOCKING

CRITICAL: This step MUST complete BEFORE any image generation. Do NOT skip or defer.

Check EXTEND.md existence (priority: project → user):

test -f .baoyu-skills/baoyu-image-gen/EXTEND.md && echo "project"
test -f "$HOME/.baoyu-skills/baoyu-image-gen/EXTEND.md" && echo "user"
ResultAction
FoundLoad, parse, apply settings. If default_model.[provider] is null → ask model only (Flow 2)
Not found⛔ Run first-time setup (references/config/first-time-setup.md) → Save EXTEND.md → Then continue

CRITICAL: If not found, complete the full setup (provider + model + quality + save location) using AskUserQuestion BEFORE generating any images. Generation is BLOCKED until EXTEND.md is created.

PathLocation
.baoyu-skills/baoyu-image-gen/EXTEND.mdProject directory
$HOME/.baoyu-skills/baoyu-image-gen/EXTEND.mdUser home

EXTEND.md Supports: Default provider | Default quality | Default aspect ratio | Default image size | Default models

Schema: references/config/preferences-schema.md

Usage

# Basic
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "A cat" --image cat.png

# With aspect ratio
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "A landscape" --image out.png --ar 16:9

# High quality
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "A cat" --image out.png --quality 2k

# From prompt files
npx -y bun ${SKILL_DIR}/scripts/main.ts --promptfiles system.md content.md --image out.png

# With reference images (Google multimodal or OpenAI edits)
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "Make blue" --image out.png --ref source.png

# With reference images (explicit provider/model)
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "Make blue" --image out.png --provider google --model gemini-3-pro-image-preview --ref source.png

# Specific provider
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "A cat" --image out.png --provider openai

# DashScope (阿里通义万象)
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "一只可爱的猫" --image out.png --provider dashscope

# Replicate (google/nano-banana-pro)
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "A cat" --image out.png --provider replicate

# Replicate with specific model
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "A cat" --image out.png --provider replicate --model google/nano-banana

Options

OptionDescription
--prompt <text>, -pPrompt text
--promptfiles <files...>Read prompt from files (concatenated)
--image <path>Output image path (required)
`--provider google\openai\
--model <id>, -mModel ID (Google: gemini-3-pro-image-preview, gemini-3.1-flash-image-preview; OpenAI: gpt-image-1.5)
--ar <ratio>Aspect ratio (e.g., 16:9, 1:1, 4:3)
--size <WxH>Size (e.g., 1024x1024)
`--quality normal\2k`
`--imageSize 1K\2K\
--ref <files...>Reference images. Supported by Google multimodal (gemini-3-pro-image-preview, gemini-3-flash-preview, gemini-3.1-flash-image-preview) and OpenAI edits (GPT Image models). If provider omitted: Google first, then OpenAI
--n <count>Number of images
--jsonJSON output

Environment Variables

VariableDescription
OPENAI_API_KEYOpenAI API key
GOOGLE_API_KEYGoogle API key
DASHSCOPE_API_KEYDashScope API key (阿里云)
REPLICATE_API_TOKENReplicate API token
OPENAI_IMAGE_MODELOpenAI model override
GOOGLE_IMAGE_MODELGoogle model override
DASHSCOPE_IMAGE_MODELDashScope model override (default: z-image-turbo)
REPLICATE_IMAGE_MODELReplicate model override (default: google/nano-banana-pro)
OPENAI_BASE_URLCustom OpenAI endpoint
GOOGLE_BASE_URLCustom Google endpoint
DASHSCOPE_BASE_URLCustom DashScope endpoint
REPLICATE_BASE_URLCustom Replicate endpoint

Load Priority: CLI args > EXTEND.md > env vars > <cwd>/.baoyu-skills/.env > ~/.baoyu-skills/.env

Model Resolution

Model priority (highest → lowest), applies to all providers:

1. CLI flag: --model <id> 2. EXTEND.md: default_model.[provider] 3. Env var: <PROVIDER>_IMAGE_MODEL (e.g., GOOGLE_IMAGE_MODEL) 4. Built-in default

EXTEND.md overrides env vars. If both EXTEND.md default_model.google: "gemini-3-pro-image-preview" and env var GOOGLE_IMAGE_MODEL=gemini-3.1-flash-image-preview exist, EXTEND.md wins.

Agent MUST display model info before each generation:

  • Show: Using [provider] / [model]
  • Show switch hint: Switch model: --model <id> | EXTEND.md default_model.[provider] | env <PROVIDER>_IMAGE_MODEL

Replicate Models

Supported model formats:

  • owner/name (recommended for official models), e.g. google/nano-banana-pro
  • owner/name:version (community models by version), e.g. stability-ai/sdxl:<version>

Examples:

# Use Replicate default model
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "A cat" --image out.png --provider replicate

# Override model explicitly
npx -y bun ${SKILL_DIR}/scripts/main.ts --prompt "A cat" --image out.png --provider replicate --model google/nano-banana

Provider Selection

1. --ref provided + no --provider → auto-select Google first, then OpenAI, then Replicate 2. --provider specified → use it (if --ref, must be google, openai, or replicate) 3. Only one API key available → use that provider 4. Multiple available → default to Google

Quality Presets

PresetGoogle imageSizeOpenAI SizeUse Case
normal1K1024pxQuick previews
2k (default)2K2048pxCovers, illustrations, infographics

Google imageSize: Can be overridden with --imageSize 1K|2K|4K

Aspect Ratios

Supported: 1:1, 16:9, 9:16, 4:3, 3:4, 2.35:1

  • Google multimodal: uses imageConfig.aspectRatio
  • Google Imagen: uses aspectRatio parameter
  • OpenAI: maps to closest supported size

Generation Mode

Default: Sequential generation (one image at a time). This ensures stable output and easier debugging.

Parallel Generation: Only use when user explicitly requests parallel/concurrent generation.

ModeWhen to Use
Sequential (default)Normal usage, single images, small batches
ParallelUser explicitly requests, large batches (10+)

Parallel Settings (when requested):

SettingValue
Recommended concurrency4 subagents
Max concurrency8 subagents
Use caseLarge batch generation when user requests parallel

Agent Implementation (parallel mode only):

# Launch multiple generations in parallel using Task tool
# Each Task runs as background subagent with run_in_background=true
# Collect results via TaskOutput when all complete

Error Handling

  • Missing API key → error with setup instructions
  • Generation failure → auto-retry once
  • Invalid aspect ratio → warning, proceed with default
  • Reference images with unsupported provider/model → error with fix hint (switch to Google multimodal: gemini-3-pro-image-preview, gemini-3.1-flash-image-preview; or OpenAI GPT Image edits)

Extension Support

Custom configurations via EXTEND.md. See Preferences section for paths and supported options.

OpenClaw Migration Notes

  • Migration mode: OpenClaw native first
  • Preserved directories:
  • scripts: yes
  • references: yes
  • prompts: no
  • External dependencies: bun, external-ai-api
  • Risk level: medium
  • Environment variables (detected): API, DASHSCOPE_API_KEY, DASHSCOPE_BASE_URL, GOOGLE_API_KEY, GOOGLE_BASE_URL, HOME, OPENAI_API_KEY, OPENAI_BASE_URL, REPLICATE_API_TOKEN, REPLICATE_BASE_URL
  • Config compatibility: keep original .baoyu-skills/<skill>/EXTEND.md behavior and add OpenClaw-compatible path in runtime wrappers when needed.

OpenClaw Preflight Checks

  • Confirm bun is available when script examples use npx -y bun ....
  • Confirm required API credentials are exported before execution.

Detected env vars: API, DASHSCOPE_API_KEY, DASHSCOPE_BASE_URL, GOOGLE_API_KEY, GOOGLE_BASE_URL, HOME, OPENAI_API_KEY, OPENAI_BASE_URL, REPLICATE_API_TOKEN, REPLICATE_BASE_URL

OpenClaw Failure Fallback

  • If runtime dependency is missing, stop execution and return exact install/setup command.
  • If API call fails, suggest retry with explicit provider/model flags and validate key scope.

Related skills

How it compares

Skill-side provider onboarding, not a hosted image CDN or standalone design app.

FAQ

Who is baoyu-image-gen for?

Developers using Baoyu skills on OpenClaw who need a one-time (or model-only) setup before automated image generation from their coding agent.

When should I use baoyu-image-gen?

At the start of Build integrations when adding AI image output, or the first time you invoke generation and EXTEND.md or default_model is missing.

Is baoyu-image-gen safe to install?

Setup touches provider choice and local EXTEND.md only; confirm API key handling in your environment and review the Security Audits panel on this Prism page.

Generative Mediallmautomation

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.