
Gpt Image 2 Prompting
- 55 installs
- 46 repo stars
- Updated May 31, 2026
- zhouwei713/gpt-image-2-prompting-skill
Helps with ai & agent building tasks.
About
gpt-image-2-prompting is a Claude Code skill for ai & agent building. It helps solo builders move faster with AI-assisted coding.
- gpt-image-2-prompting
- AI & Agent Building
- AI-coding skill
Gpt Image 2 Prompting by the numbers
- 55 all-time installs (skills.sh)
- +2 installs in the week ending Jul 28, 2026 (Skillselion tracking)
- Ranked #6,846 of 16,546 AI & Agent Building skills by installs in the Skillselion catalog
- Data as of Aug 2, 2026 (Skillselion catalog sync)
npx skills add https://github.com/zhouwei713/gpt-image-2-prompting-skill --skill gpt-image-2-promptingAdd your badge
Show developers this skill is listed on Skillselion. Paste this into your README.
| Installs | 55 |
|---|---|
| repo stars | ★ 46 |
| Last updated | May 31, 2026 |
| Repository | zhouwei713/gpt-image-2-prompting-skill ↗ |
What it does
Helps with ai & agent building tasks.
Files
GPT-Image-2 Prompting
This skill turns vague image requests into production-grade GPT-Image-2 prompts.
Core principle: A strong image prompt should read like a visual brief, not a pile of style words.
When to use
Trigger this skill whenever the user:
- asks for GPT-Image-2 prompts or better image prompts
- wants to improve a weak prompt
- wants prompts for posters, covers, UI, dashboards, information graphics, packaging, editorial layouts, concept art, or worldbuilding
- asks for prompt templates, batch prompt ideas, themed prompt packs, or reusable prompt systems
- wants funny, weird, meme-like, absurd, tabloid-style, or viral image prompts
- wants prompts for celebrity mashups, fake news scenes, cursed images, bizarre street photos, surveillance-camera style images, or humorous historical/modern contrast
- says things like “帮我写提示词”, “优化提示词”, “给我一组出图更高级的 prompt”, “做成系列提示词”, “来点搞怪的”, “整点抽象图”, “做成会传播的图”
What this skill optimizes for
1. Strong image type definition 2. Clear visual hierarchy 3. Specific layout and information modules 4. Better conversion of abstract taste words into concrete visual directions 5. Reusable prompt systems instead of one-off lines 6. Strong contrast, absurdity, and meme potential when the goal is funny or viral content 7. Real-photo plausibility for fake-news, candid, paparazzi, and surveillance-style images
Default workflow
Step 1: Identify the image type
Always decide the output format first. Common image types:
- poster
- UI screen / app screen
- dashboard
- infographic
- editorial / magazine cover
- concept art
- design proposal board
- archive sheet / dossier
- map / guide
- packaging / product visual
- storyboard / character sheet
If the user does not specify, infer from context.
Step 2: Build the prompt around 8 slots
Use this order whenever possible:
1. Image type 2. Core subject 3. Composition / layout 4. Supporting modules 5. Visual tone 6. Material / texture 7. Typography / labeling 8. Aspect ratio
Step 3: Translate vague taste words
Do not leave words like these unexplained:
- 高级感
- 电影感
- 氛围感
- 科技感
- 杂志感
Translate them into concrete visual instructions. Examples:
- 高级感 -> restrained layout, limited color palette, clean typography, premium materials, generous negative space
- 电影感 -> low-angle framing, dramatic lighting, foreground/background depth, emotional tension, reflective surfaces
- 科技感 -> glass panels, metal textures, interface modules, cool lighting, precision spacing
- 杂志感 -> editorial hierarchy, strong headline placement, clean grid, short support text, controlled palette
Step 4: Add visual organization
Most weak prompts only name a subject. Strong prompts also define the surrounding structure. Useful supporting modules:
- comment section
- parameter panel
- legend
- annotations
- scale bar
- charts
- ranking list
- route map
- labels
- footer note
- source note
- time axis
Step 5: Decide whether this should be a single image or a system
If the user wants stronger output, consider turning one prompt into a repeatable system:
- same structure, different era
- same structure, different city
- same structure, different profession
- same structure, different emotion
- same structure, different product line
This usually produces better series content than writing unrelated prompts.
Step 6: If the goal is funny or viral, add a contrast engine
For quirky, humorous, or shareable prompts, deliberately introduce one or more of these contrast patterns:
- serious person in a trivial daily-life scene
- historical or mythic figure in a modern low-stakes setting
- luxury visual language applied to ordinary places
- fake-news realism applied to absurd events
- surveillance, paparazzi, flash-photo, or accidental-candid framing
- animals behaving like professionals while humans act normal
Useful realism cues:
- candid photo
- flash photography
- paparazzi shot
- surveillance camera still
- tabloid photo
- awkward timing
- accidental masterpiece
- bizarre realism
Prompt writing rules
- Lead with the image type, not the style word
- Prefer concrete layout over generic taste language
- Use style words only after structural decisions are clear
- Include only details that affect the final image
- Keep the prompt dense, but organized
- If the user wants a shareable or viral image, favor contrast, worldbuilding, or system design
- If the user wants a polished commercial image, favor hierarchy, materials, labels, and composition control
Default response format
When writing prompts for the user, use this structure unless they ask for something else:
1. Brief idea
One sentence describing the concept.
2. Final prompt
A polished prompt in the user’s language.
3. Why this works
2-4 bullets explaining the structural strengths.
4. Optional variations
3 short variant directions if useful.
5. Optional image generation offer
After the final prompt output is complete, check whether the current environment exposes an image generation tool such as image_generate or an equivalent configured image model.
If image generation is available and the user did not already ask you to generate the image immediately, add one concise closing question: “要不要我直接用这条 Prompt 帮你生成一张图?”
Do not generate the image automatically unless the user explicitly asks for generation or answers yes. If the user says yes, call the image generation tool with the finalized prompt. Use the aspect ratio specified in the prompt when it maps cleanly to the tool’s supported ratios; otherwise infer the closest supported ratio and mention the choice briefly.
If no image generation tool/model is available in the environment, do not pretend generation is possible. Just provide the prompt, or say the prompt can be copied into the user’s image model if relevant.
Chinese default output style
When the user writes in Chinese, default to Chinese output unless they request another language.
Use this delivery structure by default:
1. 核心创意
Use 1-2 sentences to explain the visual concept in plain Chinese.
2. 完整 Prompt
Give one polished final prompt, usually as one continuous paragraph.
3. 为什么这样写
Explain briefly using 2-4 bullets. Focus on:
- 图像类型是否清楚
- 主体和版式是否明确
- 信息模块是否增强了层级
- 抽象审美词是否被翻译成了具体视觉语言
4. 可改写方向
Offer 2-3 short variation directions when useful, such as:
- 换时代 / 换城市
- 换职业 / 换情绪
- 换媒介形态(海报、UI、信息图、档案页)
- 换色彩和材质系统
5. 是否直接生图
如果当前环境配置了可用的生图工具或生图模型,例如 Hermes 的 image_generate 工具,在完成 Prompt 输出后,用一句话询问用户是否要直接生成图片: “要不要我直接用这条 Prompt 帮你生成一张图?”
只有在用户明确要求“直接生成/帮我出图/生成一张”或用户回答确认后,才调用生图工具。不要在用户只要求写 Prompt 时自动生图。
如果环境没有可用的生图能力,不要追加这个确认问题,也不要暗示可以在当前环境直接生成。
Chinese phrasing guidance
- Write naturally, like a skilled Chinese visual director giving a brief
- Keep explanations concise and practical
- Avoid jargon overload unless the user clearly wants professional terminology
- Prefer concrete Chinese visual instructions over vague taste words
- If the user asks for a batch of prompts, keep the same output format for each prompt unless a table is clearly better
Chinese batch mode
When the user asks for multiple prompts at once, default to a structured Chinese batch format.
Use this output structure:
分类标题
Name the category first, especially when the batch is large. Examples:
- 历史世界 × 现代界面
- 信息图 × 情绪表达
- 品牌提案 × 商业视觉
- 世界观档案页
- 城市观察 × 社会情绪
每条 Prompt 的默认格式
For each prompt, use:
[编号]. [标题]
- 核心创意:用一句话讲清楚这个画面想做什么
- 完整 Prompt:给一条可直接使用的完整 Prompt
- 可替换变量:列 2-4 个最值得替换的变量,方便用户自己扩写
Batch-size guidance
- 1-5 条:可以保留“为什么这样写”
- 6-20 条:以“核心创意 + 完整 Prompt + 可替换变量”为主,保持紧凑
- 20 条以上:优先按分类分组,减少逐条解释,除非用户明确要求详细拆解
Default consistency rules
- 同一批内容尽量保持统一结构
- 同一分类优先共享一套视觉逻辑,再替换主题变量
- 如果是做资料包,优先保证可复制、可扩写、可分类整理
- 如果是做灵感清单,可以适当放宽结构,但仍然保留图像类型和版式意识
Upgrade path for weak prompts
If the user gives a weak prompt, rewrite it by upgrading in this order: 1. clarify image type 2. define subject 3. define composition 4. add modules 5. translate vague taste words 6. add materials / labels / typography 7. choose aspect ratio
Chinese example response
核心创意
做一张“城市失眠指数”主题的信息图海报,把深夜情绪转成可以被观看的数据地图。
完整 Prompt
“城市失眠指数”信息图海报,中心是一张俯视夜景地图,按照 22:00、00:00、02:00、04:00 四个时段分层发光,商业区高亮,住宅区以昏黄窗光点阵呈现,四周嵌入咖啡销量、夜间打车热度、音乐播放峰值、社交媒体活跃度四个圆形数据模块,标题像国际杂志专题页,副标题写“谁还醒着,谁在假装睡着”,整体冷静克制,带轻微都市焦虑感,色彩控制在霓虹紫、电蓝、夜黑和少量暖黄,比例 4:5。
为什么这样写
- 先定义为“信息图海报”,模型会按专题视觉去组织画面
- 中心地图 + 四周模块,画面层级会更清楚
- “都市焦虑感”被拆成夜景、冷色、高亮区域这些具体视觉指令
- 配色和比例明确后,更容易出成品感
可改写方向
- 换成“加班热力图”主题
- 换成“地铁情绪地图”结构
- 换成杂志封面而不是信息图海报
Chinese batch example
历史世界 × 现代界面
1. 唐朝人的外卖 App
- 核心创意:把唐朝日常生活翻译成一张现代外卖首页界面,让历史感和产品感同时成立。
- 完整 Prompt:唐朝人的外卖 App,画面模拟手机外卖首页界面,顶部定位显示“长安·平康坊”,推荐位展示胡饼、炙羊肉、葡萄酿,商家头像采用工笔画掌柜半身像,评分以铜钱图标呈现,底部导航栏完整保留现代产品结构,状态栏显示“大唐信号满格”和“开元二十四年”,整体配色为赭石、石绿、金箔红,字体融合碑刻感标题字与细无衬线,界面像真实产品设计稿,同时带有历史穿越幽默感,比例 9:16。
- 可替换变量:朝代、城市、品类、界面类型
2. 明代人的求职网站首页
- 核心创意:把古代职业体系做成现代招聘平台首页,重点放在职位卡片和筛选逻辑。
- 完整 Prompt:明代人的求职网站首页,历史服饰与现代招聘平台界面融合,主界面有职位推荐“修史官、书院讲郎、船务账房、织造监督、宫廷画师”,筛选条件写“擅诗文、通算学、善骑射、可外派”,人物简历卡使用工笔肖像,薪资单位以俸禄石米表示,整体排版像真实招聘网站首页,比例 16:9。
- 可替换变量:时代、岗位、筛选条件、薪资表达
Good prompt categories
Use the reference file references/categories.md for category patterns and representative examples. Use references/templates.md for reusable fill-in-the-blank templates. Use references/examples.md for polished baseline examples. Use references/quirky-funny-100-prompts.md for funny, bizarre, celebrity-contrast, fake-news, and meme-ready prompt ideas. Use references/quirky-funny-100-prompts-part2.md for more chaotic, viral, surveillance-style, and absurd-realism prompt ideas.
Common mistakes to avoid
- starting with only style words
- leaving composition undefined
- no information hierarchy
- too many random elements with no focal point
- calling something “高级感” without specifying what creates that feeling
- writing a one-off idea when a reusable prompt system would be better
Best-practice mindset
The goal is not to produce a merely detailed prompt. The goal is to produce a prompt with visual intent, structure, and enough control that the output feels designed.
# OS
.DS_Store
Thumbs.db
# Editors
.vscode/
.idea/
*.swp
*.swo
# Python cache (if local validation scripts are added later)
__pycache__/
*.pyc
# Temporary files
*.tmp
*.temp
*.log
Changelog
v0.1.0
First documented stable release for the public skill package.
Included:
- Core GPT-Image-2 prompting skill in
SKILL.md - Bilingual
README.mdwith Codex and Hermes usage guidance - Reference files for templates, categories, examples, and prompt packs
CONTRIBUTING.mdfor future improvementsSECURITY.mdfor report handling and maintainer response
Contributing
Thanks for your interest in improving this skill.
What to contribute
Helpful contributions include:
- Better prompting structures for GPT-Image-2
- More reusable prompt templates
- Stronger Chinese and bilingual examples
- Clearer documentation and installation guidance
- Fixes for formatting, wording, or broken references
Before you submit
Please keep these principles in mind:
- Prioritize structure over keyword stuffing
- Make prompts production-oriented and reusable
- Preserve the Chinese-first usability of the skill
- Keep examples concrete, high-signal, and easy to adapt
Recommended workflow
1. Fork the repository 2. Create a feature branch 3. Make focused changes 4. Verify links, filenames, and formatting 5. Open a pull request with a clear summary
Content guidelines
When editing templates or examples:
- Explain the visual goal clearly
- Include composition, hierarchy, modules, materials, and aspect ratio when relevant
- Prefer prompts that can generalize into a system, not just one-off inspiration
- Keep tone practical and implementation-friendly
Pull request checklist
- [ ] The change is focused and easy to review
- [ ] Markdown renders correctly
- [ ] File paths and references are valid
- [ ] New examples add distinct value instead of repeating existing patterns
- [ ] Chinese wording is natural and clear
Issue reports
If you find a problem, please include:
- What file or section is affected
- What is unclear or incorrect
- A suggested revision if possible
License
By contributing, you agree that your contributions will be licensed under the repository's MIT License.
MIT License
Copyright (c) 2026 Hermes Agent
Permission is hereby granted, free of charge, to any person obtaining a copy
of this software and associated documentation files (the "Software"), to deal
in the Software without restriction, including without limitation the rights
to use, copy, modify, merge, publish, distribute, sublicense, and/or sell
copies of the Software, and to permit persons to whom the Software is
furnished to do so, subject to the following conditions:
The above copyright notice and this permission notice shall be included in all
copies or substantial portions of the Software.
THE SOFTWARE IS PROVIDED "AS IS", WITHOUT WARRANTY OF ANY KIND, EXPRESS OR
IMPLIED, INCLUDING BUT NOT LIMITED TO THE WARRANTIES OF MERCHANTABILITY,
FITNESS FOR A PARTICULAR PURPOSE AND NONINFRINGEMENT. IN NO EVENT SHALL THE
AUTHORS OR COPYRIGHT HOLDERS BE LIABLE FOR ANY CLAIM, DAMAGES OR OTHER
LIABILITY, WHETHER IN AN ACTION OF CONTRACT, TORT OR OTHERWISE, ARISING FROM,
OUT OF OR IN CONNECTION WITH THE SOFTWARE OR THE USE OR OTHER DEALINGS IN THE
SOFTWARE.
GPT-Image-2 Prompting Skill
中文 | English
A high-quality prompting skill for GPT-Image-2 that turns vague image ideas into production-grade prompts with clear structure, visual hierarchy, and reusable prompt systems.
一个面向 GPT-Image-2 的高质量提示词 Skill。 它的目标不是“多写几个风格词”,而是把模糊的图像想法升级成更像视觉总监 brief 的生产级 Prompt。
Highlights
- Structured prompting instead of keyword piles
- Better control over layout, hierarchy, modules, materials, and aspect ratio
- Strong Chinese-first output mode
- Batch prompt generation for prompt packs and content libraries
- Reusable templates, categories, and examples included
这个 Skill 特别适合下面这些场景:
- 你想为 GPT-Image-2 写出更高级、更稳定的提示词
- 你只有一个模糊概念,想把它变成真正可用的图像 Prompt
- 你想批量生成一组风格统一、结构稳定的 Prompt
- 你在做海报、UI、信息图、杂志封面、概念图、品牌视觉、世界观设定图
- 你想把“高级感、电影感、科技感”这类空泛描述,翻译成真正可执行的视觉语言
---
Quick Start
If you want to improve this repository, see CONTRIBUTING.md.
Codex
1. Clone this repository or download it as a zip file 2. Copy the repository folder into your Codex skills directory, for example $CODEX_HOME/skills/gpt-image-2-prompting-skill 3. Start a new Codex session so the skill can be discovered 4. Ask Codex to use the skill for image prompting tasks, for example:
- use gpt-image-2-prompting-skill to improve this image prompt
- use gpt-image-2-prompting-skill to create a poster prompt pack
- use gpt-image-2-prompting-skill to turn this rough idea into a production prompt
- use gpt-image-2-prompting-skill to generate 20 structured GPT-Image-2 prompts
If your Codex setup uses a custom skills path, place the folder in that configured path and keep SKILL.md at the repository root.
Hermes
1. Put this skill folder into your Hermes skills directory 2. Ensure Hermes can discover the skill 3. Ask Hermes things like:
- 帮我写一个 GPT-Image-2 Prompt
- 优化一下这个图片 Prompt
- 给我 10 条未来城市风 Prompt
- 把这个模糊想法变成可直接用的 Prompt
Expected output
For single prompts, the default Chinese structure is:
- 核心创意
- 完整 Prompt
- 为什么这样写
- 可改写方向
For batch prompts, the default structure is:
- 分类标题
- 标题
- 核心创意
- 完整 Prompt
- 可替换变量
---
中文
这个 Skill 能做什么
这个 Skill 的核心能力有 5 个:
1. 把模糊需求变成结构化 Prompt 2. 自动补全图像类型、构图、信息模块、材质、色彩、比例这些关键要素 3. 把“高级感 / 电影感 / 科技感 / 杂志感”翻译成更具体的视觉指令 4. 让单条 Prompt 更像“设计稿说明”而不是“关键词堆砌” 5. 支持单条输出和批量输出两种模式
简单说,它不是“灵感词库”,而是一套 Prompt 生成方法。
---
这个 Skill 适合哪些图像类型
它最适合下面这些图像任务:
- 海报(Poster)
- UI 页面 / App 页面
- Dashboard / 数据界面
- 信息图(Infographic)
- 杂志封面 / Editorial 视觉
- 概念图(Concept Art)
- 品牌提案图 / 包装提案图
- 档案页 / Dossier / 世界观设定页
- 地图 / 导览图 / 路线图
- 产品视觉 / 包装视觉
- 分镜板 / 角色设定页
如果用户没有明确说图像类型,Skill 会优先根据上下文推断。
---
它和普通 Prompt 写法的区别
普通写法通常像这样:
- 帮我生成一张未来城市风海报
- 帮我做一个科技感 UI
- 帮我做一张高级感图片
这种写法也能出图,但很容易飘。 因为模型需要自己猜:
- 这到底是什么类型的图
- 主体是什么
- 版式怎么排
- 什么地方该突出
- 什么地方该留白
这个 Skill 的做法不一样。 它会优先把 Prompt 组织成下面这 8 个槽位:
1. 图像类型 2. 核心主体 3. 构图 / 版式 4. 辅助模块 5. 视觉气质 6. 材质 / 纹理 7. 标题 / 标签 / 文字系统 8. 画幅比例
也就是说,它关注的是“画面组织能力”,不只是风格描述。
---
这个 Skill 的默认工作流
Step 1:先定图像类型
Skill 会先判断这是海报、UI、信息图、品牌提案还是世界观档案页。
Step 2:再搭 Prompt 骨架
Skill 默认使用 8 槽位结构来补足 Prompt。
Step 3:翻译空泛审美词
比如:
- 高级感 -> 留白、克制排版、少量配色、玻璃/金属材质、干净字体
- 电影感 -> 低机位、情绪灯光、前后景层次、反光地面、画面张力
- 科技感 -> 玻璃面板、金属结构、界面模块、冷色照明、精密 spacing
- 杂志感 -> 标题层级、编辑排版、短文案、网格感、留白控制
Step 4:增加信息层级
Skill 会优先补充这些常见模块:
- 评论区
- 参数栏
- 图例
- 注释标签
- 比例尺
- 图表
- 排行榜
- 路线图
- 时间轴
- 页脚说明
Step 5:判断要不要做成“系统”
很多情况下,一条 Prompt 不如一套 Prompt 系统好用。 比如:
- 同一结构,换时代
- 同一结构,换城市
- 同一结构,换职业
- 同一结构,换情绪
- 同一结构,换产品线
---
中文输出能力
这个 Skill 专门增强过中文输出。 如果用户用中文提问,它默认会用中文交付,并且优先按照更适合中文用户理解的方式输出。
单条 Prompt 默认格式
1. 核心创意 2. 完整 Prompt 3. 为什么这样写 4. 可改写方向
批量 Prompt 默认格式
当用户一次要多条 Prompt 时,Skill 会自动切换到“中文批量模式”。
默认结构:
- 分类标题
- 每条 Prompt 的标题
- 核心创意
- 完整 Prompt
- 可替换变量
它还会根据数量自动调整细节密度:
- 1-5 条:可以保留“为什么这样写”
- 6-20 条:以“核心创意 + 完整 Prompt + 可替换变量”为主
- 20 条以上:优先分组,减少逐条解释,方便做资料包
---
这个 Skill 解决的典型问题
它特别适合解决下面这些问题:
- Prompt 太短,信息不够,出图随机
- Prompt 只有风格词,没有结构
- 用户说了“高级感”,但没有可执行的视觉指令
- 批量写 Prompt 时,风格前后不统一
- 需要做海报、信息图、品牌提案时,不知道该怎么组织画面
- 想把一个模糊概念写成真正有成品感的 Prompt
---
内置参考内容
这个 Skill 自带 5 份参考文件:
1. references/templates.md
可复用 Prompt 模板库,适合快速扩写。 包括:
- 历史世界 × 现代界面
- 情绪 × 数据可视化海报
- 人物/职业 × 静物叙事
- 空间 × 设计提案图
- 世界观 × 档案页
- 日常事件 × 电影海报
- 年鉴封面 × 总结图
2. references/categories.md
把整套 Prompt 方法归纳成 10 个大类,方便快速匹配任务类型。
3. references/examples.md
代表性示例集合,适合直接借鉴或做结构映射。
4. references/quirky-funny-100-prompts.md
偏机灵古怪、反差感、假新闻感、名人乱入日常的 100 条提示词,适合做传播图、玩梗图、搞怪图。
5. references/quirky-funny-100-prompts-part2.md
第二批更偏发疯感、监控截图感、社交媒体 meme 感、职场抽象感的 100 条提示词,适合继续扩写成系列内容。
---
目录结构
gpt-image-2-prompting/
├── SKILL.md
├── README.md
└── references/
├── templates.md
├── categories.md
├── examples.md
├── quirky-funny-100-prompts.md
└── quirky-funny-100-prompts-part2.md---
使用示例
示例 1:用户只有一个很模糊的需求
用户: “帮我写一个未来城市风的图片提示词。”
Skill 会把它升级成更完整的输出,例如:
- 核心创意:未来城市夜景概念海报,重点放在立体交通和发光建筑系统
- 完整 Prompt:一条包含主体、构图、信息层级、材质、色彩和比例的完整 Prompt
- 为什么这样写:解释为什么这个结构更容易出精品
- 可改写方向:清晨版、交通中枢版、俯视地图版
示例 2:用户想批量写 Prompt
用户: “给我 20 条未来城市风 Prompt。”
Skill 会优先:
- 先按类别分组
- 每条保留“核心创意 + 完整 Prompt + 可替换变量”
- 控制整体结构统一,方便整理成资料包
示例 3:用户想优化已有 Prompt
用户: “我现在这条 Prompt 太普通了,帮我优化一下。”
Skill 会按这个顺序改写: 1. 定图像类型 2. 定主体 3. 定构图 4. 补信息模块 5. 翻译空泛审美词 6. 补材质 / 标签 / 字体 7. 定比例
---
最适合谁
这个 Skill 最适合:
- 经常写图片 Prompt 的人
- 做内容、做封面、做海报的人
- 做概念图、世界观设定图的人
- 做品牌视觉、提案图、信息图的人
- 想系统学 Prompt,而不是只想抄几句的人
---
安装 / 使用方式
如果你在 Codex 里使用这个 Skill:
1. 克隆这个仓库,或者下载 zip 后解压 2. 把仓库目录放到 Codex 的 skills 目录下,例如 $CODEX_HOME/skills/gpt-image-2-prompting-skill 3. 新开一个 Codex 会话,让 Codex 重新发现 Skill 4. 在图像 Prompt 任务里直接点名使用它
你可以这样说:
- use gpt-image-2-prompting-skill to improve this image prompt
- use gpt-image-2-prompting-skill to create a poster prompt pack
- use gpt-image-2-prompting-skill to turn this rough idea into a production prompt
- use gpt-image-2-prompting-skill to generate 20 structured GPT-Image-2 prompts
如果你的 Codex 使用自定义 skills 路径,把目录放到对应路径即可,并确保 SKILL.md 位于仓库根目录。
如果你在 Hermes 里使用这个 Skill:
1. 把技能目录放到你的 skills 目录下 2. 确保 Hermes 可以发现该 Skill 3. 在提到图像 Prompt 相关任务时,让 Hermes 调用它
例如你可以直接这样说:
- 帮我写一个 GPT-Image-2 Prompt
- 优化一下这个图片 Prompt
- 做一组海报风 Prompt
- 按资料包结构给我 20 条 Prompt
- 把这个模糊想法变成可直接用的 Prompt
- 给我 20 条搞怪、容易传播的 Prompt
- 做一组假新闻感 / 偷拍感 / 监控截图感的图片 Prompt
- 帮我把这条图像想法改得更抽象、更好笑一点
---
设计原则
这个 Skill 的核心设计原则只有一句话:
A strong image prompt should read like a visual brief, not a pile of style words.
翻成中文就是: 一条好 Prompt,读起来应该像视觉 brief,而不是一串风格词。
---
许可
本 Skill 当前默认采用仓库作者自行指定的开源或共享方式。 如果你准备公开发布到 GitHub,建议在仓库中补充明确的 LICENSE 文件。
---
English
What this skill does
This is a high-quality prompting skill for GPT-Image-2. Its goal is not to generate random style-heavy prompts, but to turn vague image ideas into production-grade prompts that read like visual briefs.
It is especially useful when you want to:
- write better GPT-Image-2 prompts
- upgrade a weak or overly generic prompt
- generate batches of prompts with consistent quality
- create prompts for posters, UI screens, infographics, editorial layouts, concept art, brand visuals, or worldbuilding images
- translate vague taste words like “premium”, “cinematic”, or “techy” into concrete visual instructions
---
Core capabilities
This skill is designed to do five things well:
1. Turn vague requests into structured prompts 2. Automatically complete key missing pieces such as image type, composition, support modules, materials, palette, and aspect ratio 3. Translate abstract aesthetic words into practical visual language 4. Produce prompts that feel closer to design briefs than keyword piles 5. Support both single-prompt and batch-prompt workflows
In short, this is not just a prompt collection. It is a prompting method.
---
Best-fit image types
This skill works best for:
- posters
- UI / app screens
- dashboards
- infographics
- editorial layouts / magazine covers
- concept art
- design proposal boards
- dossier / archive sheets
- maps / guides
- packaging / product visuals
- storyboards / character sheets
If the user does not explicitly specify the format, the skill infers it from context.
---
How it differs from ordinary prompting
A typical weak prompt looks like this:
- create a futuristic city poster
- make a tech-style UI
- generate a premium image
That can still produce an image, but the result often drifts because the model has to guess:
- what kind of image this actually is
- what the focal point should be
- how the layout should be structured
- what should be prominent and what should recede
This skill uses a more controlled structure. It builds prompts around 8 slots:
1. image type 2. core subject 3. composition / layout 4. supporting modules 5. visual tone 6. materials / textures 7. typography / labels 8. aspect ratio
That means it optimizes for visual organization, not just style naming.
---
Default workflow
Step 1: Define the image type
The skill first decides whether the request is a poster, interface, infographic, editorial cover, proposal board, or dossier-like image.
Step 2: Build the prompt skeleton
It uses the 8-slot structure above to complete the prompt.
Step 3: Translate vague taste language
Examples:
- premium -> restrained layout, limited palette, clean typography, premium materials, negative space
- cinematic -> low-angle framing, dramatic lighting, foreground/background depth, reflective surfaces, emotional tension
- futuristic -> interface modules, glass surfaces, metal textures, cool light, precision spacing
- editorial -> clear headline hierarchy, grid logic, support text, layout control
Step 4: Add information hierarchy
The skill often adds support modules such as:
- comments
- parameter bars
- legends
- annotation labels
- charts
- timelines
- rankings
- maps
- footer notes
Step 5: Decide whether the task should become a prompt system
In many cases, a reusable system is more valuable than a single prompt. For example:
- same structure, different era
- same structure, different city
- same structure, different profession
- same structure, different emotion
- same structure, different product line
---
Chinese-first output support
This skill has been specially tuned for Chinese users. When the user asks in Chinese, it defaults to Chinese output unless another language is explicitly requested.
Default single-prompt delivery
- Core idea
- Final prompt
- Why it works
- Variation directions
Default batch-prompt delivery
When the user asks for multiple prompts, the skill switches into a structured Chinese batch mode. It groups prompts by category and uses a stable format for each item.
This makes it suitable for:
- prompt packs
- fan handouts
- content libraries
- internal visual ideation docs
---
Included reference files
This skill ships with five reference files:
references/templates.md
Reusable prompt templates for fast expansion. Includes patterns such as:
- historical world × modern interface
- emotion × infographic poster
- subject × editorial still life
- space × proposal board
- worldbuilding × dossier page
- daily life × cinematic poster
- annual archive × cover system
references/categories.md
A category map that compresses the full prompt library into 10 reusable archetypes.
references/examples.md
Representative example prompts that can be adapted directly or used as structural references.
references/quirky-funny-100-prompts.md
A first batch of 100 funny, bizarre, contrast-heavy, fake-news, and celebrity-in-everyday-life prompts for more shareable images.
references/quirky-funny-100-prompts-part2.md
A second batch of 100 more chaotic prompts focused on surveillance-camera energy, meme realism, workplace absurdity, and viral-photo logic.
---
Directory structure
gpt-image-2-prompting/
├── SKILL.md
├── README.md
└── references/
├── templates.md
├── categories.md
├── examples.md
├── quirky-funny-100-prompts.md
└── quirky-funny-100-prompts-part2.md---
Example use cases
Example 1: The user only has a vague idea
User: “Write a future city style image prompt.”
The skill will expand that into:
- a clear concept statement
- a full production-grade prompt
- a short explanation of why the structure works
- 2-3 variation directions
Example 2: The user wants a batch of prompts
User: “Give me 20 future-city prompts.”
The skill will typically:
- group them by category
- keep the structure stable across prompts
- include core idea + final prompt + replaceable variables
- make the output easier to turn into a prompt pack
Example 3: The user wants prompt optimization
User: “My current prompt feels too weak. Improve it.”
The skill upgrades it in this order: 1. clarify image type 2. define subject 3. define composition 4. add modules 5. translate vague taste words 6. add materials / labels / typography 7. choose aspect ratio
---
Who this skill is for
This skill is a good fit for:
- people who frequently write image prompts
- content creators making covers and visual assets
- concept artists and worldbuilders
- designers creating posters, systems, or proposal visuals
- anyone who wants to learn prompting as a method, not just copy random prompts
---
Installation / usage
If you use this skill inside Hermes:
1. place the skill folder inside your skills directory 2. ensure Hermes can discover the skill 3. invoke it whenever the task involves image prompting
Typical requests:
- write a GPT-Image-2 prompt for me
- improve this image prompt
- create a poster prompt pack
- give me 20 prompts in a structured prompt-pack style
- turn this vague idea into a usable prompt
- give me 20 funny viral prompts
- make this image idea weirder, more meme-like, or more shareable
- write a fake-news-style or paparazzi-style image prompt
---
Design principle
The skill is built around one sentence:
A strong image prompt should read like a visual brief, not a pile of style words.
That principle drives the entire structure of the skill.
---
License
This skill currently uses whatever licensing model the repository author chooses. If you plan to publish it on GitHub, it is strongly recommended to add a clear LICENSE file to the repository.
GPT-Image-2 Prompt Categories and Representative Examples
This file summarizes the 100-prompt library into reusable categories.
A. Historical world × modern interface
Best for: contrast, virality, instant conceptual clarity Pattern: old-world setting + modern product logic + complete UI hierarchy Representative ideas:
- Tang dynasty food delivery app
- Ming dynasty recruitment homepage
- Warring States collaborative document
- Republican-era trending topics page
B. Infographic × emotional narrative
Best for: articles, covers, social content, editorial visuals Pattern: emotional theme translated into maps, charts, modules, and labels Representative ideas:
- city insomnia index
- metro emotion map
- breakfast stall economics
- convenience store color spectrum
C. Brand proposal × commercial visual system
Best for: pitch boards, design inspiration, packaging and retail concepts Pattern: one commercial concept expanded into identity, materials, labels, and environment Representative ideas:
- future wet market wayfinding
- cyberpunk traditional medicine brand proposal
- flower shop color operations manual
- vending machine brand universe
D. Space section × lifestyle storytelling
Best for: high-feel editorial images with narrative depth Pattern: a real or imagined space carries emotional and visual information through layout and objects Representative ideas:
- one person launching into space from a rental room
- old bookstore air cross-section
- seaside shop seasonal schedule
- hot spring inn operations panel
E. Dossier page × worldbuilding
Best for: fantasy, sci-fi, mythology, institutional design language Pattern: subject sheet + observation system + labels + measurements + archive logic Representative ideas:
- qilin containment dossier
- magical plant greenhouse log
- detective case board
- archaeologist desk archive
F. City observation × social emotion
Best for: magazine-like covers and reflective social visuals Pattern: public infrastructure or urban moments recast as structured visual essays Representative ideas:
- airport late-night stillness
- train station farewell information board
- last-train comfort guide
- city morning newspaper front page
G. Retail scene × everyday aesthetics
Best for: lifestyle content, brand feeds, visual essays on daily life Pattern: familiar scenes with stronger hierarchy, order, and semi-editorial framing Representative ideas:
- midnight bakery final batch
- fruit stall visual order
- independent record store weekly picks
- shoe repair bench tools atlas
H. Service design × future systems
Best for: innovation concepts, product/service proposals, visionary UX images Pattern: future-facing systems shown through service journeys, interface logic, and support modules Representative ideas:
- insomnia navigation system
- social-anxiety-friendly cafe ordering redesign
- future children’s room safety system
- lunar daycare schedule
I. Human observation × micro narrative
Best for: emotionally rich images that feel specific and lived-in Pattern: use objects, routines, and found details to tell a larger story Representative ideas:
- writer’s room in 12 objects
- lost-and-found micro museum
- photo studio sample wall
- late-night radio schedule board
J. Almanac / issue cover × summary image
Best for: annual reviews, issue covers, cultural summaries Pattern: multiple framed scenes under one clear editorial umbrella Representative ideas:
- visual almanac cover for a year
- parallel-world graduation cover system
- alternate-choice sticky note wall
- end-of-world supermarket restock board
What makes these prompts strong
1. They name the image format first 2. They define the main subject clearly 3. They specify layout instead of relying only on style words 4. They add supporting modules that create hierarchy 5. They use tone as a control layer, not as the whole prompt 6. They are reusable systems, not just isolated ideas
Recommended prompting move
When the user gives a weak prompt:
- identify the category above
- pick the closest pattern
- rewrite the prompt using one of the template structures
- add 2-5 support modules
- convert vague taste words into specific visual instructions
Representative GPT-Image-2 Examples
Use these examples when the user wants strong reference prompts.
Example 1 — Historical interface
“唐朝人的外卖 App”/“TANG DYNASTY FOOD DELIVERY INTERFACE”,古代生活场景与现代移动产品 UI 融合设计,画面模拟一台手机上的外卖首页界面,顶部定位显示“长安 · 平康坊”,横幅推荐位写“今夜最火:胡饼、炙羊肉、葡萄酿”,商家卡片使用唐代店招与木牌元素重构,头像是工笔画风掌柜半身像,评分以铜钱图标呈现,订单列表里出现“李白已下单:酒一壶”,底部导航栏分别是“首页/食单/跑腿/订单/我的”,状态栏显示“大唐信号满格”与“开元二十四年”,整体配色为赭石、石绿、金箔红,字体混合碑刻感标题字与现代细无衬线,界面既像真实产品设计稿,又有历史穿越幽默感,比例 9:16。
Example 2 — Emotional infographic
“一座城市的失眠指数”/“CITY INSOMNIA ATLAS”,夜生活数据可视化海报,中心是一张俯视视角的城市夜景地图,按照深夜 22:00、00:00、02:00、04:00 四个时段分层发光,商业区以霓虹紫和电蓝高亮,住宅区以昏黄窗光点阵呈现,地图四周嵌入四个圆形数据模块:咖啡销量曲线、夜间打车热度、音乐播放峰值、社交媒体活跃度,标题使用窄体大写英文与中文黑体叠排,副标题写“谁还醒着,谁在假装睡着”,底部小字注释数据来源、时间区间与城市样本数,整体视觉参考《Monocle》+ 科技智库报告风格,氛围介于都市浪漫与轻微焦虑之间,比例 4:5。
Example 3 — Commercial proposal
“未来菜市场导视系统”/“FUTURE WET MARKET WAYFINDING”,一套完整的公共空间视觉识别提案,主画面为改造后的未来感菜市场大厅透视图,顶部悬挂模块化发光导视牌,分区包括“水产/蔬果/熟食/花卉/社区厨房”,色彩系统分别对应深海蓝、黄绿、暖橙、淡紫,左下角放置 pictogram 图标系统,右侧是摊位招牌统一规范、员工围裙、购物篮、价签与地贴的展开图,标题排版像设计院竞标提案,整体兼具生活气和高级公共设计感,材质强调磨砂亚克力、回收金属、潮湿地面的反光,比例 16:9。
Example 4 — Space storytelling
“一个人在出租屋里完成宇宙远航”,超现实室内摄影与科幻概念海报融合,画面主体是一间普通小出租屋,但书桌延伸成飞船控制台,晾衣架像卫星阵列,窗帘缝隙外是星云,床边堆放的纸箱被改造成推进器模块,标题写“MISSION: MAKE DO AND LAUNCH”,副标题用极小字描述‘预算有限的浪漫主义’,整体温柔、孤独又有希望,比例 16:9。
Example 5 — Worldbuilding dossier
“神话生物收容局档案页:麒麟”,机构档案设计与东方式幻想生物设定结合,画面像绝密收容档案,左侧是麒麟正侧背三视图及步态分解,右侧是角、鳞片、鬃毛、蹄印细节与栖息环境数据模块,档案印有级别章、观察日志、食性和禁忌说明,整体兼具神圣感与现代机构冷静气质,比例 16:9。
Example 6 — Future service system
“面向社恐用户的咖啡馆点单系统重设计”,服务设计提案图,主画面展示咖啡馆柜台和自助点单屏新方案,减少语言接触,增加‘默认推荐/无交流取餐/座位偏好/音量区域’选项,右侧展开图标系统、桌卡、取餐牌、等候区动线图,整体极具现实落地感,比例 16:9。
Example 7 — Concept system image
“给猫设计的一座垂直城市”,建筑概念插画与宠物生活方式结合,超高竖版画面中是一座专为猫咪设计的立体城市:攀爬桥、晒太阳平台、透明观景泡泡、自动投喂站、隐藏睡眠舱、抓板立面、鱼形轻轨,每层配以简短功能说明和猫咪活动剪影,色彩梦幻但结构合理,比例 9:16。
Example 8 — Annual cover logic
“把 2026 年做成一本可翻页的视觉年鉴封面”,年度总结海报与编辑设计融合,画面由十二个小窗组成,分别代表一年中最鲜明的视觉记忆:城市热浪、夜跑、咖啡店开门、暴雨、演唱会荧光棒、毕业照、机场告别、海边日落、书桌凌乱、冬天围巾、第一场雪、跨年倒计时,中央标题“2026 / A VISUAL ALMANAC”压住全部画面,副标题极简,整体像一本高端文化年鉴封面,信息丰富、情绪完整,比例 4:5。
How to use these examples
- If the user needs one strong prompt, adapt the closest example
- If the user needs a system, extract the structure and swap variables
- If the user gives a weak prompt, map it to one example and then rebuild it with the same bones
机灵古怪向 100 个 GPT Image 2 提示词(第二批)
这一批更偏:
- 疯一点
- meme 一点
- 更像朋友圈疯传图
- 更像监控截图、假新闻、假纪实、事故现场
- 更适合做传播型内容和搞怪视觉素材
建议追加词:
- candid photo
- surveillance camera still
- flash photography
- tabloid photo
- cursed image
- absurd realism
- awkward composition
- chaotic scene
- viral internet photo
- accidental masterpiece
---
A. 名人彻底失控类(1-10)
1. 爱因斯坦在电梯里被静电电到头发彻底炸开
拥挤办公楼电梯里,爱因斯坦穿着西装抱着公文包,电梯按钮面板不断放电,他的头发被电得像爆炸一样,旁边上班族努力假装没看到,闪光灯抓拍感,比例 4:5。
2. 拿破仑在儿童乐园滑梯口指挥作战
拿破仑穿全套军装站在商场儿童乐园滑梯入口,表情严肃地指挥一群小孩冲锋,彩色海洋球、塑料城堡、商场灯光,像家长偷拍的疯传照片,比例 4:5。
3. 孔子在自助点餐机前研究半小时
孔子穿古装站在商场餐饮区自助点餐机前,认真盯着菜单界面,后面排队人群一脸无奈,旁边还有“请扫码点单”提示牌,真实生活抓拍风,比例 4:5。
4. 李白醉酒后骑共享单车冲进花坛
夜晚街头,李白穿古装摇摇晃晃骑共享单车,下一秒直接冲进路边花坛,旁边烧烤摊客人集体回头,路人手机抓拍感极强,比例 4:5。
5. 苏格拉底在短视频直播间和弹幕辩论
苏格拉底坐在夸张补光灯和麦克风中间,一边直播一边对弹幕疯狂反问,屏幕上全是“老师你先别绕”“你到底想说啥”,网络直播截图感,比例 9:16。
6. 梵高在夜市摆摊卖自己画的手机壳
夜晚创意市集,梵高穿旧外套站在摊位后面,桌上摆满星空、向日葵图案手机壳和帆布袋,顾客正拿手机付款,街边灯串和随手拍感,比例 4:5。
7. 牛顿在超市里被一整排苹果砸中
超市水果区,一整层苹果突然从货架滚下来砸向牛顿,他一脸震惊地抬头看监控,周围顾客停下脚步,像商场事故现场照片,比例 16:9。
8. 特斯拉在出租屋里修坏掉的电热毯
狭小出租屋里,特斯拉蹲在床边拆开一条老旧电热毯,桌上全是工具、泡面桶和延长线,房间灯光昏黄,像深夜生活纪录片截图,比例 4:5。
9. 玛丽莲梦露在菜鸟驿站找快递找崩溃
玛丽莲梦露穿优雅白裙站在快递架前疯狂翻找包裹,周围堆满快递盒和扫码枪,背景是“请先取件再离开”标语,极具反差的生活化照片,比例 4:5。
10. 秦始皇在银行柜台被要求补材料
银行大厅里,秦始皇穿帝王服饰站在窗口前,表情压抑,柜员礼貌地递给他一张“材料不全请重新排队”,像社会新闻偷拍照,比例 4:5。
---
B. 世界名画人物发疯类(11-20)
11. 蒙娜丽莎在地铁里突然开始大笑
现代地铁车厢里,蒙娜丽莎穿经典服饰坐在座位上,原本神秘微笑突然变成夸张大笑,周围乘客一脸惊恐,抓拍感极强,比例 4:5。
12. 戴珍珠耳环的少女在夜店门口查身份证
夜晚酒吧门口,戴珍珠耳环的少女穿古典服装但神情专业地拿着手电查别人的身份证,霓虹灯、保安绳、强闪光街拍,比例 4:5。
13. 《最后的晚餐》变成部门团建吃自助烤肉
一大桌文艺复兴人物围坐在自助烤肉店里,桌上全是夹子、生菜和五花肉,像公司团建合影,拍法一本正经,比例 16:9。
14. 《呐喊》人物站在游乐园鬼屋出口
著名《呐喊》人物从鬼屋冲出来,背景是一群游客和售票牌,他仍然保持那个经典惊恐姿态,像被记者拍到的真实乐园现场照,比例 4:5。
15. 《宫娥》人物在办公室偷吃下午茶
几位古典宫廷人物坐在现代办公区工位旁边偷偷吃蛋糕和奶茶,老板从远处走来,整张图像某种离谱职场连环画,比例 16:9。
16. 维纳斯站在商场中庭做新品路演
商场中庭,维纳斯雕像被包装成品牌嘉宾,周围是主持人、气球拱门、围观群众,画面像县城商演现场,比例 4:5。
17. 《戴帽子的自画像》人物在驾校科目二现场崩溃
穿艺术家服装的人物坐在教练车里,旁边教练神情麻木,车身压线,整个场面像驾校偷拍视频,比例 16:9。
18. 世界名画人物集体拍毕业旅行合照
多个经典名画人物站在海边景点石碑旁,摆出标准游客姿势,背后还有旅行团小旗,荒诞又真实,比例 4:5。
19. 《创世纪》名场面变成两个人抢手机充电线
两只经典伸出的手,在出租屋床头为了同一根充电线极限对接,构图致敬名画但内容极其现代,比例 4:5。
20. 名画人物在 KTV 包厢唱《突然好想你》
几个不同名画人物挤在 KTV 包厢里,灯光五颜六色,有人举麦克风嘶吼,有人鼓掌,有人喝多了发呆,比例 16:9。
---
C. 监控截图感类(21-30)
21. 凌晨三点便利店里一匹马在买关东煮
监控视角,空荡便利店内,一匹马站在关东煮机器前认真挑选食物,店员低头玩手机毫不惊讶,监控时间戳清晰,比例 16:9。
22. 电梯监控拍到两个宇航员搬西瓜
住宅楼电梯监控画面里,两个穿宇航服的人一人抱一个大西瓜,空间狭窄、灯光惨白,异常真实的离谱画面,比例 9:16。
23. 小区门口监控拍到兵马俑骑电动车
夜晚小区门禁监控,一尊兵马俑骑着电动车缓缓刷卡进门,门禁杆抬起,画面像未解之谜,比例 9:16。
24. 办公室监控拍到老板半夜给绿植开会
空无一人的会议室里,老板站在投影前,对着一排绿植认真讲季度复盘,监控角度、时间戳、冷白灯,全都像真的,比例 16:9。
25. 商场监控拍到一群企鹅逛家居店
大卖场家居区里,几只企鹅排成队看沙发和床垫,其他顾客非常自然,俯拍监控感,比例 16:9。
26. 地库监控拍到一位骑士给共享单车上锁
地下车库里,一名全副铠甲骑士弯腰把共享单车认真停进停车区,监控画质粗糙,比例 16:9。
27. 宿舍走廊监控拍到学生夜里牵着恐龙回寝室
大学宿舍走廊,一名学生牵着一只小型恐龙轻手轻脚走过,宿管室灯亮着,像校园疯传视频截图,比例 9:16。
28. 图书馆监控拍到白天那只鹅晚上回来续借
图书馆服务台监控视角,一只穿围巾的鹅把一本厚书推到柜台前,管理员很平静地办理续借,比例 16:9。
29. 小卖部监控拍到孙悟空偷买辣条
狭小小卖部内,孙悟空穿现代连帽衫,鬼鬼祟祟拿起一包辣条看配料表,监控颗粒感强,比例 4:5。
30. 婚宴酒店监控拍到财神爷独自吃席
大型婚宴厅一角,财神爷打扮的人独自坐在一桌吃菜,周围空座位很多,监控截图带时间码,荒诞又孤独,比例 16:9。
---
D. 假突发社会新闻类(31-40)
31. 城市主干道惊现巨型移动奶茶杯
一只巨大的奶茶杯长出双腿在街上快速移动,交警在旁边指挥交通,电视台记者举着话筒现场直播,比例 16:9。
32. 多地市民反映楼下出现会唱歌的石狮子
夜晚小区门口,两只石狮子对着麦克风深情合唱,围观群众举手机拍摄,像地方民生新闻爆图,比例 4:5。
33. 某高校食堂首次引入龙服务员
高校食堂里,一只温顺的小龙脖子上挂着“新员工培训中”的牌子,正帮学生递餐盘,校园新闻摄影风,比例 16:9。
34. 菜市场出现会算账的章鱼摊主
热闹菜市场里,一只章鱼坐在海鲜摊后面飞快打算盘,顾客们排队买菜,地方都市报视觉,比例 4:5。
35. 市民举报写字楼内出现神秘云层
一整层办公室里漂浮着低矮云雾,员工照常开会敲电脑,消防和记者都到了现场,像离谱突发新闻,比例 16:9。
36. 县城婚礼现场发现不明飞行热气球接亲队伍
婚车车队上空漂浮着奇怪热气球和彩带,乡镇婚礼场景极其热闹,像本地公众号头条图,比例 16:9。
37. 地铁列车内惊现古代赶考团
一整节地铁车厢里全是古代书生打扮的人,背着书箱看电子站牌,新闻摄影抓拍感,比例 16:9。
38. 小区地下车库出现临时瀑布景观
地下车库某个角落突然变成壮观瀑布,业主撑伞经过,物业人员一脸疲惫,像市政投诉新闻照片,比例 16:9。
39. 市中心广场大屏突然播放猫咪会议纪要
城市广场 LED 大屏上滚动播放猫咪高层会议纪要,群众驻足围观拍照,媒体记者现场连线,比例 16:9。
40. 多名市民目击外卖员骑龙准时送餐
晚高峰路口,一名外卖骑手骑着一只飞龙在红绿灯前停下等灯,路人纷纷举手机,新闻图片感很强,比例 16:9。
---
E. 社畜疯感 / 打工人抽象类(41-50)
41. 上班族背着整张床去公司开晨会
写字楼大堂里,一个疲惫上班族背着整张单人床刷闸机,保安已经见怪不怪,比例 4:5。
42. KPI 被印成横幅挂满办公室
开放办公区里四面八方都挂着巨大 KPI 横幅和红色喜报,员工坐在中间表情麻木,极具黑色幽默,比例 16:9。
43. 一位程序员在工位上搭起佛堂求编译成功
办公桌前堆满键盘、显示器和供果,一位程序员双手合十,对着屏幕上报错信息虔诚祈祷,比例 4:5。
44. 茶水间里一群人围着咖啡机做战术分析
普通公司茶水间,一群员工拿着马克杯围着咖啡机像在开军事会议,神情极度严肃,比例 16:9。
45. 老板在周会里放出一只写着“降本增效”的老虎
会议室大屏前,一只老虎背上挂着“降本增效”横幅,员工神情镇定地继续汇报,超现实职场图,比例 16:9。
46. 实习生抱着 18 杯奶茶穿越暴风雨走进办公楼
暴雨街头,一位狼狈实习生双手抱满外卖和奶茶冲向写字楼旋转门,像电影级悲壮场面,比例 4:5。
47. 打工人把工位改造成露营地
办公室工位上搭着小帐篷、挂着营地灯、铺着折叠椅和睡袋,员工仍在认真开会,比例 16:9。
48. 复盘会议上每个人都带着算盘和香炉
商务会议室里,投影上是漂亮图表,但所有参会者桌上都放着算盘、香炉和护身符,比例 16:9。
49. 人事面试现场坐着一个穿龙袍的求职者
现代面试室里,一位穿着帝王龙袍的求职者一本正经递简历,HR 微笑聆听,离谱中带着真实感,比例 4:5。
50. 领导说简单汇报一下,结果背后升起巨型 PPT 城墙
一位年轻员工站在会议室里,背后是夸张到像城墙一样高的 PPT 页面,视觉隐喻非常直白,比例 4:5。
---
F. 宠物和动物更抽象类(51-60)
51. 猫咪在派出所按爪印备案
派出所办事窗口,一只橘猫被民警温柔抱着,在文件上按下爪印,场景庄严又可爱,比例 4:5。
52. 柯基在机场安检口被要求单独开包检查
一只柯基背着小书包站在安检传送带旁,工作人员认真检查它的行李,机场真实纪实感,比例 4:5。
53. 长颈鹿在奶茶店弯腰点单
高高的长颈鹿把头低到收银台上方认真看电子菜单,店员神情平静,社交媒体疯传照感觉,比例 4:5。
54. 黑猫深夜在楼道开会,桌上摊着小区地图
昏暗楼道里,几只黑猫围坐在小桌旁认真研究地图和鱼干,像阴谋组织现场,比例 4:5。
55. 企鹅穿着雨衣在暴雨中送快递
一只企鹅套着黄色雨衣拖着快递车在暴雨城市街头赶路,职业精神拉满,比例 4:5。
56. 鸵鸟在银行自助区缩着头办理转账
银行大厅自动设备旁,一只鸵鸟努力操作转账页面,排队人群很淡定,比例 4:5。
57. 金毛在婚礼现场叼着戒指跑错方向
婚礼草坪上,一只金毛叼着戒指突然朝远处狂奔,新郎新娘和宾客全体惊呼,抓拍感十足,比例 16:9。
58. 海鸥在码头抢走游客薯条后召开庆功会
码头栏杆上,一群海鸥围着抢来的薯条像在分赃庆功,游客在远处无语,比例 16:9。
59. 熊猫在办公室里因为打印机卡纸而沉思
安静办公室里,一只熊猫站在打印机前,怀里抱着一沓文件,表情极其无奈,比例 4:5。
60. 兔子在深夜直播带货卖胡萝卜榨汁机
直播间里,一只白兔戴耳麦坐在桌前,背景是夸张销量大屏和补光灯,9:16 直播截图感。
---
G. 县城国际大片反差类(61-70)
61. 007 电影级镜头拍县城修鞋摊
极其高级的电影布光和构图,一个修鞋摊老板在路边低头补鞋,旁边停着电动车和塑料板凳,比例 4:5。
62. 巴黎时装周气质拍乡镇赶集日
高级时尚大片质感,模特穿着夸张高定在乡镇赶集路上走秀,背景是卖菜、卖鸡和大喇叭宣传车,比例 4:5。
63. 超现实大片拍理发店门口等位人群
小区理发店门口一群人披着毛巾等位,被拍成杂志封面般高级,闪光灯、强对比、比例 4:5。
64. 奢侈品广告拍五金市场砍价现场
一个气质极强的模特站在五金市场里,认真和老板砍电钻价格,镜头语言非常高级,比例 4:5。
65. 电影海报级别拍夜市套圈摊
夜晚霓虹闪烁的夜市里,一个人神情严肃地瞄准套圈,被拍得像命运之战,比例 4:5。
66. 国际体育大片拍大爷公园甩鞭子
清晨公园,一位大爷甩鞭子的瞬间被拍得像奥运冠军海报,光影夸张,比例 16:9。
67. 好莱坞动作片质感拍电动车充电棚
雨夜中,一排电动车在临时充电棚下安静充电,被拍得像末日生存基地,比例 16:9。
68. 时尚杂志封面拍学校门口炸串摊
校门口炸串摊和滚烫油锅,被极度时尚化拍摄,主角手里拿着一串炸淀粉肠,比例 4:5。
69. 黑帮电影气质拍麻将馆门口抽烟阿姨
夜色中,一群阿姨站在麻将馆门口抽烟聊天,灯光、构图、人物关系像黑帮片海报,比例 16:9。
70. 世界末日史诗感拍宿舍停电一小时
学生宿舍停电后,所有人举着手机闪光灯走廊聚集,被拍得像灾难片求生现场,比例 16:9。
---
H. 互联网抽象具象化类(71-80)
71. “已读不回”变成客厅里的幽灵
一个年轻人坐在沙发上,身后漂浮着几个半透明聊天气泡幽灵,上面写着“已读”,冷幽默现实感,比例 4:5。
72. “社恐”变成一件巨大的隐身斗篷
人群密集的聚会里,一个人披着巨大透明斗篷缩在角落,别人都看不清他,视觉隐喻强烈,比例 4:5。
73. “班味”变成一种灰色雾气从写字楼飘出
下班高峰的办公楼门口,成群上班族走出大楼,每个人身后都拖着灰色雾气尾巴,比例 16:9。
74. “内耗”变成一个人和自己拔河
办公室中央,一个人和自己的多个分身进行拔河,周围同事继续办公,超现实又很真实,比例 16:9。
75. “拖延症”变成房间里不断长大的沙发
一个年轻人坐在家中,沙发像生物一样越长越大把他吞进去,桌上待办事项堆积,比例 4:5。
76. “情绪稳定”变成头顶悬浮着一块平静蓝天
办公室里其他人都炸裂崩溃,只有一个人头顶悬浮着一小片平静蓝天,视觉反差明显,比例 4:5。
77. “松弛感”变成地铁里唯一一张沙滩椅
拥挤地铁车厢里,只有一个人坐在沙滩椅上戴墨镜喝椰子水,松弛到离谱,比例 4:5。
78. “破防了”变成玻璃心脏掉在地上
一个人在房间里低头看手机,胸口透明玻璃心脏碎了一地,超现实情绪图,比例 4:5。
79. “热搜体质”变成头顶持续旋转的 LED 跑马灯
街上一个普通人头顶不停转动热搜词条 LED 灯牌,路人围观拍照,比例 4:5。
80. “低气压”变成一个人头上压着乌云上班
地铁站台上,一个年轻人头顶紧贴着一朵小乌云,别人都神色正常,现实魔幻主义,比例 4:5。
---
I. 婚礼、聚会、节日翻车类(81-90)
81. 婚礼接亲队伍里混进了一匹斑马
热闹接亲现场,一匹斑马站在人群中央格外镇定,大家都很忙根本没时间管它,比例 16:9。
82. 生日派对上蛋糕是一个真的小火山
家庭生日现场,中间摆着一座会冒烟的小火山蛋糕,宾客表情复杂,抓拍感强,比例 4:5。
83. 年会抽奖抽到一条龙的使用权
公司年会舞台上,大屏显示一等奖是一条龙,舞台边停着一只温顺但巨大的龙,员工们震惊鼓掌,比例 16:9。
84. 中秋家庭聚餐时月亮降落在阳台上
普通家庭阳台上,一轮真实巨大的圆月停靠在那里,全家人端着菜出来围观,生活化超现实,比例 4:5。
85. 婚礼司仪突然被鹅抢走话筒
婚宴舞台上,一只鹅冲上去叼住麦克风,司仪和新人都愣住,强烈抓拍感,比例 16:9。
86. 春节拜年现场财神和外卖员撞衫
楼道里,穿红金配色制服的外卖员和装扮成财神的人站在门口尴尬对视,比例 4:5。
87. 毕业典礼上学位帽全部飞成鸽子
操场毕业典礼瞬间,抛向空中的学位帽在空中变成一群鸽子飞走,梦幻但真实,比例 16:9。
88. 团建露营现场天幕下坐着一位皇帝
现代露营地里,一群年轻人穿冲锋衣围炉煮茶,中间坐着穿龙袍的皇帝,毫无违和,比例 16:9。
89. 朋友聚餐时有人带来一扇门当礼物
餐厅包厢里,一位朋友认真抬着一整扇门走进来,其他人举杯欢迎,网络迷惑图风格,比例 4:5。
90. 圣诞老人深夜在小区门口被保安登记
圣诞老人背着大袋子站在门岗填写访客登记表,保安一脸认真,冬夜真实抓拍感,比例 4:5。
---
J. 可继续扩写成系列的离谱主题类(91-100)
91. “如果古代名人开始过现代穷日子”系列封面
历史名人出现在合租房、便利店、图文店、地铁、外卖站、菜市场等现实场景里,强调反差、抓拍感、真实纪实风,比例 4:5。
92. “如果名画人物被迫融入现代社交生活”系列
名画人物出现在婚礼、同学聚会、地铁、夜店、KTV、咖啡馆、直播间等现代社交场景,像真实照片而不是绘画,比例 4:5。
93. “如果监控拍到的离谱东西都是真的”系列
用电梯监控、小区监控、商场监控、图书馆监控的视角,拍各种离谱但真实的场面,比例多用 9:16 或 16:9。
94. “如果动物都在认真上班”系列
动物穿着制服进入完全真实的人类职业场景,表情专业,周围人类见怪不怪,适合做一整套传播图,比例 4:5。
95. “如果互联网热词全部具象化”系列
把社恐、内耗、班味、已读不回、热搜体质、低气压、松弛感、破防等做成现实中的可拍照片,比例 4:5。
96. “如果县城气质被拍成国际大片”系列
夜市、理发店、婚礼、菜市场、修鞋摊、炸串摊、麻将馆、公园晨练,全都用电影级布光和时尚摄影手法拍,比例 4:5。
97. “如果职场焦虑变成怪物”系列
把 KPI、周报、绩效、汇报、开会、加班、deadline 变成办公室里的实体怪物或超现实现象,比例 16:9。
98. “一本正经地报道荒唐事件”系列新闻头图
用社会新闻、都市报、地方电视台、突发现场摄影语言,去拍最离谱的东西,强调真实性和滑稽感共存,比例 16:9。
99. “如果所有重要场合都突然跑偏”系列
婚礼、年会、毕业典礼、生日会、接亲、节日家庭聚餐,正常流程中突然出现巨大反差元素,抓拍感强,比例 4:5。
100. “越认真拍,越好笑”系列总纲
用真实抓拍、纪实摄影、监控截图、强闪光街拍、地方新闻、社交媒体疯传图的方式,去拍那些本来极其荒谬的事情,比如皇帝面试、企鹅开会、名画人物唱 K、老板给绿植开会,重点不是梦幻,而是像真的发生过,比例 4:5。
机灵古怪向 100 个 GPT Image 2 提示词
这份清单不是走“正规设计稿”路线,而是偏反差、离谱、魔性、抓拍感、假纪实、互联网传播感。
适合的方向:
- 名人反差图
- 荒诞日常
- 假新闻现场
- 复古抓拍
- 社交媒体 meme 感
- 一本正经但内容离谱
建议用法:
- 可以直接复制一条使用
- 也可以把人物、地点、职业、时代、道具替换掉
- 如果想更像“真实照片”,可以追加:street photography, candid photo, flash photography, imperfect framing, motion blur, paparazzi shot, archival photograph, tabloid style
- 如果想更像“网络疯传图”,可以追加:chaotic composition, accidental masterpiece, meme-worthy, bizarre realism, awkward timing
---
A. 名人离谱日常类(1-10)
1. 爱因斯坦深夜便利店值班
爱因斯坦穿着便利店店员制服,在凌晨两点的 24 小时便利店里给顾客热关东煮,头发依旧炸开,收银台上贴着“今日咖啡第二杯半价”,真实手机抓拍感,冷白荧光灯,轻微噪点,新闻随手拍风格,比例 4:5。
2. 拿破仑在小区门口收外卖
拿破仑穿着经典军装,站在中国普通居民小区门口,双手接过一大袋炸鸡和奶茶,神情严肃但场面很生活化,门口保安亭、共享单车、电动车背景,仿佛被路人偷拍到的照片,真实纪实感,比例 4:5。
3. 莎士比亚在奶茶店改文案
莎士比亚坐在现代奶茶店角落,用羽毛笔认真修改新品海报文案,桌上放着芋泥波波奶茶和一台粉色笔记本电脑,古典服装和现代商业环境形成强烈反差,杂志抓拍风,比例 3:4。
4. 牛顿在水果摊研究苹果打折
牛顿站在菜市场水果摊前,认真盯着一堆打折苹果,左手拿笔记本,右手挑苹果,摊主一脸不耐烦,画面像菜市场偷拍视频截图,生活气特别强,比例 4:5。
5. 达芬奇在五金店挑电钻
达芬奇穿着文艺复兴服装,在现代五金建材店里认真比较三款电钻参数,旁边货架写着“今日特价”,灯光普通,真实商场监控截图质感,荒诞但一本正经,比例 16:9。
6. 贝多芬在广场舞现场当 DJ
贝多芬站在社区广场舞音响设备旁边,神情投入地调音,阿姨们穿统一舞蹈服,背景是夜晚广场和彩灯,整个画面像地方新闻图,带一点闪光灯直打脸的粗粝质感,比例 4:5。
7. 梵高在理发店染蓝发
梵高坐在小区理发店里,头上裹着染发膜,旁边镜子前贴着“洗剪吹 29 元”,理发师神情平静,画面色彩夸张但很真实,像短视频截图,比例 3:4。
8. 居里夫人在厨房研究空气炸锅
居里夫人穿着旧时代长裙站在现代厨房里,严肃地盯着空气炸锅的加热面板,旁边摆着失败的薯条和写满公式的小本子,家庭灯光,生活流摄影,比例 4:5。
9. 特斯拉在网吧修插线板
尼古拉·特斯拉蹲在网吧桌子底下修一个乱七八糟的插线板,周围年轻人正在打游戏,蓝紫霓虹、键盘灯、泡面桶,像一张奇怪又极其真实的网吧偷拍视频,比例 16:9。
10. 玛丽莲梦露在早餐摊等煎饼果子
玛丽莲梦露穿经典白裙站在路边早餐摊前,手里拿着取餐小票,风把裙摆吹起但她一脸平静,背景是豆浆机、塑料凳、路边晨雾,带一点胶片偷拍感,比例 4:5。
---
B. 名人参加中国式生活场景类(11-20)
11. 苏格拉底在小区业主群里吵架
苏格拉底穿着长袍坐在客厅沙发上,一边拿手机发语音一边神情激动,茶几上有半杯凉茶和一堆物业通知,画面像家庭纪实照片,比例 4:5。
12. 凯撒在火锅店排号崩溃
凯撒大帝穿着盔甲坐在火锅店门口塑料椅上,手里捏着“前方还需等待 46 桌”的号码单,表情复杂,旁边一群现代年轻人低头刷手机,强烈反差,比例 3:4。
13. 莫扎特在婚礼现场弹电子琴
莫扎特穿宫廷服装,在县城婚礼舞台旁边弹电子琴,背景是大红色 LED 屏和“新婚快乐”,整个场面土味又华丽,像婚庆公司宣传照,比例 16:9。
14. 维纳斯在商场试穿羽绒服
断臂维纳斯出现在商场羽绒服专柜,导购热情介绍最新款长款羽绒服,周围灯光明亮,带一种高级雕塑突然掉进日常消费世界的荒诞感,比例 4:5。
15. 孔子在家长会上拿着成绩单沉思
孔子穿古装坐在小学教室最后一排,桌上摆着成绩单和保温杯,前方老师正在投影成绩分析 PPT,场景真实到像朋友圈热图,比例 4:5。
16. 李白在深夜烧烤摊举杯自拍
李白坐在露天烧烤摊,桌上全是烤串和啤酒,左手举杯右手拿手机自拍,眼神迷离,旁边朋友都笑疯了,夜市灯光、油烟、抓拍感极强,比例 4:5。
17. 曹雪芹在图文店打印论文
曹雪芹站在学校附近图文打印店里,看着打印机吐出一大叠论文,神情憔悴,柜台上写着“装订 5 元”,学生气氛浓厚,真实生活流,比例 3:4。
18. 秦始皇在高铁站找身份证
秦始皇穿着帝王服饰,站在高铁安检口前翻包找身份证,工作人员一脸职业微笑,旁边旅客匆匆经过,画面像被偶然拍到的爆款照片,比例 16:9。
19. 清明上河图人物在地铁里刷短视频
一整群宋代市井人物穿着古装挤在现代地铁车厢里,所有人都低头刷手机,车厢广告和古人服装形成强烈幽默感,像一张大型行为艺术抓拍,比例 16:9。
20. 武则天在直播间卖护肤品
武则天坐在现代直播间里,背景是夸张打光和“今日宠粉价”灯牌,桌上摆满高端护肤品,她神情威严但正在讲解精华液,直播电商视觉,比例 9:16。
---
C. 假新闻头条类(21-30)
21. 月球发现第一家火锅店开业现场
假新闻摄影,宇航员和外星顾客一起在月球表面围着火锅桌吃毛肚,背景有红色横幅“隆重开业”,像极不靠谱但拍得很真的突发新闻照片,比例 16:9。
22. 南极企鹅集体参加马拉松
一群企鹅胸前别着号码牌在冰面上奔跑,旁边有志愿者递水,现场摄影记者抓拍,像体育新闻封面,荒诞但真实,比例 16:9。
23. 城市高楼之间发现巨型晾衣绳
一条横跨两栋摩天楼之间的超长晾衣绳,上面挂满被子和花衬衫,楼下人群围观拍照,媒体直播车停在路边,像社会新闻,比例 16:9。
24. 大爷发明会自己排队的折叠凳
街头采访新闻图,一位大爷站在小发明旁边,几张折叠凳自动排成一队,路人表情震惊,地方电视台采访风格,比例 4:5。
25. 世界首场猫咪股东大会召开
一群猫戴着小领带坐在会议桌边,桌上有麦克风和矿泉水瓶,记者闪光灯乱飞,像财经新闻现场照片,比例 16:9。
26. 城市突现会遛人的狗
一只巨兴奋的狗拽着主人狂奔穿过街区,主人双脚离地,周围路人手机抓拍,标题党新闻感极强,比例 4:5。
27. 博物馆雕像半夜集体下班买夜宵
石雕人物排队站在夜市小摊前买烤冷面,背后就是博物馆侧门,灯光昏黄,像监控截图流出,比例 16:9。
28. 小区电梯里发现一只穿西装的鹅
监控视角,一只白鹅穿合身西装站在电梯中央,表情冷静,旁边住户刻意假装不看它,极度荒诞,比例 9:16。
29. 图书馆自习区惊现骑马赶论文的人
夜晚图书馆里,一名学生骑着一匹马冲向打印区,其他同学继续低头写作业,像校园疯传图,比例 16:9。
30. 古堡里首次举办楼下广场舞联赛
中世纪古堡庭院里,一群穿宫廷服饰的人排成整齐方阵跳广场舞,裁判席和观众席非常正式,像离谱国际赛事新闻,比例 16:9。
---
D. 复古抓拍 / 伪历史照片类(31-40)
31. 1920 年代的上班族地铁通勤照
黑白胶片风,一群穿 1920 年代西装礼帽的人在拥挤地铁里低头看现代智能手机,历史与现代强烈错位,仿佛一张被篡改的档案照片,比例 4:5。
32. 古罗马人第一次见到自动贩卖机
古罗马市民围在一台现代自动贩卖机前研究按钮,托加长袍、石柱街道、纪实老照片质感,轻微划痕和颗粒,比例 4:5。
33. 70 年代 disco 现场混入一位修仙者
复古舞厅里,全员爆炸头和亮片衣服,只有中间一位仙风道骨的人盘腿打坐,闪光灯抓拍,诡异又合理,比例 3:4。
34. 清朝宫廷成员第一次拍大头贴
狭窄的大头贴机器里,一群清朝服饰人物挤在镜头前做古怪表情,四格照片版式,复古又搞笑,比例 3:4。
35. 文艺复兴画家集体团建打卡照
一群文艺复兴大师站在景区网红打卡点前统一比耶,后面是巨大的彩色塑料装置,旅游照构图,比例 4:5。
36. 维多利亚时代的侦探在网吧查资料
昏黄灯光下,一位穿长风衣的维多利亚侦探坐在烟味很重的老网吧里查案,旁边全是游戏界面,胶片纪实风,比例 16:9。
37. 老上海名媛排队买盲盒
民国风服装的名媛们站在商场快闪店门口排队买限量盲盒,纸袋、霓虹招牌、胶片颗粒,仿佛时代穿越街拍,比例 4:5。
38. 古埃及祭司主持现代剪彩仪式
古埃及祭司服装、金色饰品、红色剪彩带、现代商场门口,媒体相机闪烁,画面像一张旧报纸上的离谱新闻配图,比例 16:9。
39. 中世纪骑士在便利店门口抽电子烟
夜晚便利店门外,一名全副盔甲骑士靠着冰柜抽电子烟,停车位上停着共享单车,城市街头抓拍风,比例 4:5。
40. 古希腊哲学家们集体拍毕业照
一群古希腊哲学家穿长袍站成毕业照队形,前排蹲姿、后排站姿,手里拿着现代毕业证书和鲜花,严肃脸却很荒诞,比例 4:5。
---
E. 离谱合影 / 群像闹剧类(41-50)
41. 世界伟人一起参加社区运动会
爱因斯坦、拿破仑、莎士比亚、达芬奇等历史名人穿着统一运动服站在塑胶跑道上拍开幕式合影,背后横幅写“第二届友谊第一比赛第二”,比例 16:9。
42. 所有童话人物在火车站集体误车
白雪公主、匹诺曹、小红帽、美人鱼等童话角色拖着大箱子在火车站狂奔,广播屏幕显示“列车已停止检票”,现场混乱抓拍,比例 16:9。
43. 不同时代的皇帝一起参加公司年会
古代帝王们坐在圆桌宴席前抽奖,舞台大屏写着“年度冲刺 再创辉煌”,每个人表情都很严肃,像企业活动摄影,比例 16:9。
44. 各国神话人物挤在出租车后排
后排座位上挤着宙斯、哪吒、美杜莎、阿努比斯等神话角色,司机一脸麻木,夜晚城市霓虹倒影,像电影幕后偷拍照,比例 16:9。
45. 文豪们一起拍海边游客照
鲁迅、莎士比亚、托尔斯泰、海明威等文学巨匠站在海边景点巨石前拍游客纪念照,背后有人放风筝,反差非常大,比例 4:5。
46. 所有发明家参加电子产品发布会
特斯拉、爱迪生、达芬奇、图灵等坐在前排看一场现代手机发布会,灯光酷炫,所有人都很认真,仿佛他们真是业内嘉宾,比例 16:9。
47. 古代将军们一起做核酸排队照
不同朝代的武将全副盔甲但非常配合地站在社区检测点排队,手里拿着手机二维码,像社会纪实摄影,比例 4:5。
48. 神仙妖怪团建吃自助餐
一群中国神话角色在大型自助餐厅端盘子夹菜,玉皇大帝在夹寿司,孙悟空拿着冰淇淋,场面热闹又滑稽,比例 16:9。
49. 全世界侦探集体参加密室逃脱
福尔摩斯、波洛、柯南式侦探风人物们站在密室逃脱门口合影,人人都露出“这题太简单”的表情,霓虹店招很现代,比例 4:5。
50. 历史名人拍公司工牌证件照
一整排历史名人的工牌证件照墙,所有人穿统一商务装,姓名牌一本正经,像互联网公司入职系统截图,比例 4:5。
---
F. 动物成精 / 假纪实奇观类(51-60)
51. 柴犬主持晚间新闻
一只柴犬穿正装坐在新闻主播台后面,严肃播报国际新闻,提词器反光、演播室灯光、专业镜头感,荒诞但毫无违和,比例 16:9。
52. 猫咪在婚礼上担任摄影师
一只胖橘猫挂着专业相机在婚礼现场跑来跑去抓拍,宾客们都默认它是工作人员,画面像真实婚礼摄影侧拍,比例 4:5。
53. 哈士奇在警局做笔录
哈士奇坐在桌前,前爪搭在警局桌面上,一脸无辜,警察正在记录,旁边墙上有“坦白从宽”字样,像本地新闻截图,比例 4:5。
54. 企鹅在机场商务舱休息室办公
一只企鹅坐在机场 lounge 的桌边敲笔记本电脑,旁边是咖啡、登机牌、行李箱,商务精英感和企鹅本体形成强烈反差,比例 4:5。
55. 鸭子成为小区保安队长
一只体型很大的鸭子戴着保安帽站在门禁杆旁巡逻,居民们非常自然地从它身边刷卡进门,监控截图感,比例 9:16。
56. 羊驼参加脱口秀开放麦
一家小酒馆舞台上,一只羊驼站在立麦前讲段子,台下观众笑疯了,暖黄色酒吧灯光,像真实演出照片,比例 4:5。
57. 金鱼坐在鱼缸里远程开会
透明鱼缸里有一条一本正经的金鱼,鱼缸前摆着电脑和降噪耳机,视频会议界面开着,办公室工位真实感,比例 4:5。
58. 奶牛在美术馆认真看抽象画
美术馆白盒子空间里,一头奶牛很认真地站在一幅极简抽象画前,旁边还有解说牌和安静的观众,荒诞高级感,比例 3:4。
59. 兔子在菜市场砍价买胡萝卜
一只白兔站在菜市场摊位前跟摊主认真砍价,旁边塑料袋和零钱散落,手机抓拍感极强,比例 4:5。
60. 乌鸦在法庭上当书记员
庄严法庭里,一只乌鸦站在书记员位置敲打键盘记录庭审,整体非常正式,黑色羽毛和木质法庭氛围很有戏剧感,比例 16:9。
---
G. 网络迷因 / 魔性社交媒体类(61-70)
61. 朋友聚会里只有一个人穿宇航服
普通家庭客厅生日聚会,一群人都穿日常衣服,只有角落里一个人穿完整宇航服端着蛋糕,像朋友圈疯传的照片,比例 4:5。
62. 健身房里最努力的是一只鹅
现代健身房中,一只鹅在跑步机上拼命奔跑,其他会员一脸平静继续锻炼,构图像偷拍视频,比例 4:5。
63. 婚礼上新郎突然换成纸板立牌
婚礼交换戒指现场,新娘旁边站着的是一个等身纸板新郎,司仪和宾客表情都努力保持正常,极其魔性,比例 4:5。
64. 办公室团建合照混入一只羊
互联网公司会议室合影,所有员工都穿工牌微笑,最后一排站着一只白羊,大家像完全没觉得有问题,比例 16:9。
65. 在图书馆自习的人头上长出 Wi-Fi 信号
深夜图书馆里,几位学生正埋头写作业,但每个人头顶都隐约长出发光的 Wi-Fi 信号标志,安静又诡异,比例 16:9。
66. 地铁里一个人举着鱼缸当公文包
早高峰地铁,一位西装上班族手里拎着一个装着金鱼的小鱼缸当公文包,周围人都假装没看见,抓拍感强,比例 4:5。
67. 全班同学都在考试,只有一人带着露营装备
安静考场里,所有学生正常答题,只有一个人搭了小帐篷、带保温壶和头灯,离谱但构图真实,比例 16:9。
68. 超市收银台前排着人和一匹马
普通超市收银区,一匹马安静地站在队伍里等待结账,前后顾客表情毫无波澜,比例 4:5。
69. 会议室里所有人都开电脑,只有老板拿着算盘
现代商务会议室,员工都用笔记本电脑做汇报,老板坐在主位上认真拨算盘,整个画面像某种离谱职场隐喻 meme,比例 16:9。
70. 医院候诊区有人带着一盆仙人掌看病
候诊区座椅上,一个年轻人抱着一大盆仙人掌认真等叫号,场景十分真实,带一点冷幽默,比例 4:5。
---
H. 假时尚大片 / 反差广告类(71-80)
71. 高定时装大片,但拍摄地点是五金建材城
超模级人物穿夸张高定礼服站在五金建材城货架之间,背后全是水管、瓷砖、电线卷,时尚大片布光,荒诞高级感,比例 4:5。
72. 顶级香水广告,但主角是一碗螺蛳粉
极致奢华香水广告摄影风格,主角却是一碗热气腾腾的螺蛳粉,被当作高端时尚产品拍摄,暗色背景、珠宝级布光,比例 3:4。
73. 跑车广告,但车主是卖菜阿姨
一位菜市场阿姨非常自然地靠在顶级跑车旁,手里还拎着青菜和塑料袋,拍成奢侈品牌广告气质,比例 16:9。
74. 高级珠宝大片,但模特在小卖部吃辣条
一个气场很强的模特佩戴顶级珠宝,站在小卖部门口边吃辣条边看镜头,强闪光、时尚杂志封面风,比例 4:5。
75. 奢侈手表广告,但背景是工地食堂
极简奢侈手表特写,佩戴者穿工装坐在工地食堂里吃盒饭,质感非常高级但环境非常接地气,比例 4:5。
76. 护肤品广告,但模特是兵马俑
顶级商业美容广告风格,一尊兵马俑正在被柔和灯光照亮,旁边摆着玻尿酸精华瓶,极致反差,比例 3:4。
77. 高端家居广告,但房间里养着一头牛
极简主义室内设计空间、奶油色家具和大片自然光,中间却有一头奶牛非常优雅地站着,像一本正经的家居品牌广告,比例 16:9。
78. 法式杂志封面,但主角是刚出锅的包子
一本高级时尚杂志封面视觉,主角不是模特而是一笼白胖热腾的包子,标题排版很认真,比例 4:5。
79. 户外运动品牌大片,但人物在办公室爬工位
户外探险广告的视觉语言,主角穿专业登山装备,在公司开放办公区沿着工位和隔断攀爬,史诗感夸张,比例 16:9。
80. 豪华酒店宣传片风,但拍的是学校宿舍
四人间大学宿舍被拍成超五星酒店宣传片风格,镜头语言极度精致,但细节里仍有晾袜子、桶装泡面和风扇,比例 16:9。
---
I. 互联网抽象脑洞类(81-90)
81. 浏览器弹窗实体化追着人跑
街头场景里,一堆巨大的浏览器弹窗和广告横幅变成实体物体追着路人狂跑,像超现实灾难喜剧,比例 16:9。
82. 表情包人物从手机里爬出来
一个年轻人坐在沙发上刷手机,经典夸张表情包角色正从手机屏幕里往外爬,家庭灯光,带点恐怖喜剧味道,比例 4:5。
83. “撤回不了的消息”变成街头追捕现场
城市夜晚,一条巨大的聊天气泡写着“对方已看到”,像怪物一样在街上追人,荒诞电影海报感,比例 4:5。
84. 打工人的待办事项长成一面墙
一个上班族站在办公室里,背后是无数便利贴和待办清单堆成的巨大墙体,几乎像怪物压下来,幽默又共鸣,比例 4:5。
85. 社交媒体点赞变成真实奖牌挂满全身
一位普通博主站在房间中央,身上挂满夸张的金属“点赞”奖牌,沉重到走不动路,视觉隐喻很强,比例 3:4。
86. 外卖骑手在末日废土里准时送达
末日废土世界,一位外卖骑手仍然骑着电动车准时穿过废墟和火光送餐,头盔、保温箱、严肃职业感,比例 16:9。
87. KPI 具象化成办公室里的巨型怪兽
现代办公室里,一只由图表、柱状图、报表、红色箭头组成的怪兽站在老板身后,员工们继续假装在开会,比例 16:9。
88. 热搜榜变成古代皇榜张贴现场
古代城门口,百姓围观一张巨大的“今日热搜榜”皇榜,榜上全是现代网络词汇,古今错位感很强,比例 4:5。
89. 低电量模式变成城市黄昏状态
整个城市像手机进入低电量模式一样变暗、变黄、运转缓慢,路人神情疲惫,像现实世界被系统设定影响,比例 16:9。
90. 云端备份变成天上掉文件夹
城市上空乌云密布,但云里掉下来的不是雨,而是大量系统文件夹、压缩包和上传进度条,荒诞科幻现实主义,比例 16:9。
---
J. 离谱系列海报 / 可批量扩写类(91-100)
91. “如果历史名人都住在同一个合租屋”海报
一张系列海报主视觉,爱因斯坦、拿破仑、莎士比亚、李白等历史名人一起住在现代合租公寓里,公共厨房、混乱客厅、冰箱贴满值日表,像喜剧剧集宣传海报,比例 4:5。
92. “古代人第一次用现代 App”系列封面
一张夸张有趣的系列封面,古代人物认真使用外卖、打车、短视频、招聘、地图导航 App,界面真实,人物反应荒诞,比例 4:5。
93. “神仙下凡做基层工作”系列海报
神仙角色穿着工服做普通基层职业,比如客服、保安、收银员、外卖员、物业维修,视觉既庄重又搞笑,比例 4:5。
94. “如果动物突然拥有正式编制”系列证件照
不同动物穿职业制服拍官方证件照,柴犬主播、企鹅律师、乌鸦书记员、鸭子保安,背景统一,像政府官网人物页,比例 4:5。
95. “互联网热词实体化”系列图
把拖延症、内耗、已读不回、热搜体质、社恐、班味等词变成具象场景,半纪实半超现实,适合做一组病毒感海报,比例 4:5。
96. “如果博物馆文物夜里偷偷打工”系列海报
兵马俑送外卖、维纳斯做导购、青铜器开直播、石狮子当门卫,城市夜晚纪实抓拍感,比例 4:5。
97. “假如世界名画人物进入现代社会”系列图
名画人物出现在地铁、便利店、写字楼、奶茶店、图文店、婚礼现场,画风真实抓拍,不要油画质感,要像真实照片,比例 4:5。
98. “假如所有神话人物都要参加年终考核”系列海报
各路神仙妖怪穿商务正装,在会议室、汇报厅、打印区、茶水间里为 KPI 和汇报材料焦头烂额,像职场黑色幽默大片,比例 16:9。
99. “县城气质 × 国际大片”系列图
把国际大片的布光、构图、人物气场,放进县城婚礼、建材市场、夜宵摊、理发店、五金店、广场舞现场,强反差视觉,比例 4:5。
100. “一本正经地拍很荒唐的事”系列封面
电影级或杂志级摄影手法,去拍那些极其生活化但很荒唐的事情,比如皇帝挤地铁、哲学家抢优惠券、骑士抽电子烟、宇航员吃路边摊,要求真实、严肃、像真的新闻现场,比例 4:5。
GPT-Image-2 Prompt Templates
Use these templates when the user wants reusable prompt structures instead of one-off prompts.
Template 1: Historical world × modern interface
“[era/civilization] [people/use case] [interface type]” / “[English title]”, [historical setting] fused with modern [product/interface] design, the image simulates a [device] showing a [home screen/detail page/list page], top area displays [place/time/status], the main area includes [module 1], [module 2], [module 3], icons and controls are rebuilt with [historical motif/material], portraits use [art style], typography combines [headline style] with [body style], overall palette is [color 1], [color 2], [color 3], the result should feel like a real product design screen with strong historical contrast, aspect ratio [ratio].
Template 2: Emotion × infographic poster
“[city/topic/emotion] index” / “[English title]”, data-visualization poster fused with [urban observation / emotional narrative / lifestyle editorial], center composition uses [map / time wheel / floor plan / axis], layered by [time / intensity / geography], surrounded by [data module 1], [data module 2], [data module 3], [data module 4], headline styled like [magazine / report / exhibition title], subtitle reads “[short line]”, restrained but emotionally precise, aspect ratio [ratio].
Template 3: Subject × editorial still-life spread
“[subject/persona] room” / “[English title]”, editorial still-life narrative layout, top-down composition on a [table/surface material], arranged objects include [object 1], [object 2], [object 3], [object 4], [object 5], [object 6], each object gets a fine annotation line and a small note describing [memory / function / clue], typography is [font style], palette is [palette], should feel like a magazine feature spread or exhibition caption page, aspect ratio [ratio].
Template 4: Space × design proposal board
“[space/project] wayfinding system” / “[English title]”, complete visual identity proposal board for [space type], the main image is a perspective view of the space, showing [signage], [icons], [zoning], [directional flow], one side expands into [tickets / tags / uniforms / floor stickers / maps], material language emphasizes [material 1], [material 2], [material 3], overall look should feel like a professional design pitch deck, aspect ratio [ratio].
Template 5: Worldbuilding × dossier page
“[organization] dossier: [subject name]”, structured archive sheet for a [myth / sci-fi / fantasy / research] subject, left side shows [front / side / back / action breakdown], right side includes [detail module 1], [detail module 2], [behavior/ecology/system data], bottom area has [observation log / warning note / numbered labels], stamped with [clearance level / catalog mark / institution label], should feel like an internal archive page, aspect ratio [ratio].
Template 6: Daily life × cinematic poster
“[ordinary event]” film poster, the main image shows [scene], subject is [pose/action/emotional state], visual tone references [art-house / suspense / quiet apocalypse / black comedy], title is “[film title]”, subtitle/tagline is “[line]”, lower section includes cast billing and release details, strong cinematic composition, aspect ratio [ratio].
Template 7: Future service × user journey board
“[future service] experience design”, service-design presentation board showing [space / interface / device], side column contains a user journey from [step 1] to [step 2] to [step 3], each step annotated with [emotion / waiting time / action / feedback], should look like a professional innovation presentation, aspect ratio [ratio].
Template 8: Annual archive × cover system
“[year] visual almanac cover” / “[English title]”, cover composed of [number] framed scenes, each frame represents one visual memory from the year: [scene 1], [scene 2], [scene 3], [scene 4], [scene 5], central title overlays all frames, subtitle remains restrained, should feel like a premium annual issue cover, aspect ratio [ratio].
Quick replacement vocabulary
Image types
- poster
- infographic
- dashboard
- interface screen
- concept art board
- magazine cover
- archive sheet
- design proposal
- guide map
- packaging visual
Structural modules
- comment section
- parameter bar
- annotation labels
- heat map
- ranking list
- route map
- timeline
- legend
- source note
- footer caption
Tone words to concretize
- premium
- editorial
- cinematic
- scientific
- restrained
- playful but serious
- nostalgic
- futuristic
- documentary-like
Security Policy
Supported versions
The current public skill package is supported from v0.1.0 onward while this repository remains active.
Reporting a vulnerability
For documentation bugs, broken references, or unsafe prompt examples, open a public GitHub issue with the affected file and a short reproduction note.
For sensitive reports, avoid posting secrets, private prompts, API keys, or account data in a public issue. Contact the maintainer through the GitHub profile first, then share only the minimum details needed to reproduce the problem.
Scope
Security review for this repository focuses on:
- Prompt examples that could encourage unsafe usage
- Documentation that may lead users to expose API keys or private content
- Repository files that affect Codex skill discovery or execution
- Dependency or script changes if future versions add executable tooling
Maintainer response
The maintainer will triage valid reports, document the fix in the relevant issue or release notes, and prioritize changes that reduce user risk.