Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
lovstudio avatar

Lovstudio:Document Illustrator

  • 2 installs
  • Updated May 30, 2026
  • lovstudio/document-illustrator-skill

Inserts AI-generated illustrations into a document in place, planning insertion points globally and generating all images in parallel.

About

Reads a document, plans illustration insertion points globally, generates all images in parallel, and inserts them back into the source. A user uses it to add cover images and inline illustrations to articles or notes with configurable ratios and styles.

  • Globally plans insertion points then generates images in parallel
  • Supports cover images, custom ratios, and three styles

Lovstudio:Document Illustrator by the numbers

  • 2 all-time installs (skills.sh)
  • Ranked #1,166 of 1,335 Generative Media skills by installs in the Skillselion catalog
  • Data as of Jul 8, 2026 (Skillselion catalog sync)
npx skills add https://github.com/lovstudio/document-illustrator-skill --skill lovstudiodocument-illustrator

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs2
Last updatedMay 30, 2026
Repositorylovstudio/document-illustrator-skill

What it does

Inserts AI-generated illustrations into a document in place, planning insertion points globally and generating all images in parallel.

Files

SKILL.mdMarkdownGitHub ↗

Document Illustrator Skill

基于 AI 智能分析的文档配图生成工具。全局规划、并行生成、异步插入,高效为文档添加配图。

核心流程(5 步)

备份 → 全局规划插入点 → 并行生成图片 → 异步插入原文 → 清理备份

Step 0: 备份原文件

在修改前先创建备份,确保安全回滚:

import shutil
backup_path = f"{doc_path}.illustrator-backup"
shutil.copy2(doc_path, backup_path)

所有后续操作直接在原文件上进行。

Step 1: 全局确定所有插入位置

读取完整文档,一次性规划所有图片的插入位置:

1. 使用 Read 工具读取完整文档 2. AI 分析内容结构,识别核心主题 3. 为每个主题确定精确的插入锚点(行号 + 上下文文本) 4. 输出一份插入计划表:

插入计划:
  [1] 行 15 后 | 锚点: "## Rules 的诞生" | 主题: Rules 演化历程
  [2] 行 42 后 | 锚点: "## Commands 打包" | 主题: 工作流打包
  [3] 行 78 后 | 锚点: "## MCP 动态能力" | 主题: 第三方集成
  ...
  [cover] 行 1 前 | 封面图 | 主题: 全文概要

关键:插入锚点使用上下文文本(而非纯行号),这样即使前面的插入导致行号偏移,后续插入仍可通过锚点定位。

Step 2: 并行生成所有图片

用 Agent 工具并行启动所有图片生成子任务:

对每个插入计划项,同时启动一个 Agent:
  Agent 1: generate_single_image.py --title "..." --content "..." --output images/illustration-01.png
  Agent 2: generate_single_image.py --title "..." --content "..." --output images/illustration-02.png
  Agent 3: generate_single_image.py --title "..." --content "..." --output images/illustration-03.png
  ...
  • 所有 Agent 并发执行,不互相等待
  • 每个 Agent 完成后返回图片路径或错误信息
  • 预期总耗时 = 单张耗时(10-20s),而非 N * 单张耗时

Step 3: 异步插入原文

每个 Agent 完成后立即插入,不等待其他 Agent:

1. Agent 完成 → 获得图片路径 2. 在原文档中通过锚点文本定位插入位置(不依赖行号) 3. 使用 Edit 工具在锚点后插入 Markdown 图片引用:

   ![主题描述](images/illustration-01.png)

4. 插入使用锚点文本匹配,所以前面的插入不影响后面的定位

位置偏移处理

  • 每次插入会增加文档行数
  • 使用锚点文本(如 ## Rules 的诞生)而非行号来定位
  • 从文档末尾向开头方向插入也可避免偏移问题

Step 4: 验证与清理

所有图片插入完成后:

1. 验证:检查原文档中所有计划的 ![...]() 引用都已插入 2. 验证:检查所有图片文件都存在于 images/ 目录 3. 成功 → 删除备份文件 {doc_path}.illustrator-backup 4. 失败 → 保留备份文件,报告哪些图片未能生成/插入,用户可用备份恢复

完成: 6/6 张配图已插入原文档
已清理备份文件

配置选项

执行前 Claude 会询问(或从用户消息中推断):

选项默认
图片比例16:9 / 3:416:9
是否封面图是/否
内容配图数量3-10根据文档长度推荐
风格gradient-glass / ticket / vector-illustrationgradient-glass

如果用户在请求中已指定(如"竖屏、票据风格、8张"),直接使用,不再询问。

风格速查

风格关键词适合
gradient-glass玻璃拟态、极光渐变、科技感技术文档、产品介绍
ticket黑白对比、票券结构、极简数据报告、信息图表
vector-illustration扁平插画、复古配色、几何化教程、故事、品牌

风格文件位于 styles/ 目录。

技术细节

项目
API 模型Gemini 2.0 Flash Image Preview
16:9 分辨率2560x1440 (2K) / 3840x2160 (4K)
3:4 分辨率1920x2560 (2K) / 2880x3840 (4K)
单张耗时~10-20s
并行耗时~10-20s(总,不乘 N)
依赖pip install google-genai pillow python-dotenv
API Key.envGEMINI_API_KEY 或环境变量

脚本

  • scripts/generate_single_image.py — 单张图片生成(供 Agent 并行调用)
  • scripts/generate_illustrations.py — 旧版批量顺序生成(保留兼容)

Related skills

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.