
Nano Banana 2
- 588 installs
- 680 repo stars
- Updated August 3, 2026
- inference-sh/skills
nano-banana-2 is a Claude Code skill that generates and edits images via Google Gemini 3.1 Flash Image Preview using the inference.sh belt CLI for developers automating multimodal agent workflows.
About
nano-banana-2 is a generative media skill from inference-sh/skills that drives Google Gemini 3.1 Flash Image Preview, also called Nano Banana 2, through the inference.sh belt CLI. Capabilities include text-to-image generation, image editing, multi-image input supporting up to 14 images, and Google Search grounding for grounded visuals. The skill requires installing the belt CLI via npx skills add belt-sh/cli and allows Bash belt commands inside agent sessions. Developers reach for nano-banana-2 when pipelines need fast, low-cost image synthesis or edits without standing up a separate image API integration from scratch.
- Runs tiny LLMs with minimal latency and cost
- Designed for on-device or edge inference inside agents
- Compatible with Claude Code, Cursor, and generic agent hosts
- Reduces token spend on repetitive micro-tasks
- 498 developers have installed this skill
Nano Banana 2 by the numbers
- 588 all-time installs (skills.sh)
- +15 installs in the week ending Aug 2, 2026 (Skillselion tracking)
- Ranked #1,605 of 16,546 AI & Agent Building skills by installs in the Skillselion catalog
- Data as of Aug 5, 2026 (Skillselion catalog sync)
npx skills add https://github.com/inference-sh/skills --skill nano-banana-2Add your badge
Show developers this skill is listed on Skillselion. Paste this into your README.
| Installs | 588 |
|---|---|
| repo stars | ★ 680 |
| Last updated | August 3, 2026 |
| Repository | inference-sh/skills ↗ |
How do you generate images with Gemini via inference.sh CLI?
Get fast, low-cost inference from lightweight language models directly inside agent workflows.
Who is it for?
Developers wiring agent workflows who need Gemini 3.1 Flash image generation and editing through the inference.sh belt CLI.
Skip if: Teams avoiding shell-based CLI integrations, video generation workloads, or projects requiring non-Google image model providers.
When should I use this skill?
The user mentions nano banana 2, Gemini 3.1 flash image, text-to-image, or image editing via inference.sh belt.
What you get
Generated or edited image files, belt CLI commands, and grounded multimodal outputs from up to 14 input images.
- Generated image files
- Edited image outputs
- belt CLI command recipes
By the numbers
- Supports multi-image input up to 14 images
- Uses Google Gemini 3.1 Flash Image Preview model
Files
Install the belt CLI skill: npx skills add belt-sh/cliNano Banana 2 - Gemini 3.1 Flash Image Preview
Generate images with Google Gemini 3.1 Flash Image Preview via inference.sh CLI.
Quick Start
Requires inference.sh CLI (belt). Install instructionsbelt login
belt app run google/gemini-3-1-flash-image-preview --input '{"prompt": "a banana in space, photorealistic"}'Examples
Basic Text-to-Image
belt app run google/gemini-3-1-flash-image-preview --input '{
"prompt": "A futuristic cityscape at sunset with flying cars"
}'Multiple Images
belt app run google/gemini-3-1-flash-image-preview --input '{
"prompt": "Minimalist logo design for a coffee shop",
"num_images": 4
}'Custom Aspect Ratio
belt app run google/gemini-3-1-flash-image-preview --input '{
"prompt": "Panoramic mountain landscape with northern lights",
"aspect_ratio": "16:9"
}'Image Editing (with input images)
belt app run google/gemini-3-1-flash-image-preview --input '{
"prompt": "Add a rainbow in the sky",
"images": ["https://example.com/landscape.jpg"]
}'High Resolution (4K)
belt app run google/gemini-3-1-flash-image-preview --input '{
"prompt": "Detailed illustration of a medieval castle",
"resolution": "4K"
}'With Google Search Grounding
belt app run google/gemini-3-1-flash-image-preview --input '{
"prompt": "Current weather in Tokyo visualized as an artistic scene",
"enable_google_search": true
}'Input Options
| Parameter | Type | Description |
|---|---|---|
prompt | string | Required. What to generate or change |
images | array | Input images for editing (up to 14). Supported: JPEG, PNG, WebP |
num_images | integer | Number of images to generate |
aspect_ratio | string | Output ratio: "1:1", "16:9", "9:16", "4:3", "3:4", "auto" |
resolution | string | "1K", "2K", "4K" (default: 1K) |
output_format | string | Output format for images |
enable_google_search | boolean | Enable real-time info grounding (weather, news, etc.) |
Output
| Field | Type | Description |
|---|---|---|
images | array | The generated or edited images |
description | string | Text description or response from the model |
output_meta | object | Metadata about inputs/outputs for pricing |
Prompt Tips
Styles: photorealistic, illustration, watercolor, oil painting, digital art, anime, 3D render
Composition: close-up, wide shot, aerial view, macro, portrait, landscape
Lighting: natural light, studio lighting, golden hour, dramatic shadows, neon
Details: add specific details about textures, colors, mood, atmosphere
Sample Workflow
# 1. Generate sample input to see all options
belt app sample google/gemini-3-1-flash-image-preview --save input.json
# 2. Edit the prompt
# 3. Run
belt app run google/gemini-3-1-flash-image-preview --input input.jsonPython SDK
from inferencesh import inference
client = inference()
# Basic generation
result = client.run({
"app": "google/gemini-3-1-flash-image-preview@0c7ma1ex",
"input": {
"prompt": "A banana in space, photorealistic"
}
})
print(result["output"])
# Stream live updates
for update in client.run({
"app": "google/gemini-3-1-flash-image-preview@0c7ma1ex",
"input": {
"prompt": "A futuristic cityscape at sunset"
}
}, stream=True):
if update.get("progress"):
print(f"progress: {update['progress']}%")
if update.get("output"):
print(f"output: {update['output']}")Related Skills
# Original Nano Banana (Gemini 3 Pro Image, Gemini 2.5 Flash Image)
npx skills add inference-sh/skills@nano-banana
# Full platform skill (all 250+ apps)
npx skills add inference-sh/skills@infsh-cli
# All image generation models
npx skills add inference-sh/skills@ai-image-generationBrowse all image apps: belt app store --category image
Documentation
- Running Apps - How to run apps via CLI
- Streaming Results - Real-time progress updates
- File Handling - Working with images
Related skills
Forks & variants (2)
Nano Banana 2 has 2 known copies in the catalog totaling 74 installs. They canonicalize to this original listing.
- qu-skills - 70 installs
- skills-shell - 4 installs
FAQ
What CLI does nano-banana-2 require?
nano-banana-2 requires the inference.sh belt CLI. Install the belt CLI skill with npx skills add belt-sh/cli, then invoke Gemini 3.1 Flash Image Preview generation and editing commands from agent sessions.
How many input images does nano-banana-2 support?
nano-banana-2 supports multi-image input with up to 14 images for Gemini 3.1 Flash Image Preview. The skill also covers text-to-image, image editing, and Google Search grounding workflows.