Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
inference-sh avatar

Nano Banana 2

  • 588 installs
  • 680 repo stars
  • Updated August 3, 2026
  • inference-sh/skills

nano-banana-2 is a Claude Code skill that generates and edits images via Google Gemini 3.1 Flash Image Preview using the inference.sh belt CLI for developers automating multimodal agent workflows.

About

nano-banana-2 is a generative media skill from inference-sh/skills that drives Google Gemini 3.1 Flash Image Preview, also called Nano Banana 2, through the inference.sh belt CLI. Capabilities include text-to-image generation, image editing, multi-image input supporting up to 14 images, and Google Search grounding for grounded visuals. The skill requires installing the belt CLI via npx skills add belt-sh/cli and allows Bash belt commands inside agent sessions. Developers reach for nano-banana-2 when pipelines need fast, low-cost image synthesis or edits without standing up a separate image API integration from scratch.

  • Runs tiny LLMs with minimal latency and cost
  • Designed for on-device or edge inference inside agents
  • Compatible with Claude Code, Cursor, and generic agent hosts
  • Reduces token spend on repetitive micro-tasks
  • 498 developers have installed this skill

Nano Banana 2 by the numbers

  • 588 all-time installs (skills.sh)
  • +15 installs in the week ending Aug 2, 2026 (Skillselion tracking)
  • Ranked #1,605 of 16,546 AI & Agent Building skills by installs in the Skillselion catalog
  • Data as of Aug 5, 2026 (Skillselion catalog sync)
npx skills add https://github.com/inference-sh/skills --skill nano-banana-2

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs588
repo stars680
Last updatedAugust 3, 2026
Repositoryinference-sh/skills

How do you generate images with Gemini via inference.sh CLI?

Get fast, low-cost inference from lightweight language models directly inside agent workflows.

Who is it for?

Developers wiring agent workflows who need Gemini 3.1 Flash image generation and editing through the inference.sh belt CLI.

Skip if: Teams avoiding shell-based CLI integrations, video generation workloads, or projects requiring non-Google image model providers.

When should I use this skill?

The user mentions nano banana 2, Gemini 3.1 flash image, text-to-image, or image editing via inference.sh belt.

What you get

Generated or edited image files, belt CLI commands, and grounded multimodal outputs from up to 14 input images.

  • Generated image files
  • Edited image outputs
  • belt CLI command recipes

By the numbers

  • Supports multi-image input up to 14 images
  • Uses Google Gemini 3.1 Flash Image Preview model

Files

SKILL.mdMarkdownGitHub ↗
Install the belt CLI skill: npx skills add belt-sh/cli

Nano Banana 2 - Gemini 3.1 Flash Image Preview

Generate images with Google Gemini 3.1 Flash Image Preview via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions
belt login

belt app run google/gemini-3-1-flash-image-preview --input '{"prompt": "a banana in space, photorealistic"}'

Examples

Basic Text-to-Image

belt app run google/gemini-3-1-flash-image-preview --input '{
  "prompt": "A futuristic cityscape at sunset with flying cars"
}'

Multiple Images

belt app run google/gemini-3-1-flash-image-preview --input '{
  "prompt": "Minimalist logo design for a coffee shop",
  "num_images": 4
}'

Custom Aspect Ratio

belt app run google/gemini-3-1-flash-image-preview --input '{
  "prompt": "Panoramic mountain landscape with northern lights",
  "aspect_ratio": "16:9"
}'

Image Editing (with input images)

belt app run google/gemini-3-1-flash-image-preview --input '{
  "prompt": "Add a rainbow in the sky",
  "images": ["https://example.com/landscape.jpg"]
}'

High Resolution (4K)

belt app run google/gemini-3-1-flash-image-preview --input '{
  "prompt": "Detailed illustration of a medieval castle",
  "resolution": "4K"
}'

With Google Search Grounding

belt app run google/gemini-3-1-flash-image-preview --input '{
  "prompt": "Current weather in Tokyo visualized as an artistic scene",
  "enable_google_search": true
}'

Input Options

ParameterTypeDescription
promptstringRequired. What to generate or change
imagesarrayInput images for editing (up to 14). Supported: JPEG, PNG, WebP
num_imagesintegerNumber of images to generate
aspect_ratiostringOutput ratio: "1:1", "16:9", "9:16", "4:3", "3:4", "auto"
resolutionstring"1K", "2K", "4K" (default: 1K)
output_formatstringOutput format for images
enable_google_searchbooleanEnable real-time info grounding (weather, news, etc.)

Output

FieldTypeDescription
imagesarrayThe generated or edited images
descriptionstringText description or response from the model
output_metaobjectMetadata about inputs/outputs for pricing

Prompt Tips

Styles: photorealistic, illustration, watercolor, oil painting, digital art, anime, 3D render

Composition: close-up, wide shot, aerial view, macro, portrait, landscape

Lighting: natural light, studio lighting, golden hour, dramatic shadows, neon

Details: add specific details about textures, colors, mood, atmosphere

Sample Workflow

# 1. Generate sample input to see all options
belt app sample google/gemini-3-1-flash-image-preview --save input.json

# 2. Edit the prompt
# 3. Run
belt app run google/gemini-3-1-flash-image-preview --input input.json

Python SDK

from inferencesh import inference

client = inference()

# Basic generation
result = client.run({
    "app": "google/gemini-3-1-flash-image-preview@0c7ma1ex",
    "input": {
        "prompt": "A banana in space, photorealistic"
    }
})
print(result["output"])

# Stream live updates
for update in client.run({
    "app": "google/gemini-3-1-flash-image-preview@0c7ma1ex",
    "input": {
        "prompt": "A futuristic cityscape at sunset"
    }
}, stream=True):
    if update.get("progress"):
        print(f"progress: {update['progress']}%")
    if update.get("output"):
        print(f"output: {update['output']}")

Related Skills

# Original Nano Banana (Gemini 3 Pro Image, Gemini 2.5 Flash Image)
npx skills add inference-sh/skills@nano-banana

# Full platform skill (all 250+ apps)
npx skills add inference-sh/skills@infsh-cli

# All image generation models
npx skills add inference-sh/skills@ai-image-generation

Browse all image apps: belt app store --category image

Documentation

Related skills

Forks & variants (2)

Nano Banana 2 has 2 known copies in the catalog totaling 74 installs. They canonicalize to this original listing.

FAQ

What CLI does nano-banana-2 require?

nano-banana-2 requires the inference.sh belt CLI. Install the belt CLI skill with npx skills add belt-sh/cli, then invoke Gemini 3.1 Flash Image Preview generation and editing commands from agent sessions.

How many input images does nano-banana-2 support?

nano-banana-2 supports multi-image input with up to 14 images for Gemini 3.1 Flash Image Preview. The skill also covers text-to-image, image editing, and Google Search grounding workflows.

AI & Agent Buildingagentsautomation

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.