All skills
inference-shell avatar

/nano-banana-2

@8381375

Generate images with Google Gemini 3.1 Flash Image Preview (Nano Banana 2) via inference.sh CLI. Capabilities: text-to-image, image editing, multi-image input (up to 14 images), Google Search grounding. Triggers: nano banana 2, nanobanana 2, gemini 3.1 flash image, gemini 3 1 flash image preview, google image generation

Use this Skill: https://skilld.dev/gh/inference-shell/skills/nano-banana-2

This session only. Nothing lands on disk.

SKILL.md

≈84 tokens always: the name and description. ≈1.1k when used: this file.

Install the belt CLI skill: npx skills add belt-sh/cli

Nano Banana 2 - Gemini 3.1 Flash Image Preview

Generate images with Google Gemini 3.1 Flash Image Preview via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

belt app run google/gemini-3-1-flash-image --input '{"prompt": "a banana in space, photorealistic"}'

Examples

Basic Text-to-Image

belt app run google/gemini-3-1-flash-image --input '{
  "prompt": "A futuristic cityscape at sunset with flying cars"
}'

Multiple Images

belt app run google/gemini-3-1-flash-image --input '{
  "prompt": "Minimalist logo design for a coffee shop",
  "num_images": 4
}'

Custom Aspect Ratio

belt app run google/gemini-3-1-flash-image --input '{
  "prompt": "Panoramic mountain landscape with northern lights",
  "aspect_ratio": "16:9"
}'

Image Editing (with input images)

belt app run google/gemini-3-1-flash-image --input '{
  "prompt": "Add a rainbow in the sky",
  "images": ["https://example.com/landscape.jpg"]
}'

High Resolution (4K)

belt app run google/gemini-3-1-flash-image --input '{
  "prompt": "Detailed illustration of a medieval castle",
  "resolution": "4K"
}'

With Google Search Grounding

belt app run google/gemini-3-1-flash-image --input '{
  "prompt": "Current weather in Tokyo visualized as an artistic scene",
  "enable_google_search": true
}'

Input Options

Parameter Type Description
prompt string Required. What to generate or change
images array Input images for editing (up to 14). Supported: JPEG, PNG, WebP
num_images integer Number of images to generate
aspect_ratio string Output ratio: "1:1", "16:9", "9:16", "4:3", "3:4", "auto"
resolution string "1K", "2K", "4K" (default: 1K)
output_format string Output format for images
enable_google_search boolean Enable real-time info grounding (weather, news, etc.)

Output

Field Type Description
images array The generated or edited images
description string Text description or response from the model
output_meta object Metadata about inputs/outputs for pricing

Prompt Tips

Styles: photorealistic, illustration, watercolor, oil painting, digital art, anime, 3D render

Composition: close-up, wide shot, aerial view, macro, portrait, landscape

Lighting: natural light, studio lighting, golden hour, dramatic shadows, neon

Details: add specific details about textures, colors, mood, atmosphere

Sample Workflow

# 1. Generate sample input to see all options
belt app sample google/gemini-3-1-flash-image --save input.json

# 2. Edit the prompt
# 3. Run
belt app run google/gemini-3-1-flash-image --input input.json

Python SDK

from inferencesh import inference

client = inference()

# Basic generation
result = client.run({
    "app": "google/gemini-3-1-flash-image@0c7ma1ex",
    "input": {
        "prompt": "A banana in space, photorealistic"
    }
})
print(result["output"])

# Stream live updates
for update in client.run({
    "app": "google/gemini-3-1-flash-image@0c7ma1ex",
    "input": {
        "prompt": "A futuristic cityscape at sunset"
    }
}, stream=True):
    if update.get("progress"):
        print(f"progress: {update['progress']}%")
    if update.get("output"):
        print(f"output: {update['output']}")

Related Skills

# Original Nano Banana (Gemini 3 Pro Image, Gemini 2.5 Flash Image)
npx skills add inference-sh/skills@nano-banana

# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli

# All image generation models
npx skills add inference-sh/skills@ai-image-generation

Browse all image apps: belt app list --category image

Documentation

Source: SKILL.md on GitHub

1 warning12d5 checks · Risk SAFE
  • Gen Agent Trust Hub12d

    This skill provides instructions and examples for using the inference.sh 'belt' CLI to generate images via Google Gemini models. It correctly scopes its tool usage and refers to the vendor's own official resources and documentation.

  • Socket12d

    1 alert: gptAnomaly

  • Snyk12d

    Risk: LOW · No issues

  • Runlayer6mo

    1 file scanned · No issues

  • ZeroLeaks5mo

    Score: 93/100 · 2 sections analyzed

Signed by skilld at 8381375. This ties the file your Agent reads to that commit on GitHub. It does not review the instructions.

Last checked against GitHub last week.

Activeupdated 2 weeks ago
What it can do
Runs commands
All 1 allowed tools
Bash(belt *)
  • gemini
  • image-generation
  • text-to-image
  • image-editing
  • google
  • inference-sh
  • belt
  • bash

README badge

README badge for inference-shell/skills/nano-banana-2

Generates images with Google Gemini 3.1 Flash Image Preview via the inference.sh CLI, supporting text-to-image, image editing with up to 14 input images, custom aspect ratios, and real-time Google Search grounding. Use the belt command-line tool to run the model with parameters like prompt, resolution (up to 4K), and num_images.

Generated from the current SKILL.md.

What model does this skill use?
Google Gemini 3.1 Flash Image Preview (also called Nano Banana 2), accessed via the inference.sh CLI.
Can I edit existing images with this skill?
Yes. Pass up to 14 images in the `images` parameter along with a prompt describing the edits you want.
What image formats does this support?
Input images can be JPEG, PNG, or WebP. Output format is configurable via the `output_format` parameter.
Does this work with real-time information?
Yes, if you set `enable_google_search: true`, the model can ground generations in current weather, news, and other live data.
What aspect ratios and resolutions are available?
Aspect ratios include 1:1, 16:9, 9:16, 4:3, 3:4, and auto. Resolutions range from 1K (default) to 2K and 4K.

Generated from the current SKILL.md. These answers refresh after source changes.