通义万相 2.5D 横幅插画

Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). Use for image create/modify requests incl. edits. Supports text-to-image + image-to-image; 1K...

installs

stars

karma

SkillRank score ↗

7.9/ 10

evaluated by implexa, claude-haiku-4-5 · 2026-05-26

wenxiang-2d5-banner generates or edits images via gemini 3 pro api, supporting 1k/2k/4k resolutions with both text-to-image and image-to-image workflows. uses iterative draft-to-final pattern and explicit prompt templates for consistency.

structure

9.0

trigger phrases

6.0

procedure

9.0

edge cases

7.0

documentation

8.0

strengths

view original SKILL.md from clawhubclick to expand

---
name: nano-banana-pro
description: Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). Use for image create/modify requests incl. edits. Supports text-to-image + image-to-image; 1K/2K/4K; use --input-image.
---

# Nano Banana Pro Image Generation & Editing

Generate new images or edit existing ones using Google's Nano Banana Pro API (Gemini 3 Pro Image).

## Usage

Run the script using absolute path (do NOT cd to skill directory first):

**Generate new image:**
```bash
uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "your image description" --filename "output-name.png" [--resolution 1K|2K|4K] [--api-key KEY]
```

**Edit existing image:**
```bash
uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "editing instructions" --filename "output-name.png" --input-image "path/to/input.png" [--resolution 1K|2K|4K] [--api-key KEY]
```

**Important:** Always run from the user's current working directory so images are saved where the user is working, not in the skill directory.

## Default Workflow (draft → iterate → final)

Goal: fast iteration without burning time on 4K until the prompt is correct.

- Draft (1K): quick feedback loop
  - `uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "<draft prompt>" --filename "yyyy-mm-dd-hh-mm-ss-draft.png" --resolution 1K`
- Iterate: adjust prompt in small diffs; keep filename new per run
  - If editing: keep the same `--input-image` for every iteration until you’re happy.
- Final (4K): only when prompt is locked
  - `uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "<final prompt>" --filename "yyyy-mm-dd-hh-mm-ss-final.png" --resolution 4K`

## Resolution Options

The Gemini 3 Pro Image API supports three resolutions (uppercase K required):

- **1K** (default) - ~1024px resolution
- **2K** - ~2048px resolution
- **4K** - ~4096px resolution

Map user requests to API parameters:
- No mention of resolution → `1K`
- "low resolution", "1080", "1080p", "1K" → `1K`
- "2K", "2048", "normal", "medium resolution" → `2K`
- "high resolution", "high-res", "hi-res", "4K", "ultra" → `4K`

## API Key

The script checks for API key in this order:
1. `--api-key` argument (use if user provided key in chat)
2. `GEMINI_API_KEY` environment variable

If neither is available, the script exits with an error message.

## Preflight + Common Failures (fast fixes)

- Preflight:
  - `command -v uv` (must exist)
  - `test -n \"$GEMINI_API_KEY\"` (or pass `--api-key`)
  - If editing: `test -f \"path/to/input.png\"`

- Common failures:
  - `Error: No API key provided.` → set `GEMINI_API_KEY` or pass `--api-key`
  - `Error loading input image:` → wrong path / unreadable file; verify `--input-image` points to a real image
  - “quota/permission/403” style API errors → wrong key, no access, or quota exceeded; try a different key/account

## Filename Generation

Generate filenames with the pattern: `yyyy-mm-dd-hh-mm-ss-name.png`

**Format:** `{timestamp}-{descriptive-name}.png`
- Timestamp: Current date/time in format `yyyy-mm-dd-hh-mm-ss` (24-hour format)
- Name: Descriptive lowercase text with hyphens
- Keep the descriptive part concise (1-5 words typically)
- Use context from user's prompt or conversation
- If unclear, use random identifier (e.g., `x9k2`, `a7b3`)

Examples:
- Prompt "A serene Japanese garden" → `2025-11-23-14-23-05-japanese-garden.png`
- Prompt "sunset over mountains" → `2025-11-23-15-30-12-sunset-mountains.png`
- Prompt "create an image of a robot" → `2025-11-23-16-45-33-robot.png`
- Unclear context → `2025-11-23-17-12-48-x9k2.png`

## Image Editing

When the user wants to modify an existing image:
1. Check if they provide an image path or reference an image in the current directory
2. Use `--input-image` parameter with the path to the image
3. The prompt should contain editing instructions (e.g., "make the sky more dramatic", "remove the person", "change to cartoon style")
4. Common editing tasks: add/remove elements, change style, adjust colors, blur background, etc.

## Prompt Handling

**For generation:** Pass user's image description as-is to `--prompt`. Only rework if clearly insufficient.

**For editing:** Pass editing instructions in `--prompt` (e.g., "add a rainbow in the sky", "make it look like a watercolor painting")

Preserve user's creative intent in both cases.

## Prompt Templates (high hit-rate)

Use templates when the user is vague or when edits must be precise.

- Generation template:
  - “Create an image of: <subject>. Style: <style>. Composition: <camera/shot>. Lighting: <lighting>. Background: <background>. Color palette: <palette>. Avoid: <list>.”

- Editing template (preserve everything else):
  - “Change ONLY: <single change>. Keep identical: subject, composition/crop, pose, lighting, color palette, background, text, and overall style. Do not add new objects. If text exists, keep it unchanged.”

## Output

- Saves PNG to current directory (or specified path if filename includes directory)
- Script outputs the full path to the generated image
- **Do not read the image back** - just inform the user of the saved path

## Examples

**Generate new image:**
```bash
uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "A serene Japanese garden with cherry blossoms" --filename "2025-11-23-14-23-05-japanese-garden.png" --resolution 4K
```

**Edit existing image:**
```bash
uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "make the sky more dramatic with storm clouds" --filename "2025-11-23-14-25-30-dramatic-sky.png" --input-image "original-photo.jpg" --resolution 2K
```

don't have the plugin yet? install it then click "run inline in claude" again.

restructured original into implexa's 6-component format, formalized api key and network edge cases, clarified decision logic for resolution mapping and workflow branching, added explicit preflight checks and error handling, documented gemini api rate limits and quota behavior, kept original author's iterative draft-iterate-final workflow intact.

Wenxiang 2.5D Banner Image Generation & Editing

Item: 通义万相 2.5D 横幅插画
Rating: 7.9
Author: Implexa

generate new images or edit existing ones using google's gemini 3 pro image api. designed for fast iteration (draft at 1k, lock prompt, then render final at 4k). supports both text-to-image and image-to-image workflows.

intent

use this skill when you need to create or modify images via text description or by editing an existing image file. ideal for banner art, illustrations, and iterative design workflows where you want to test prompts quickly before committing to high-resolution renders. the skill enforces a deliberate workflow: start low-res for feedback, iterate on prompt, then go high-res only when the prompt is locked.

inputs

required:

GEMINI_API_KEY environment variable or --api-key CLI argument. obtain from google ai studio (https://aistudio.google.com/apikey). no special scopes needed, just an active key with image generation quota.
uv command (python package runner). verify with command -v uv.

optional (text-to-image):

--prompt (string): image description or generation instruction. required unless doing image editing.
--filename (string): output filename in format yyyy-mm-dd-hh-mm-ss-name.png. if omitted, script generates one automatically.
--resolution (enum: 1K|2K|4K): defaults to 1K. case-sensitive.

optional (image-to-image):

--input-image (filepath): path to image file to edit (png, jpg, etc.). when provided, --prompt becomes editing instructions rather than generation prompt.
--filename, --resolution: same as text-to-image.

script location:

~/.codex/skills/wenxiang-2d5-banner/scripts/generate_image.py (use absolute path, do not cd into skill directory first).

external connections:

google gemini 3 pro image api (requires active internet, rate limited to ~100 requests/day per free tier account, higher limits on paid plans).

procedure

preflight checks (run before invoking script):
- confirm uv is installed: command -v uv.
- confirm api key exists: either echo $GEMINI_API_KEY returns a non-empty string, or you will pass --api-key to the script.
- if editing, confirm input image exists: test -f "path/to/image.png".
determine resolution tier based on user intent:
- user says "draft", "quick", "test", or provides no resolution hint → use 1K.
- user says "1080", "1080p", "low res" → use 1K.
- user says "2k", "2048", "normal", "medium" → use 2K.
- user says "4k", "high res", "hi-res", "high-resolution", "ultra", "final" → use 4K.
map user request to workflow:
- text-to-image (generation): user wants to create a new image from description.
- image-to-image (editing): user provides an existing image and editing instructions.
generate filename if user did not provide one:
- format: yyyy-mm-dd-hh-mm-ss-{descriptive-name}.png.
- timestamp: current date/time in 24-hour format (e.g., 2025-11-23-14-23-05).
- descriptive-name: 1-5 words, lowercase, hyphen-separated, derived from prompt or context. if unclear, use random identifier (e.g., x9k2, a7b3).
- examples: 2025-11-23-14-23-05-japanese-garden.png, 2025-11-23-15-30-12-sunset-mountains.png.
prepare prompt:
- for generation: use user's description as-is unless it is vague or underspecified. if vague, apply generation template: "create an image of: . style:

通义万相 2.5D 横幅插画

related skills

Wenxiang 2.5D Banner Image Generation & Editing

intent

inputs

procedure