跳到主内容
智客 ZICQ

技能库 智客分类:媒体内容 nano-banana-pro OpenClaw

Nano Banana Pro

以纳米·香蕉Pro(Gemini 3 Pro Image)来生成/编辑图像. 用于创建/修改请求,包括编辑。 支持文本到图像+图像到图像; 1K/2K/4K; 使用 -- input-image.

3104 安装量 · 420 星标

官方网址:ClawHub

技能介绍

先看中文介绍;官方 description 原文单独保留,不改写 SKILL.md。

做什么

以纳米·香蕉Pro(Gemini 3 Pro Image)来生成/编辑图像. 用于创建/修改请求,包括编辑。 支持文本到图像+图像到图像; 1K/2K/4K; 使用 -- input-image.

何时用

官方 description 未单独写出 Use when。按规范,代理会在用户任务与这段 description 的关键词匹配时激活本技能。

代理如何加载

按 Agent Skills 渐进披露:启动时只加载 name 与 description(约 100 token);任务匹配后才读入整份 SKILL.md 正文;scripts/、references/、assets/ 仅在需要时再读。 本文件正文结构:Nano Banana Pro Image Generation & Editing、Usage、Default Workflow (draft → iterate → final)、Resolution Options、API Key、Preflight + Common Failures (fast fixes)。 其中含规范建议的小节:分步指令、输入输出示例。

文件分析

文件分析:除 SKILL.md 外,正文引用了 scripts/generate_image.py,属于带资源的技能包,这些文件按需再读。

官方 description(原文)

Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). Use for image create/modify requests incl. edits. Supports text-to-image + image-to-image; 1K/2K/4K; use --input-image.

Nano Banana Pro Image Generation & EditingUsageDefault Workflow (draft → iterate → final)Resolution OptionsAPI KeyPreflight + Common Failures (fast fixes)Filename GenerationImage EditingPrompt HandlingPrompt Templates (high hit-rate)OutputExamples

来源分类:ClawHub Image Generation

SKILL.md 与 Agent 调用

官方规范 ↗
name
nano-banana-pro
description
Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). Use for image create/modify requests incl. edits. Supports text-to-image + image-to-image; 1K/2K/4K; use --input-image.
  1. 发现技能客户端向 Agent 提供名称与描述目录。
  2. 匹配与调用用户指定或任务匹配后,载入 SKILL.md 指令。
  3. 按需加载按步骤读取参考文档、使用脚本与素材。
指令中引用的文件 · 1
  • scripts/generate_image.py

以下路径提取自原文;文件是否齐全请以来源仓库中的完整目录为准。

具体调用语法与可用工具以目标 Agent 客户端为准。 查看调用机制说明 ↗

安装这个技能

Skills CLI ↗

先选择目标 Agent 和安装范围,保留技能包的附属文件,安装后检查客户端能否发现该技能。

该技能引用了附属文件,请从来源获取完整目录;仅复制 SKILL.md 可能缺少依赖。

交给 Agent 安装

复制安装指令给支持 Agent Skills 的代理,确认其中的目标目录与客户端匹配。

把 Agent Skill「nano-banana-pro」安装到我的项目:SKILL.md 原文与官方 description 见 https://zicq.com/zh/skills/skl-c5924f7fad2e8655-Nano-Banana-Pro.html
请存为 .cursor/skills/nano-banana-pro/SKILL.md 或 .claude/skills/nano-banana-pro/SKILL.md,frontmatter 的 name 与 description 保持原样,不要改写。
该技能还带 scripts/、references/、assets/ 等文件,请从 https://clawhub.ai/skills/nano-banana-pro 取完整目录,不要只建一个 SKILL.md。

当前没有明确的 GitHub 技能包地址,请按来源页面的安装器说明操作。

ClawHub ↗

阅读排版
--- name: nano-banana-pro description: Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). Use for image create/modify requests incl. edits. Supports text-to-image + image-to-image; 1K/2K/4K; use --input-image. --- # Nano Banana Pro Image Generation & Editing Generate new images or edit existing ones using Google's Nano Banana Pro API (Gemini 3 Pro Image). ## Usage Run the script using absolute path (do NOT cd to skill directory first): **Generate new image:** ```bash uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "your image description" --filename "output-name.png" [--resolution 1K|2K|4K] [--api-key KEY] ``` **Edit existing image:** ```bash uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "editing instructions" --filename "output-name.png" --input-image "path/to/input.png" [--resolution 1K|2K|4K] [--api-key KEY] ``` **Important:** Always run from the user's current working directory so images are saved where the user is working, not in the skill directory. ## Default Workflow (draft → iterate → final) Goal: fast iteration without burning time on 4K until the prompt is correct. - Draft (1K): quick feedback loop - `uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "" --filename "yyyy-mm-dd-hh-mm-ss-draft.png" --resolution 1K` - Iterate: adjust prompt in small diffs; keep filename new per run - If editing: keep the same `--input-image` for every iteration until you’re happy. - Final (4K): only when prompt is locked - `uv run ~/.codex/skills/nano-banana-pro/scripts/generate_image.py --prompt "" --filename "yyyy-mm-dd-hh-mm-ss-final.png" --resolution 4K` ## Resolution Options The Gemini 3 Pro Image API supports three resolutions (uppercase K required): - **1K** (default) - ~1024px resolution - **2K** - ~2048px resolution - **4K** - ~4096px resolution Map user requests to API parameters: - No mention of resolution → `1K` - "low resolution", "1080", "1080p", "1K" → `1K` - "2K", "2048", "normal", "medium resolution" → `2K` - "high resolution", "high-res", "hi-res", "4K", "ultra" → `4K` ## API Key The script checks for API key in this order: 1. `--api-key` argument (use if user provided key in chat) 2. `GEMINI_API_KEY` environment variable If neither is available, the script exits with an error message. ## Preflight + Common Failures (fast fixes) - Preflight: - `command -v uv` (must exist) - `test -n \"$GEMINI_API_KEY\"` (or pass `--api-key`) - If editing: `test -f \"path/to/input.png\"` - Common failures: - `Error: No API key provided.` → set `GEMINI_API_KEY` or pass `--api-key` - `Error loading input image:` → wrong path / unreadable file; verify `--input-image` points to a real image - “quota/permission/403” style API errors → wrong key, no access, or quota exceeded; try a different key/account ## Filename Generation Generate filenames with the pattern: `yyyy-mm-dd-hh-mm-ss-name.png` **Format:** `{timestamp}-{descriptive-name}.png` - Timestamp: Current date/time in format `yyyy-mm-dd-hh-mm-ss` (24-hour format) - Name: Descriptive lowercase text with hyphens - Keep the descriptive part concise (1-5 words typically) - Use context from user's prompt or conversation - If unclear, use random identifier (e.g., `x9k2`, `a7b3`) Examples: - Prompt "A serene Japanese garden" → `2025-11-23-14-23-05-japanese-garden.png` - Prompt "sunset over mountains" → `2025-11-23-15-30-12-sunset-mountains.png` - Prompt "create an image of a robot" → `2025-11-23-16-45-33-robot.png` - Unclear context → `2025-11-23-17-12-48-x9k2.png` ## Image Editing When the user wants to modify an existing image: 1. Check if they provide an image path or reference an image in the current directory 2. Use `--input-image` parameter with the path to the image 3. The prompt should contain editing instructions (e.g., "make the sky more dramatic", "remove the person", "change to cartoon style") 4. Common editing tasks: add/remove elements, change style, adjust colors, blur background, etc. ## Prompt Handling **For generation:** Pass user's image description as-is to `--prompt`. Only rework if clearly insufficient. **For editing:** Pass editing instructions in `--prompt` (e.g., "add a rainbow in the sky", "make it look like a watercolor painting") Preserve user's creative intent in both cases. ## Prompt Templates (high hit-rate) Use templates when the user is vague or when edits must be precise. - Generation template: - “Create an image of: . Style:

相关技能

媒体内容

Youtube Watcher

从YouTube视频中获取并读取文字记录. 需要汇总视频时使用,回答有关视频内容的问题,或者从中提取信息.

媒体内容

Atxp

Access ATXP支付API工具用于网页搜索,AI图像生成,音乐创建,视频生成,X/Twitter搜索,电子邮件,代理账户管理. 当用户需要实时网络搜索时使用,AI生成的媒体(图像,音乐,视频),X/Twitter搜索,发送/接收电子邮件,或创建和资助代理账户. 需要通过“ …

媒体内容

Youtube

YouTube Data API与管理的OAuth的集成. 搜索视频,管理播放列表,访问频道数据,并与评论互动. 用户想与YouTube互动时使用此技能. 对于其他第三方应用,使用api-gateway技能(https://clawhub.ai/byungkyu/api-gate…