nano-banana-edit
๐ฏSkillfrom agentspace-so/runcomfy-agent-skills
Edit images with Google Nano Banana 2 on RunComfy, supporting batch edits of up to 20 images per call with strong identity preservation. Features localized edits using spatial language, background swaps, and configurable resolution up to 4K.
Overview
A RunComfy skill for editing images with Google Nano Banana 2's image-to-image edit endpoint. Its standout capability is batch editing of up to 20 images per call with strong identity preservation, making it well-suited for consistent product photography edits across entire SKU galleries. The skill supports spatial language for localized edits ("change only the left object") and preserves subject identity through "keep X unchanged" prompt patterns. It bundles the model's schema and prompting best practices for sharper results.
Key Features
- Batch editing of up to 20 input images per call for consistent multi-image workflows
- Strong identity preservation for edits like background swap, clothing change, and localized object modifications
- Spatial language support for directing edits to specific regions ("upper-right corner", "the left object")
- Configurable resolution (0.5K to 4K), aspect ratio, output format, and safety tolerance
- Built-in routing guidance for when to use GPT Image 2 Edit, Flux Kontext, or Nano Banana 2 text-to-image instead
Who is this for?
- E-commerce teams that need to batch-edit product images (background swaps, lighting adjustments) across entire catalogs consistently
- Visual content creators producing A/B variant sets from a single source image with consistent identity
- Developers integrating batch image editing into automated pipelines through the RunComfy CLI
Same repository
agentspace-so/runcomfy-agent-skills(30 items)
Installation
npx vibeindex add agentspace-so/runcomfy-agent-skills --skill nano-banana-editnpx skills add agentspace-so/runcomfy-agent-skills --skill nano-banana-edit~/.claude/skills/nano-banana-edit/SKILL.mdSKILL.md
More from this repository10
A RunComfy skill that generates images using Google Nano Banana 2, the flash-tier text-to-image model in the Gemini family. Optimized for rapid iteration, social thumbnails, and in-image typography with configurable resolution tiers and safety tolerance.
A smart intent-routing skill for image editing on RunComfy that selects the best model based on the editing task. Routes to Nano Banana Edit for batch edits up to 20 images, GPT Image 2 for multilingual text rewrite, Flux Kontext Pro for single-shot precise edits, or Z-Image Turbo for mask-driven inpainting.
Provides Kling 3.0 video generation on RunComfy, covering all six endpoints across three quality tiers (Standard, Pro, 4K) and two modes (text-to-video, image-to-video) for Kuaishou's third-generation cinematic video model with native synchronized audio.
Generate text-to-video with Wan-AI's Wan 2.7 on RunComfy, featuring multi-reference conditioning and audio-driven lip-sync via custom audio tracks. Supports prompt expansion, negative prompts, and up to 1080p resolution through the RunComfy CLI.
Edit images with OpenAI GPT Image 2 on RunComfy, excelling at multilingual in-image text editing across any script (Latin, kana, CJK, Cyrillic, Arabic) and multi-reference composition with up to 10 input images. Ideal for identity-preserving edits and layout-precise repositioning.
Generate text-to-video with HappyHorse 1.0 on RunComfy, currently ranked #1 on Artificial Analysis Video Arena. Supports native 1080p with in-pass synchronized audio, multi-shot character consistency, and 6-language prompt support via the RunComfy CLI.
Generate cinematic short-form video with ByteDance Seedance 2.0 Pro on RunComfy, supporting multi-modal references including up to 9 images, 3 videos, and 3 audio tracks. Features native lip-synced audio generation and is ideal for brand-consistent multi-language narratives.
A RunComfy skill for generating images with Black Forest Labs' Flux 2 Klein, the distilled low-latency variant of Flux 2. Supports 9B and 4B model variants with sub-second inference for real-time art direction, rapid concepting, and multi-reference brand styling.
The foundation skill for the RunComfy platform, providing a single CLI to install, authenticate, and invoke hundreds of model endpoints including image generation, video, face-swap, lip-sync, and LoRA training.
A Claude Code skill for swapping faces in images and videos via RunComfy CLI, routing across multiple model endpoints including Wan 2-2 Animate, GPT Image 2 Edit, Flux Kontext, and Kling Motion Control based on use case.