๐ŸŽฏ

gpt-image-edit

๐ŸŽฏSkill

from agentspace-so/runcomfy-agent-skills

VibeIndex|
What it does
|

Edit images with OpenAI GPT Image 2 on RunComfy, excelling at multilingual in-image text editing across any script (Latin, kana, CJK, Cyrillic, Arabic) and multi-reference composition with up to 10 input images. Ideal for identity-preserving edits and layout-precise repositioning.

Overview

A RunComfy skill for editing images with OpenAI GPT Image 2's edit endpoint (ChatGPT Images 2.0 image-to-image). It excels at multilingual in-image text editing across all writing systems (Latin, kana, CJK, Cyrillic, Arabic) and supports up to 10 reference images per call for multi-reference composition. The skill is strongest in class at preserving identity through targeted edits and rewriting embedded text, making it the go-to choice when typography precision and multilingual support matter.

Key Features

  • Best-in-class multilingual in-image text editing across Latin, kana, CJK, Cyrillic, and Arabic scripts
  • Identity preservation through targeted edits with "keep X unchanged" prompt patterns
  • Up to 10 reference images per call: first image is primary, rest are auxiliary for composition cues
  • Layout-precise editing: move headlines, swap CTAs, and rearrange visual elements with spatial accuracy
  • Built-in routing guidance for when to use Nano Banana Edit, Flux Kontext, or GPT Image 2 text-to-image instead

Who is this for?

  • Global marketing teams that need to localize ad creatives by rewriting in-image text across multiple languages and writing systems
  • Brand designers creating translated headline variants while maintaining exact visual identity and layout
  • E-commerce teams updating product labels, signage, and embedded text across diverse markets
๐Ÿ“ฆ

Same repository

agentspace-so/runcomfy-agent-skills(30 items)

gpt-image-edit

Installation

Vibe Index InstallInstalls to .claude/skills/
npx vibeindex add agentspace-so/runcomfy-agent-skills --skill gpt-image-edit
skills.sh Installโš  Installs to .agents/skills/
npx skills add agentspace-so/runcomfy-agent-skills --skill gpt-image-edit
Manual InstallCopy SKILL.md content and save to the path below
~/.claude/skills/gpt-image-edit/SKILL.md

SKILL.md

324,146Installs
26
-
Last UpdatedMay 15, 2026

More from this repository10

๐ŸŽฏ
nano-banana-2๐ŸŽฏSkill

A RunComfy skill that generates images using Google Nano Banana 2, the flash-tier text-to-image model in the Gemini family. Optimized for rapid iteration, social thumbnails, and in-image typography with configurable resolution tiers and safety tolerance.

๐ŸŽฏ
image-edit๐ŸŽฏSkill

A smart intent-routing skill for image editing on RunComfy that selects the best model based on the editing task. Routes to Nano Banana Edit for batch edits up to 20 images, GPT Image 2 for multilingual text rewrite, Flux Kontext Pro for single-shot precise edits, or Z-Image Turbo for mask-driven inpainting.

๐ŸŽฏ
kling-3-0๐ŸŽฏSkill

Provides Kling 3.0 video generation on RunComfy, covering all six endpoints across three quality tiers (Standard, Pro, 4K) and two modes (text-to-video, image-to-video) for Kuaishou's third-generation cinematic video model with native synchronized audio.

๐ŸŽฏ
nano-banana-edit๐ŸŽฏSkill

Edit images with Google Nano Banana 2 on RunComfy, supporting batch edits of up to 20 images per call with strong identity preservation. Features localized edits using spatial language, background swaps, and configurable resolution up to 4K.

๐ŸŽฏ
wan-2-7๐ŸŽฏSkill

Generate text-to-video with Wan-AI's Wan 2.7 on RunComfy, featuring multi-reference conditioning and audio-driven lip-sync via custom audio tracks. Supports prompt expansion, negative prompts, and up to 1080p resolution through the RunComfy CLI.

๐ŸŽฏ
happyhorse-1-0๐ŸŽฏSkill

Generate text-to-video with HappyHorse 1.0 on RunComfy, currently ranked #1 on Artificial Analysis Video Arena. Supports native 1080p with in-pass synchronized audio, multi-shot character consistency, and 6-language prompt support via the RunComfy CLI.

๐ŸŽฏ
seedance-v2๐ŸŽฏSkill

Generate cinematic short-form video with ByteDance Seedance 2.0 Pro on RunComfy, supporting multi-modal references including up to 9 images, 3 videos, and 3 audio tracks. Features native lip-synced audio generation and is ideal for brand-consistent multi-language narratives.

๐ŸŽฏ
flux-2-klein๐ŸŽฏSkill

A RunComfy skill for generating images with Black Forest Labs' Flux 2 Klein, the distilled low-latency variant of Flux 2. Supports 9B and 4B model variants with sub-second inference for real-time art direction, rapid concepting, and multi-reference brand styling.

๐ŸŽฏ
runcomfy-cli๐ŸŽฏSkill

The foundation skill for the RunComfy platform, providing a single CLI to install, authenticate, and invoke hundreds of model endpoints including image generation, video, face-swap, lip-sync, and LoRA training.

๐ŸŽฏ
face-swap๐ŸŽฏSkill

A Claude Code skill for swapping faces in images and videos via RunComfy CLI, routing across multiple model endpoints including Wan 2-2 Animate, GPT Image 2 Edit, Flux Kontext, and Kling Motion Control based on use case.