🎯

ai-image-generation

🎯Skill

from prime-skills/runcomfy-agent-skills

Installation

Vibe Index InstallInstalls to .claude/skills/
$npx vibeindex add prime-skills/runcomfy-agent-skills --skill ai-image-generation
skills.sh Install⚠ Installs to .agents/skills/
$npx skills add prime-skills/runcomfy-agent-skills --skill ai-image-generation
Manual InstallCopy SKILL.md content and save to the path below
path
~/.claude/skills/ai-image-generation/SKILL.md
VibeIndex|
What it does
|

A Claude Code skill that acts as a smart router for AI image generation and editing across 11+ models on the RunComfy platform. Supports text-to-image and image-to-image modes with models including FLUX 2, GPT Image 2, Seedream 5, Nano Banana 2, and more.

Overview

AI Image Generation is a Claude Code skill that provides a unified interface for generating and editing images across 11+ AI models through the RunComfy CLI. It acts as a smart router, automatically selecting the best model based on the user's intent, whether that is typography precision, photorealistic portraits, sub-second iteration, multi-reference brand styling, or open-weights workflows. The skill covers both text-to-image (t2i) and image-to-image/edit (i2i) endpoints, with models including FLUX 2 (Klein, Pro, Dev, Flash, Turbo, Max), Google Nano Banana 2/Pro, OpenAI GPT Image 2, ByteDance Seedream 5/4.5/4.0, Alibaba Qwen Image and Z-Image Turbo, and Wan 2.7.

Key Features

  • Smart model routing: Recommends the best model based on user intent with a decision matrix covering typography, photorealism, speed, open weights, and brand consistency
  • 11+ model catalog: Includes FLUX 2 family (Klein 9B/4B, Pro, Dev, Flash, Turbo, Max), GPT Image 2, Nano Banana 2/Pro, Seedream 5/4.5/4.0, Dreamina 4.0, Qwen Image, Z-Image Turbo, and Wan 2.7
  • Text-to-image and image-to-image: Covers both generation from prompts and editing/restyling of existing images across supported models
  • Per-model prompt patterns: Documents each model's optimal prompting style, parameter schema, supported resolutions, and pricing
  • Single CLI workflow: All models accessed via runcomfy run with consistent authentication, input format, and output directory handling
  • Speed-optimized options: Includes sub-second models like FLUX 2 Klein 4B and Z-Image Turbo for rapid iteration and A/B testing

Who is this for?

  • Designers and developers who need to generate images across different AI models without managing separate APIs or credentials
  • Teams building image generation pipelines that require switching between models for different tasks (typography, portraits, editing)
  • Content creators looking for a single skill that covers the full spectrum from fast drafts to high-fidelity final images
📦

Same repository

prime-skills/runcomfy-agent-skills(25 items)

ai-image-generation

SKILL.md

350,574Installs
-
AddedAug 11, 2026

More from this repository10

🎯
video-edit🎯Skill

A smart intent-routing skill for video editing on RunComfy that automatically selects the best model (Wan 2.7 Edit-Video, Kling 2.6 Pro Motion Control, or Lucy Edit Restyle) based on what the user wants to do, from general restyling and background swaps to precise motion transfer and outfit changes.

🎯
image-to-video🎯Skill

A smart intent-routing skill that animates still images on RunComfy by selecting the optimal model: HappyHorse 1.0 I2V for general portrait and product animation with native audio, Wan 2.7 for custom-voiceover lip-sync, or Seedance 2.0 Pro for multi-modal composition with image, video, and audio references.

🎯
nano-banana-2🎯Skill

A skill for generating images with Google Nano Banana 2, the flash-tier text-to-image model in the Gemini family on RunComfy, optimized for rapid iteration, batch ideation, in-image typography rendering, and optional web-grounded context with support for resolutions from 0.5K to 4K.

🎯
image-edit🎯Skill

A smart intent-routing skill for image editing on RunComfy that selects the optimal model from Nano Banana Edit (batch up to 20 images), GPT Image 2 Edit (multilingual text rewrite, multi-ref composition), Flux Kontext Pro (single-shot precise local edit), or Z-Image Turbo Inpaint (mask-driven region edit) based on user intent.

🎯
flux-kontext🎯Skill

A Claude Code skill that enables precise single-image editing using Black Forest Labs's Flux 1 Kontext Pro model through the RunComfy CLI, with built-in prompting patterns for high-fidelity local edits.

🎯
wan-2-7🎯Skill

A skill for generating text-to-video with Wan 2.7, Wan-AI's flagship motion model on RunComfy, featuring multi-reference conditioning, audio-driven lip-sync via custom audio tracks, smooth transition physics, prompt expansion control, and up to 1080p output at 15 seconds duration.

🎯
happyhorse-1-0🎯Skill

A Claude Code skill for generating text-to-video clips with HappyHorse 1.0 via the RunComfy CLI. HappyHorse 1.0 is currently ranked #1 on the Artificial Analysis Video Arena, producing native 1080p video with in-pass synchronized audio and multi-shot character consistency.

🎯
kling-3-0🎯Skill

A Claude Code skill for generating video with Kling 3.0 via the RunComfy CLI. Covers all six Kling 3.0 endpoints across three rendering tiers (Standard, Pro, 4K) and two modes (text-to-video, image-to-video), with native synchronized audio and multi-shot character consistency.

🎯
codex-pet🎯Skill

A Claude Code skill that generates custom OpenAI Codex Pets from a single reference image via the RunComfy CLI. It produces a Codex-compatible spritesheet (1536x1872, 8 columns x 9 rows) and pet.json manifest using GPT Image 2 and ImageMagick, without requiring Codex Pro or an OPENAI_API_KEY.

🎯
ai-video-generation🎯Skill

A Claude Code skill that acts as a smart router for AI video generation across the full RunComfy model catalog. Supports text-to-video, image-to-video, and video-extend modes with models including HappyHorse 1.0, Kling 3.0, Wan 2.7, Seedance, Veo 3.1, and more.