๐ŸŽฏ

image-to-video

๐ŸŽฏSkill

from prime-skills/runcomfy-agent-skills

Installation

Vibe Index InstallInstalls to .claude/skills/
$npx vibeindex add prime-skills/runcomfy-agent-skills --skill image-to-video
skills.sh Installโš  Installs to .agents/skills/
$npx skills add prime-skills/runcomfy-agent-skills --skill image-to-video
Manual InstallCopy SKILL.md content and save to the path below
path
~/.claude/skills/image-to-video/SKILL.md
VibeIndex|
What it does
|

A smart intent-routing skill that animates still images on RunComfy by selecting the optimal model: HappyHorse 1.0 I2V for general portrait and product animation with native audio, Wan 2.7 for custom-voiceover lip-sync, or Seedance 2.0 Pro for multi-modal composition with image, video, and audio references.

Overview

Image-to-Video is an intent-routing skill for the RunComfy platform that converts still images into animated video clips by automatically selecting the best model for the task. It routes to HappyHorse 1.0 I2V (ranked #1 on Artificial Analysis Arena with Elo 1392) for general portrait and product animations with native audio synthesis, Wan 2.7 with audio_url for lip-syncing to custom voiceover tracks, or Seedance 2.0 Pro for multi-modal compositions combining up to 9 image references, 3 video references, and 3 audio references in a single call.

Key Features

  • HappyHorse 1.0 I2V route: Arena #1 model for portrait and product animation with native synchronized audio, up to 1080p and 15 seconds, strong facial fidelity and geometry preservation
  • Wan 2.7 lip-sync route: Accepts custom WAV/MP3 audio tracks (3-30s) to drive lip-sync on generated talking-head clips, enabling multi-language dub variants from the same prompt
  • Seedance 2.0 Pro route: Multi-modal composition combining subject images, scene reference videos, and voice reference audio in a single call for brand-consistent narrative clips
  • Automatic intent classification: Determines whether the user wants general animation, custom-voiceover lip-sync, or multi-modal composition and routes accordingly
  • Optimized prompting guidance: Each route includes model-specific tips such as leading with motion verbs, using preservation goals, and numbering multi-modal references
  • Flexible output control: Supports various aspect ratios, resolutions (up to 1080p), durations (up to 15s), and seed-locked reproducibility

Who is this for?

  • Content creators and social media teams who need to quickly animate product photos, portraits, or brand assets into short video clips
  • Marketing professionals producing lip-synced spokesperson videos or multi-language ad variants from a single image and voiceover
  • Creative teams building multi-modal narratives that combine character images, scene references, and voice tones into cohesive clips
๐Ÿ“ฆ

Same repository

prime-skills/runcomfy-agent-skills(25 items)

image-to-video

SKILL.md

406,100Installs
-
AddedAug 11, 2026

More from this repository10

๐ŸŽฏ
video-edit๐ŸŽฏSkill

A smart intent-routing skill for video editing on RunComfy that automatically selects the best model (Wan 2.7 Edit-Video, Kling 2.6 Pro Motion Control, or Lucy Edit Restyle) based on what the user wants to do, from general restyling and background swaps to precise motion transfer and outfit changes.

๐ŸŽฏ
nano-banana-2๐ŸŽฏSkill

A skill for generating images with Google Nano Banana 2, the flash-tier text-to-image model in the Gemini family on RunComfy, optimized for rapid iteration, batch ideation, in-image typography rendering, and optional web-grounded context with support for resolutions from 0.5K to 4K.

๐ŸŽฏ
image-edit๐ŸŽฏSkill

A smart intent-routing skill for image editing on RunComfy that selects the optimal model from Nano Banana Edit (batch up to 20 images), GPT Image 2 Edit (multilingual text rewrite, multi-ref composition), Flux Kontext Pro (single-shot precise local edit), or Z-Image Turbo Inpaint (mask-driven region edit) based on user intent.

๐ŸŽฏ
flux-kontext๐ŸŽฏSkill

A Claude Code skill that enables precise single-image editing using Black Forest Labs's Flux 1 Kontext Pro model through the RunComfy CLI, with built-in prompting patterns for high-fidelity local edits.

๐ŸŽฏ
wan-2-7๐ŸŽฏSkill

A skill for generating text-to-video with Wan 2.7, Wan-AI's flagship motion model on RunComfy, featuring multi-reference conditioning, audio-driven lip-sync via custom audio tracks, smooth transition physics, prompt expansion control, and up to 1080p output at 15 seconds duration.

๐ŸŽฏ
happyhorse-1-0๐ŸŽฏSkill

A Claude Code skill for generating text-to-video clips with HappyHorse 1.0 via the RunComfy CLI. HappyHorse 1.0 is currently ranked #1 on the Artificial Analysis Video Arena, producing native 1080p video with in-pass synchronized audio and multi-shot character consistency.

๐ŸŽฏ
kling-3-0๐ŸŽฏSkill

A Claude Code skill for generating video with Kling 3.0 via the RunComfy CLI. Covers all six Kling 3.0 endpoints across three rendering tiers (Standard, Pro, 4K) and two modes (text-to-video, image-to-video), with native synchronized audio and multi-shot character consistency.

๐ŸŽฏ
codex-pet๐ŸŽฏSkill

A Claude Code skill that generates custom OpenAI Codex Pets from a single reference image via the RunComfy CLI. It produces a Codex-compatible spritesheet (1536x1872, 8 columns x 9 rows) and pet.json manifest using GPT Image 2 and ImageMagick, without requiring Codex Pro or an OPENAI_API_KEY.

๐ŸŽฏ
ai-video-generation๐ŸŽฏSkill

A Claude Code skill that acts as a smart router for AI video generation across the full RunComfy model catalog. Supports text-to-video, image-to-video, and video-extend modes with models including HappyHorse 1.0, Kling 3.0, Wan 2.7, Seedance, Veo 3.1, and more.

๐ŸŽฏ
ai-image-generation๐ŸŽฏSkill

A Claude Code skill that acts as a smart router for AI image generation and editing across 11+ models on the RunComfy platform. Supports text-to-image and image-to-image modes with models including FLUX 2, GPT Image 2, Seedream 5, Nano Banana 2, and more.