controlnet-pose
🎯Skillfrom agentspace-so/runcomfy-agent-skills
A Claude Code skill for pose-conditioned image and video generation via the RunComfy CLI, routing across Kling motion control, Wan 2-2 Animate, and Z-Image ControlNet endpoints for pose, skeleton, and motion reference-driven generation.
Overview
ControlNet Pose is a Claude Code skill for pose-conditioned image and video generation via the RunComfy CLI. It routes across Kling 2-6 Motion Control Pro/Standard for transferring a reference video's motion onto a target character, community Wan 2-2 Animate for audio-driven character animation with pose conditioning, and Z-Image Turbo ControlNet LoRA for pose-conditioned still image generation from OpenPose, DWPose, canny, or depth control images. The skill picks the right route based on video vs still output and stylized vs photoreal intent.
Key Features
- Video motion/pose transfer - Kling 2-6 Motion Control Pro (default for video) takes a reference performance video plus a target character image and produces video of the target performing the reference motion, ideal for dance choreography re-shots and sports motion transfer.
- Pose-conditioned image generation - Z-Image Turbo ControlNet LoRA generates still images conditioned on OpenPose, DWPose, canny edge, or depth control images, enabling precise pose control in image creation.
- Audio-driven pose animation - Community Wan 2-2 Animate combines pose conditioning with audio-driven animation for character performances that sync body movement to audio tracks.
- Quality/cost tiers - Choose between Kling Motion Control Pro (premium quality) and Standard (cheaper) for video, with the skill guiding selection based on intended use.
Who is this for?
- Animators and motion designers who need to transfer choreography or body motion from reference footage onto new characters
- Game developers and concept artists creating pose-specific character renders from skeleton or depth references
- Video producers building character animation pipelines who need CLI-level control over pose-conditioned generation across multiple model endpoints
Same repository
agentspace-so/runcomfy-agent-skills(30 items)
Installation
npx vibeindex add agentspace-so/runcomfy-agent-skills --skill controlnet-posenpx skills add agentspace-so/runcomfy-agent-skills --skill controlnet-pose~/.claude/skills/controlnet-pose/SKILL.mdSKILL.md
More from this repository10
A smart intent-routing skill for video editing on RunComfy that selects the best model based on the user's intent. Routes to Wan 2.7 Edit-Video for restyle and background swaps, Kling 2.6 Pro for precise motion transfer, or Lucy Edit for lightweight identity-stable restyle and outfit swaps.
A smart intent-routing skill for image-to-video generation on RunComfy that automatically selects the best model for the task. Routes to HappyHorse 1.0 I2V for general animations, Wan 2.7 for custom-voiceover lip-sync, or Seedance 2.0 Pro for multi-modal composition from image, video, and audio references.
A smart intent-routing skill for image editing on RunComfy that selects the best model based on the editing task. Routes to Nano Banana Edit for batch edits up to 20 images, GPT Image 2 for multilingual text rewrite, Flux Kontext Pro for single-shot precise edits, or Z-Image Turbo for mask-driven inpainting.
Edit images with Black Forest Labs' Flux 1 Kontext Pro on RunComfy, specializing in single-reference precise local edits with high-fidelity source preservation. Ideal for targeted changes like adding objects or modifying details while keeping the rest of the image unchanged.
A RunComfy skill that generates images using Google Nano Banana 2, the flash-tier text-to-image model in the Gemini family. Optimized for rapid iteration, social thumbnails, and in-image typography with configurable resolution tiers and safety tolerance.
Edit images with Google Nano Banana 2 on RunComfy, supporting batch edits of up to 20 images per call with strong identity preservation. Features localized edits using spatial language, background swaps, and configurable resolution up to 4K.
Generate text-to-video with HappyHorse 1.0 on RunComfy, currently ranked #1 on Artificial Analysis Video Arena. Supports native 1080p with in-pass synchronized audio, multi-shot character consistency, and 6-language prompt support via the RunComfy CLI.
Generate text-to-video with Wan-AI's Wan 2.7 on RunComfy, featuring multi-reference conditioning and audio-driven lip-sync via custom audio tracks. Supports prompt expansion, negative prompts, and up to 1080p resolution through the RunComfy CLI.
Generate cinematic short-form video with ByteDance Seedance 2.0 Pro on RunComfy, supporting multi-modal references including up to 9 images, 3 videos, and 3 audio tracks. Features native lip-synced audio generation and is ideal for brand-consistent multi-language narratives.
Edit images with OpenAI GPT Image 2 on RunComfy, excelling at multilingual in-image text editing across any script (Latin, kana, CJK, Cyrillic, Arabic) and multi-reference composition with up to 10 input images. Ideal for identity-preserving edits and layout-precise repositioning.