wan-2-7
๐ฏSkillfrom prime-skills/runcomfy-agent-skills
Installation
npx vibeindex add prime-skills/runcomfy-agent-skills --skill wan-2-7npx skills add prime-skills/runcomfy-agent-skills --skill wan-2-7~/.claude/skills/wan-2-7/SKILL.mdA skill for generating text-to-video with Wan 2.7, Wan-AI's flagship motion model on RunComfy, featuring multi-reference conditioning, audio-driven lip-sync via custom audio tracks, smooth transition physics, prompt expansion control, and up to 1080p output at 15 seconds duration.
Overview
Wan 2.7 is a skill for generating video from text prompts using Wan-AI's flagship motion model, hosted on the RunComfy Model API. It excels at multi-reference conditioning with up to 5 reference media (images, video, and voice combined), audio-driven lip-sync where custom WAV/MP3 tracks (3-30 seconds) drive character lip movements, and smooth physics-aware motion transitions. The skill supports up to 1080p resolution, 15-second clips, 5 aspect ratios, automatic prompt expansion for short prompts, and negative prompting for targeted issue exclusion.
Key Features
- Audio-driven lip-sync: Supply a custom WAV/MP3 voiceover track (3-30s, up to 15MB) via
audio_urlto generate talking-head videos with synchronized lip movements, enabling multi-language dub variants from the same prompt - Multi-reference conditioning: Combine up to 5 reference media (images, video clips, and voice) for fine-grained motion control and character consistency
- Prompt expansion control: Automatic prompt rewriting is on by default to enhance short prompts; disable with
enable_prompt_expansion: falsefor literal, brand-strict control - Negative prompting: Target specific issues to avoid with concrete exclusions like "no subtitles, no watermark, no flicker" for cleaner output
- Physics-aware motion: Strong motion priors for smooth transitions, natural camera movements (dolly, orbit, tilt, handheld follow), and realistic physics
- Flexible output specs: 5 aspect ratios (16:9, 9:16, 1:1, 4:3, 3:4), 720p or 1080p resolution, 2-15 second duration, and seed-locked reproducibility
Who is this for?
- Marketing and ad teams producing lip-synced spokesperson videos with custom voiceovers, especially those needing multi-language variants from the same visual prompt
- Video content creators who need physics-aware, cinema-quality motion with precise camera control for product showcases, brand narratives, or social media shorts
- Developers building automated video generation pipelines on RunComfy who need fine-grained control over motion, audio sync, and prompt behavior
Same repository
prime-skills/runcomfy-agent-skills(25 items)
SKILL.md
More from this repository10
A smart intent-routing skill for video editing on RunComfy that automatically selects the best model (Wan 2.7 Edit-Video, Kling 2.6 Pro Motion Control, or Lucy Edit Restyle) based on what the user wants to do, from general restyling and background swaps to precise motion transfer and outfit changes.
A smart intent-routing skill that animates still images on RunComfy by selecting the optimal model: HappyHorse 1.0 I2V for general portrait and product animation with native audio, Wan 2.7 for custom-voiceover lip-sync, or Seedance 2.0 Pro for multi-modal composition with image, video, and audio references.
A skill for generating images with Google Nano Banana 2, the flash-tier text-to-image model in the Gemini family on RunComfy, optimized for rapid iteration, batch ideation, in-image typography rendering, and optional web-grounded context with support for resolutions from 0.5K to 4K.
A smart intent-routing skill for image editing on RunComfy that selects the optimal model from Nano Banana Edit (batch up to 20 images), GPT Image 2 Edit (multilingual text rewrite, multi-ref composition), Flux Kontext Pro (single-shot precise local edit), or Z-Image Turbo Inpaint (mask-driven region edit) based on user intent.
A Claude Code skill that enables precise single-image editing using Black Forest Labs's Flux 1 Kontext Pro model through the RunComfy CLI, with built-in prompting patterns for high-fidelity local edits.
A Claude Code skill for generating text-to-video clips with HappyHorse 1.0 via the RunComfy CLI. HappyHorse 1.0 is currently ranked #1 on the Artificial Analysis Video Arena, producing native 1080p video with in-pass synchronized audio and multi-shot character consistency.
A Claude Code skill for generating video with Kling 3.0 via the RunComfy CLI. Covers all six Kling 3.0 endpoints across three rendering tiers (Standard, Pro, 4K) and two modes (text-to-video, image-to-video), with native synchronized audio and multi-shot character consistency.
A Claude Code skill that generates custom OpenAI Codex Pets from a single reference image via the RunComfy CLI. It produces a Codex-compatible spritesheet (1536x1872, 8 columns x 9 rows) and pet.json manifest using GPT Image 2 and ImageMagick, without requiring Codex Pro or an OPENAI_API_KEY.
A Claude Code skill that acts as a smart router for AI video generation across the full RunComfy model catalog. Supports text-to-video, image-to-video, and video-extend modes with models including HappyHorse 1.0, Kling 3.0, Wan 2.7, Seedance, Veo 3.1, and more.
A Claude Code skill that acts as a smart router for AI image generation and editing across 11+ models on the RunComfy platform. Supports text-to-image and image-to-image modes with models including FLUX 2, GPT Image 2, Seedream 5, Nano Banana 2, and more.