Create videos where a person appears to speak or sing using only a photo and an audio file. This skill picks the best AI model for your project, whether you need a virtual presenter, a dubbed product demo, or a lip-synced character. No technical background required.
Behind the scenes the skill uses the RunComfy CLI to call models like ByteDance OmniHuman, Wan 2-7, HappyHorse, and Seedance v2. Each model handles different types of input. The skill reads your intent and chooses the right one automatically.
You can start with a written script or an existing audio recording. The result is a video with natural mouth movements and gestures. This is useful for UGC content, virtual spokespersons, or any project that needs a face to match a voice.
Global
mkdir -p ~/.claude/skills/ai-avatar-videoProject
mkdir -p .claude/skills/ai-avatar-videoSource Repository
Remotion Best Practicesremotion-dev/skills
Best practices for building videos with React and Remotion
Remotion Renderhalt-catch-fire/skills
Render React code to MP4 videos with Remotion and the belt CLI
Ai Video Generationhalt-catch-fire/skills
Create AI videos from text or images with 40+ models via inference.sh CLI
Ai Image Generationhalt-catch-fire/skills
Create AI images from text with 50+ models via simple CLI commands
Image To Videoagentspace-so/runcomfy-agent-skills
Pick the right AI model to animate your image into video
Nano Banana 2agentspace-so/runcomfy-agent-skills
Create rapid image drafts with strong typography using Google Nano Banana 2 on RunComfy
Nano Banana Editagentspace-so/runcomfy-agent-skills
Edit images with AI preserve subjects swap backgrounds batch up to 20
Flux Kontextagentspace-so/runcomfy-agent-skills
Flux Kontext Pro edits one image part while keeping the rest identical