Loading...
Loading...
在RunComfy上将任意静态图片动起来——该技能是一个智能路由工具,可将用户意图与RunComfy模型库中合适的i2v模型进行匹配。针对通用动画场景选择HappyHorse 1.0 I2V(竞技场排名第1,支持原生音频、身份保留);针对自定义旁白唇形同步场景选择带`audio_url`参数的Wan 2.7;针对图像+参考视频+参考音频的多模态动画场景选择Seedance 2.0 Pro。该技能整合了每个模型的文档提示模板,让调用者无需在错误模型上反复尝试即可获得更优质的输出。通过本地RunComfy CLI调用`runcomfy run <vendor>/<model>/image-to-video`(或其端点变体)。当触发词为"image to video"、"image-to-video"、"i2v"、"animate image"、"make this move"或任何明确要求将静态图转为视频的指令时,该技能启动。
npx skill4agent add skills-collective/skills image-to-videonpx skills add agentspace-so/runcomfy-skills --skill image-to-video -g| 用户意图 | 模型 | 选择理由 |
|---|---|---|
| 制作人像动画——保持身份特征稳定 | HappyHorse 1.0 I2V | Artificial Analysis 竞技场排名第1(Elo 1392);面部保真度出色 |
| 产品展示/360度旋转/微距运动 | HappyHorse 1.0 I2V | 几何形态保留+流畅镜头移动 |
| 一站式生成带同步环境音的动画 | HappyHorse 1.0 I2V | 支持同步音频合成 |
| 制作动画并与自定义旁白音轨做唇形同步 | Wan 2.7 + | 支持上传自有MP3/WAV文件(时长3–30秒,≤15MB)并驱动唇形同步 |
| 多语言配音变体(同一图片,每次调用更换音频) | Wan 2.7 + | 同一画面,更换 |
| 多模态合成——图像+参考视频+参考音频结合 | Seedance 2.0 Pro | 支持最多9张参考图、3个参考视频(每个2–15秒)、3个参考音频 |
| 品牌一致性叙事——结合角色参考+场景参考+声音参考 | Seedance 2.0 Pro | 图像保留身份特征,视频保留场景风格,音频保留声音特质 |
| 用户未明确指定时的默认选择 | HappyHorse 1.0 I2V | 综合质量最佳+支持原生音频 |
npm i -g @runcomfy/cliruncomfy loginRUNCOMFY_TOKEN=<token>happyhorse/happyhorse-1-0/image-to-video| 字段 | 类型 | 是否必填 | 默认值 | 说明 |
|---|---|---|---|---|
| string | 是 | — | JPEG/JPG/PNG/WEBP格式。最小300px。宽高比1:2.5–2.5:1。≤10MB。 |
| string | 是 | — | 非CJK字符≤5000个,CJK字符≤2500个。需描述运动/镜头/灯光效果。 |
| 枚举值 | 否 | | 可选 |
| 整数 | 否 | 5 | 时长3–15秒。 |
| 整数 | 否 | 0 | 重复使用可对比不同变体效果。 |
| 布尔值 | 否 | true | 控制是否添加服务商水印。 |
runcomfy run happyhorse/happyhorse-1-0/image-to-video \
--input '{
"image_url": "https://.../portrait.jpg",
"prompt": "Gentle camera drift around the subject'\''s face, subtle breathing motion, identity-stable features, soft natural light."
}' \
--output-dir <absolute/path>audio_urlwan-ai/wan-2-7/text-to-video/image-to-videoaudio_url| 字段 | 类型 | 是否必填 | 默认值 | 说明 |
|---|---|---|---|---|
| string | 是 | — | 最多约5000字符。描述说话人镜头:构图、灯光、动作。 |
| string | 是(用于唇形同步) | — | WAV/MP3格式,时长3–30秒,≤15MB。驱动唇形同步。 |
| 枚举值 | 否 | | 可选 |
| 枚举值 | 否 | | 可选 |
| 枚举值 | 否 | | 时长2–15秒(整秒)。需与音频长度匹配。 |
| string | 否 | — | 明确要避免的问题(例如"no subtitles, no flicker")。 |
| 整数 | 否 | — | 用于结果重现。 |
runcomfy run wan-ai/wan-2-7/text-to-video \
--input '{
"prompt": "Medium close-up of a confident spokesperson in a softly-lit recording booth, leaning slightly toward the camera, locked tripod, shallow DOF, warm key light from camera-left.",
"audio_url": "https://.../voiceover-en.mp3",
"duration": 12,
"aspect_ratio": "9:16"
}' \
--output-dir <absolute/path>durationnegative_prompt"no subtitles, no flicker, no distorted hands"audio_urlbytedance/seedance-v2/pro| 字段 | 类型 | 是否必填 | 默认值 | 说明 |
|---|---|---|---|---|
| string | 是 | — | 中文≤500字符 或 英文≤1000词。 |
| 数组 | 是(用于图像转视频) | | 0–9张图片。第一张为主体图像。 |
| 数组 | 否 | | 0–3个参考片段(MP4/MOV格式),每个时长2–15秒。 |
| 数组 | 否 | | 0–3个参考音频(WAV/MP3格式),时长2–15秒,单个≤15MB。 |
| 枚举值 | 否 | | 可选 |
| 整数 | 否 | 5 | 时长4–15秒(整秒)。 |
| 枚举值 | 否 | | 可选 |
| 布尔值 | 否 | true | 同步生成语音/音效/音乐。 |
| 整数 | 否 | — | 用于结果重现。 |
runcomfy run bytedance/seedance-v2/pro \
--input '{
"prompt": "Subject from image 1 walks through the café in video 1, voice tone matches audio 1. Medium close-up, slow push-in, warm light, gentle ambience.",
"image_url": ["https://.../subject.jpg"],
"video_url": ["https://.../cafe-locked-shot.mp4"],
"audio_url": ["https://.../voice-tone.mp3"],
"duration": 8
}' \
--output-dir <absolute/path>image_urlprompt"subject from image 1, lighting from video 1, voice from audio 1"wan-2-7seedance-v2| 代码 | 含义 |
|---|---|
| 0 | 成功 |
| 64 | CLI参数错误 |
| 65 | 输入JSON错误/schema不匹配 |
| 69 | 上游服务5xx错误 |
| 75 | 可重试:超时/429限流 |
| 77 | 未登录或令牌被拒绝 |
runcomfy run <model_id>.runcomfy.net.runcomfy.com--output-dirCtrl-Cruncomfy login~/.config/runcomfy/token.jsonRUNCOMFY_TOKEN--inputmodel-api.runcomfy.net*.runcomfy.net*.runcomfy.com