skill
Image To Video
نبذة
# Image-to-Video — Pro Pack on RunComfy
[runcomfy.com](https://www.runcomfy.com/?utm_source=skills.sh&utm_medium=skill&utm_campaign=image-to-video) · [HappyHorse I2V](https://www.runcomfy.com/models/happyhorse/happyhorse-1-0/image-to-video?utm_source=skills.sh&utm_medium=skill&utm_campaign=image-to-video) · [Wan 2.7](https://www.runcomfy.com/models/wan-ai/wan-2-7/text-to-video?utm_source=skills.sh&utm_medium=skill&utm_campaign=image-to-video) · [Seedance 2.0 Pro](https://www.runcomfy.com/models/bytedance/seedance-v2/pro?utm_source=skills.sh&utm_medium=skill&utm_campaign=image-to-video) · [GitHub](https://github.com/agentspace-so/runcomfy-skills/tree/main/image-to-video)
**Image-to-video, intent-routed.** This skill doesn't lock you to one model — it picks the right i2v model in the RunComfy catalog based on what the user actually wants: portrait animation, custom-voiceover lip-sync, or multi-modal composition.
```bash npx skills add agentspace-so/runcomfy-skills --skill image-to-video -g ```
## Pick the right model for the user's intent
| User intent | Model | Why | |---|---|---| | Animate a portrait — keep identity stable | **HappyHorse 1.0 I2V** | #1 on Artificial Analysis Arena (Elo 1392); strong facial fidelity | | Product reveal / 360 / macro motion | **HappyHorse 1.0 I2V** | Geometry preservation + smooth camera moves | | Native synchronized ambient audio in one pass | **HappyHorse 1.0 I2V** | In-pass audio synthesis | | Animate **and** lip-sync to a **custom voiceover track** | **Wan 2.7 + `audio_url`** | Accepts your own MP3/WAV (3–30s, ≤15MB) and drives lip-sync to it | | Multi-language dub variants (same image, different audio per call) | **Wan 2.7 + `audio_url`** | Same shot, swap `audio_url` per language | | Multi-modal — image + reference video + reference audio together | **Seedance 2.0 Pro** | Up to 9 image refs, 3 video refs (2–15s each), 3 audio refs | | Brand-consistent narrative with character ref + scene ref + voice ref | **Seedance 2.0 Pro** | Image holds identity, video holds scene, audio holds voice | | Default if unspecified | **HappyHorse 1.0 I2V** | Best all-round quality + native audio |
The agent reads this table, classifies the user's intent, and picks the matching subsection below.
## Prerequisites
1. **RunComfy CLI** — `npm i -g @runcomfy/cli` 2. **RunComfy account** — `runcomfy login` opens a browser device-code flow. 3. **CI / containers** — set `RUNCOMFY_TOKEN=<token>`. 4. **A source image URL** — JPEG/PNG/WebP, min 300px, ≤10MB; aspect 1:2.5 to 2.5:1 (HappyHorse) — other models have similar specs.
---
## Route 1: HappyHorse 1.0 I2V — default for portrait / product / general animation
**Model**: `happyhorse/happyhorse-1-0/image-to-video` · **Arena rank**: #1 (Elo 1392)
### Schema
| Field | Type | Required | Default | Notes | |---|---|---|---|---| | `image_url` | string | yes | — | JPEG/JPG/PNG/WEBP. Min 300px. Aspect 1:2.5–2.5:1. ≤10MB. | | `prompt` | string | yes | — | ≤5000 non-CJK or 2500 CJK chars. **Motion / camera / lighting** description. | | `resolution` | enum | no | `1080P` | `720P` or `1080P`. | | `duration` | int | no | 5 | 3–15 seconds. | | `seed` | int | no | 0 | Reuse for variant comparisons. | | `watermark` | bool | no | true | Provider watermark toggle. |
Output aspect = input aspect. No independent reframing.
### Invoke
```bash runcomfy run happyhorse/happyhorse-1-0/image-to-video \ --input '{ "image_url": "https://.../portrait.jpg", "prompt": "Gentle camera drift around the subject'\''s face, subtle breathing motion, identity-stable features, soft natural light." }' \ --output-dir <absolute/path> ```
### Prompting tips
- **Lead with motion verbs**: "drift", "dolly in", "orbit", "tilt up", "reveal", "blink", "breathe". Front-load what's MOVING. - **Don't restate the image** — the model sees it. Focus tokens on what changes. - **Preservation goals explicit**: "identity-stable features", "packaging unchanged", "background geometry stable". - **Lighting evolution**: "rim light intensifying", "shadows shortening as camera rises". - **One beat per clip** — single primary motion (orbit OR dolly OR tilt OR character action).
---
## Route 2: Wan 2.7 + `audio_url` — when the user has a custom voiceover
**Model**: `wan-ai/wan-2-7/text-to-video` (NOT `/image-to-video` — Wan 2.7's t2v endpoint accepts an `audio_url` that drives lip-sync)
**Note on i2v with Wan 2.7**: Wan 2.7's primary i2v animation isn't on a dedicated endpoint here. For pure i2v (image animated by motion prompt only), prefer **HappyHorse i2v**. Use Wan 2.7 specifically when the user has a custom audio track they want lip-synced to a generated talking-head clip.
### Schema (Wan 2.7 t2v with audio)
| Field | Type | Required | Default | Notes | |---|---|---|---|---| | `prompt` | string | yes | — | Up to ~5000 chars. Describe the talking-head shot: framing, lighting, motion. | | `audio_url` | string | yes (for lip-sync) | — | WAV/MP3, 3–30s, ≤15MB. **Drives lip-sync
التثبيت
شغل هذا الأمر
npx skills add genmedia-labs/skillsيعمل مع
خطوات التثبيت
Install with `npx skills add genmedia-labs/skills`, or clone the repository and copy the `image-to-video` folder into your Claude skills directory.
أسئلة شائعة
كيف أثبت Image To Video؟
شغل هذا الأمر في الطرفية:
npx skills add genmedia-labs/skillsمع أي أدوات ذكاء اصطناعي تعمل Image To Video؟
تعمل مع claude_app، claude_code، claude_api، cursor، codex، windsurf، cline، zed.
من طور Image To Video؟
طورها genmedia-labs، وتصدر بترخيص MIT.
هل Image To Video مجانية؟
نعم، يمكنك استخدامها مجانا وفق ترخيص MIT.
npx skills add vercel-labs/skills
npx skills add mattpocock/skills
npx skills add mattpocock/skills
npx skills add mattpocock/skills
npx skills add mattpocock/skills
npx skills add genmedia-labs/skills
افحص قبل التثبيت
شغل أي مصدر عبر فحوصاتنا - الظهور في الذكاء الاصطناعي والأمان والأداء واكتشاف التقنيات.
فحص أمني تلقائي للموقع
الأمان
محلل سرعة الصفحة
الأداء
اختبار جودة المحتوى العربي بالذكاء الاصطناعي
جودة المحتوى
مختبر وكلاء الذكاء الاصطناعي
اختبار الذكاء الاصطناعي
كاشف منصة الموقع
الترحيل
تدقيق الظهور في محركات الذكاء الاصطناعي
الظهور في الذكاء الاصطناعي
مولد ملف llms.txt
الظهور في الذكاء الاصطناعي
مقياس سهولة القراءة بالعربية
جودة المحتوى
منشئ البيانات المنظمة
الظهور في الذكاء الاصطناعي
حاسبة تكاليف الذكاء الاصطناعي
اختبار الذكاء الاصطناعي
محلل العناوين العربية
جودة المحتوى