skill
AI Avatar Video
نبذة
# AI Avatar & Talking Head Video
Put words in a face. This skill routes across RunComfy's audio-driven avatar models — OmniHuman, Wan 2-7 with audio_url, HappyHorse, Seedance v2 — picking the right path for the user's intent and shipping the documented prompts + the exact `runcomfy run` invoke for each.
[runcomfy.com](https://www.runcomfy.com/?utm_source=skills.sh&utm_medium=skill&utm_campaign=ai-avatar-video) · [Lip-sync feature](https://www.runcomfy.com/models/feature/lip-sync?utm_source=skills.sh&utm_medium=skill&utm_campaign=ai-avatar-video) · [CLI docs](https://docs.runcomfy.com/cli/introduction?utm_source=skills.sh&utm_medium=skill&utm_campaign=ai-avatar-video)
## Powered by the RunComfy CLI
```bash # 1. Install (see runcomfy-cli skill for details) npm i -g @runcomfy/cli # or: npx -y @runcomfy/cli --version
# 2. Sign in runcomfy login # or in CI: export RUNCOMFY_TOKEN=<token>
# 3. Generate an avatar video runcomfy run <vendor>/<model>/<endpoint> \ --input '{"prompt": "...", "audio_url": "https://...", "image_url": "https://..."}' \ --output-dir ./out ```
CLI deep dive: [`runcomfy-cli`](https://www.skills.sh/agentspace-so/runcomfy-agent-skills/runcomfy-cli) skill.
## Install this skill
```bash npx skills add agentspace-so/runcomfy-agent-skills --skill ai-avatar-video -g ```
---
## Pick the right model for the user's intent
Listed newest first. The agent classifies user intent — pre-recorded audio file or just a script? Photoreal portrait or stylized character? Single shot or cinematic composition? — and picks one route below.
**OmniHuman** — `bytedance/omnihuman/api` *(default)* > ByteDance audio-driven full-body avatar. Feed one portrait + one audio file, get back a video where the subject speaks / sings / gestures naturally. Listed on RunComfy's `/feature/lip-sync` as the curated default. > Pick for: UGC voiceover, virtual presenter, dubbed product demo, multi-language clips from same portrait. > Avoid for: no audio file available (need to generate speech from a script) — use **HappyHorse 1.0**.
**HappyHorse 1.0** — `happyhorse/happyhorse-1-0/text-to-video` (t2v) · `happyhorse/happyhorse-1-0/image-to-video` (i2v) > Arena #1 t2v / i2v with in-pass audio generated from prompt. No external audio file required — quote the spoken line inside the prompt. > Pick for: written script with no audio file, "write a script → get a video", concept clips, i2v talking-head from an existing portrait. > Avoid for: precise lip-sync to a specific MP3 — audio is regenerated each call, not locked.
**Seedance v2 Pro** — `bytedance/seedance-v2/pro` > ByteDance multi-modal flagship — up to 9 reference images, 3 reference videos, 3 reference audio tracks composed in one pass with cinematic motion / lens / lighting control. > Pick for: cinematic monologue with reference subject + reference audio + reference scene; ad creative. > Avoid for: simple "portrait + audio" jobs — overpowered, slower. Use **OmniHuman**.
**Wan 2-7 with `audio_url`** — `wan-ai/wan-2-7/text-to-video` > Open-weights with `audio_url` field — prompt describes the scene, audio file drives the mouth. > Pick for: full scene control (not just a portrait), specific voiceover MP3, open-weights pipeline. > Avoid for: simplest portrait-talks job — use **OmniHuman**.
**Wan 2-2 Animate** — `community/wan-2-2-animate/api` > Community-published variant on the Wan 2-2 base. Audio-driven full-body animation of stylized characters (illustration, anime, mascot). > Pick for: stylized / illustrated character + audio (not a photoreal portrait). > Avoid for: photoreal subjects — use **OmniHuman** or **Wan 2-7**.
---
## Route 1: OmniHuman — default audio-driven avatar
**Model**: `bytedance/omnihuman/api` **Catalog**: [omnihuman](https://www.runcomfy.com/models/bytedance/omnihuman/api?utm_source=skills.sh&utm_medium=skill&utm_campaign=ai-avatar-video) · [`/feature/lip-sync`](https://www.runcomfy.com/models/feature/lip-sync?utm_source=skills.sh&utm_medium=skill&utm_campaign=ai-avatar-video)
ByteDance OmniHuman is the strongest single-shot path: feed it **one portrait image + one audio file**, get back a video where the subject speaks / sings / gestures naturally to the audio. No prompt required beyond the inputs.
### Invoke
```bash runcomfy run bytedance/omnihuman/api \ --input '{ "image_url": "https://your-cdn.example/presenter.jpg", "audio_url": "https://your-cdn.example/voiceover.mp3" }' \ --output-dir ./out ```
### Tips
- **Portrait framing works best** — head-and-shoulders or upper body. Full-body still works but expects more "presenter" energy. - **Audio quality drives output quality** — clean voiceover (no music bed) → cleaner mouth sync. If your audio is a mix, isolate the voice stem first. - **No prompt field** — the model derives everything from image + audio. Don't fight that. - See the full input schema on the [model page](https://www.runcomfy.com/models/bytedance/omnihuman/api?utm_source=skills.sh&utm_medi
التثبيت
شغل هذا الأمر
npx skills add prime-skills/runcomfy-agent-skillsيعمل مع
خطوات التثبيت
Install with `npx skills add prime-skills/runcomfy-agent-skills`, or clone the repository and copy the `ai-avatar-video` folder into your Claude skills directory.
أسئلة شائعة
كيف أثبت AI Avatar Video؟
شغل هذا الأمر في الطرفية:
npx skills add prime-skills/runcomfy-agent-skillsمع أي أدوات ذكاء اصطناعي تعمل AI Avatar Video؟
تعمل مع claude_app، claude_code، claude_api، cursor، codex، windsurf، cline، zed.
من طور AI Avatar Video؟
طورها prime-skills، وتصدر بترخيص MIT.
هل AI Avatar Video مجانية؟
نعم، يمكنك استخدامها مجانا وفق ترخيص MIT.
npx skills add vercel-labs/skills
npx skills add mattpocock/skills
npx skills add mattpocock/skills
npx skills add mattpocock/skills
npx skills add mattpocock/skills
npx skills add genmedia-labs/skills
افحص قبل التثبيت
شغل أي مصدر عبر فحوصاتنا - الظهور في الذكاء الاصطناعي والأمان والأداء واكتشاف التقنيات.
فحص أمني تلقائي للموقع
الأمان
محلل سرعة الصفحة
الأداء
اختبار جودة المحتوى العربي بالذكاء الاصطناعي
جودة المحتوى
مختبر وكلاء الذكاء الاصطناعي
اختبار الذكاء الاصطناعي
كاشف منصة الموقع
الترحيل
تدقيق الظهور في محركات الذكاء الاصطناعي
الظهور في الذكاء الاصطناعي
مولد ملف llms.txt
الظهور في الذكاء الاصطناعي
مقياس سهولة القراءة بالعربية
جودة المحتوى
منشئ البيانات المنظمة
الظهور في الذكاء الاصطناعي
حاسبة تكاليف الذكاء الاصطناعي
اختبار الذكاء الاصطناعي
محلل العناوين العربية
جودة المحتوى