Mohamed Nagy

Mohamed Nagy https://www.youtube.com/
Ai Specialist & Film Maker , دكتور جامعي بكلية الفنون التطبيقية

الورك فلو والتطبيق  في أول تعليق مجاناً بالكامل 🚀 Fastest MiniMax H3 Image to Video — 15 Secondsورك فلو يحوّل صورة واحدة...
12/08/2026

الورك فلو والتطبيق في أول تعليق مجاناً بالكامل
🚀 Fastest MiniMax H3 Image to Video — 15 Seconds

ورك فلو يحوّل صورة واحدة ثابتة إلى فيديو سينمائي مدته حتى 15 ثانية، مع توليد الحركة، الحوار، الأصوات والمؤثرات المتزامنة باستخدام MiniMax H3.

✅ يحتاج إلى صورة واحدة فقط.
✅ يدعم البرومبت التفصيلي والحوار العربي.
✅ يولّد الفيديو والصوت معًا.
✅ يحافظ على ملامح الشخصية والملابس وفقًا للبرومبت.
✅ يدعم نسب عرض ودقات متعددة.
✅ يصدر النتيجة بصيغة H.264 MP4 عند 24 FPS.

━━━━━━━━━━━━━━━━━━━━

⚙️ شرح الورك فلو بالعربية

1️⃣ رفع الصورة — Load Image

ارفع الصورة الأساسية داخل نود:

Image — الصورة

هذه الصورة تحدد الشخصية أو العنصر الذي سيظهر ويتحرك في الفيديو.

للحفاظ على الهوية، وضّح داخل البرومبت أن الصورة تتحكم في:

- ملامح الوجه.
- الشعر والملابس.
- العمر والبنية الجسدية.
- الألوان والعناصر المهمة.

وحدّد أيضًا العناصر التي لا تريد نقلها، مثل الخلفية أو وضعية الجسم أو الميكروفون.

━━━━━━━━━━━━━━━━━━━━

2️⃣ كتابة البرومبت — MiniMaxH3ImageToVideo

اكتب وصف الفيديو داخل نود:

MiniMaxH3ImageToVideo

يُفضّل تنظيم البرومبت بهذا الترتيب:

- وظيفة الصورة المرجعية.
- مواصفات الشخصية المطلوب تثبيتها.
- المكان والأجواء.
- تسلسل الأحداث زمنيًا.
- حركة الشخصية.
- حركة الكاميرا.
- الإضاءة والأسلوب البصري.
- الحوار والمؤثرات الصوتية.
- العناصر الممنوعة.

مثال لتعريف الصورة:

[REFERENCE USE]

Image1 defines the subject’s exact facial identity, hairstyle, clothing, body proportions, and visible appearance only.

Do not inherit the original background, pose, framing, microphone, or lighting from Image1.

إذا كنت لا تريد أصواتًا بشرية، أضف:

No dialogue, no narration, no singing, no speech, and no generated vocals. Generate only synchronized environmental ambience and physical sound effects.

━━━━━━━━━━━━━━━━━━━━

3️⃣ تحديد المدة — Duration

حدّد مدة الفيديو من نود:

Float — Duration

الإعداد الحالي:

15 Seconds

يحوّل الورك فلو المدة تلقائيًا إلى عدد فريمات مناسب للنموذج عند 24 FPS.

عند اختيار 15 ثانية:

15 × 24 = 360 Frames

ثم تعدّل معادلة ComfyMathExpression العدد تلقائيًا ليصبح:

362 Frames

وذلك ليتوافق مع البنية الزمنية المطلوبة داخل MiniMax H3.

━━━━━━━━━━━━━━━━━━━━

4️⃣ اختيار الحجم — ResolutionSelector

اختر نسبة العرض والدقة من:

ResolutionSelector — حجم الفيديو

الإعداد الافتراضي:

- Aspect Ratio: 16:9
- Megapixels: 0.3 MP
- Output Resolution: 736×416
- Size Multiple: 32

ابدأ الاختبار بدقة 0.2 أو 0.3 MP لتقليل وقت الرندر والتكلفة، ثم ارفع الدقة بعد التأكد من نجاح الحركة والبرومبت.

Megapixels| دقة 16:9
0.2 MP| 608×352
0.3 MP| 736×416
0.4 MP| 864×480
0.5 MP| 960×544
0.6 MP| 1056×608
0.7 MP| 1152×640
0.8 MP| 1216×672
0.9 MP| 1280×736
0.98 MP| 1344×768
1.0 MP| 1376×768
1.2 MP| 1504×832
1.5 MP| 1664×928
1.8 MP| 1824×1024
2.0 MP| 1920×1088

━━━━━━━━━━━━━━━━━━━━

🧠 كيف يعمل الورك فلو؟

1️⃣ تستقبل نود Load Image الصورة المرجعية.

2️⃣ تستقبل MiniMaxH3ImageToVideo الصورة والبرومبت والمدة والدقة.

3️⃣ يستخدم الورك فلو نموذج الفيديو:

minimax_h3_fl2va_bf16.safetensors

4️⃣ يستخدم نموذج النص والفهم البصري:

qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors

5️⃣ يستخدم مسارين منفصلين لفك الترميز:

- MiniMax H3 Video VAE FP16
- MiniMax H3 Audio VAE FP32

6️⃣ تطبق نود:

MiniMaxH3MemoryEfficientSageAttentionPatch

تقنية Attention موفرة للذاكرة.

7️⃣ تستخدم نود:

MiniMaxH3BlockCacheT8

الـBlock Cache لتقليل العمليات المتكررة والمساعدة في تسريع التوليد.

8️⃣ يستخدم الـSampler:

- Sampler: res_multistep
- Scheduler: simple
- Sampling Steps: 20
- Denoise: 1.0

9️⃣ يتم فك ترميز الفيديو داخل VAEDecode وفك الصوت داخل VAEDecodeAudio.

🔟 تجمع نود Video Combine الفيديو والصوت داخل ملف نهائي واحد.

إعدادات التصدير:

- Format: H.264 MP4
- Frame Rate: 24 FPS
- Pixel Format: YUV420P
- CRF: 19

━━━━━━━━━━━━━━━━━━━━

▶️ طريقة تشغيل الورك فلو

1️⃣ ارفع الصورة داخل Image — الصورة.
2️⃣ اكتب وصف الفيديو داخل MiniMaxH3ImageToVideo.
3️⃣ حدد المدة من Duration.
4️⃣ اختر نسبة العرض والدقة من ResolutionSelector.
5️⃣ ابدأ بدقة 0.2 أو 0.3 MP.
6️⃣ اضغط Queue Prompt.
7️⃣ انتظر حتى يكتمل توليد الفيديو والصوت.
8️⃣ شاهد النتيجة من Video Combine.
9️⃣ حمّل الفيديو النهائي بصيغة MP4.

━━━━━━━━━━━━━━━━━━━━

📌 ملاحظات مهمة

- هذا إصدار Image-to-Video ويحتاج إلى صورة واحدة.
- لا يحتاج إلى Last Frame أو ملف صوتي خارجي.
- الصورة تحدد المظهر، بينما البرومبت يحدد الحركة والمكان والأحداث والصوت.
- قيمة الـSeed مضبوطة على Randomize، لذلك قد تختلف النتيجة في كل تشغيل.
- اكتب الأحداث بترتيب زمني واضح.
- لا تضع أحداثًا أكثر مما تسمح به مدة 15 ثانية.
- وزّع الحوار على توقيت واقعي حتى يُنطق بصورة طبيعية.
- اذكر اللغة واللهجة المطلوبة بوضوح.
- هذه النسخة تستخدم 20 Sampling Steps وليست Turbo 4-Step.
- لا تعدّل منطقة «إعدادات ملناش دعوة بيها» إلا إذا كنت تعرف وظيفة النودز.
- رفع الدقة يزيد وقت الرندر واستهلاك الذاكرة والتكلفة.

━━━━━━━━━━━━━━━━━━━━

🚀 Fastest MiniMax H3 Image to Video — 15 Seconds

This workflow transforms one still image into a cinematic video of up to 15 seconds, generating motion, dialogue, environmental ambience, and synchronized sound effects with MiniMax H3.

✅ Requires only one input image.
✅ Supports detailed prompts and Arabic dialogue.
✅ Generates video and synchronized audio together.
✅ Preserves identity and clothing through prompt instructions.
✅ Supports multiple resolutions and aspect ratios.
✅ Exports an H.264 MP4 video at 24 FPS.

━━━━━━━━━━━━━━━━━━━━

⚙️ Workflow Explanation in English

1️⃣ Upload the Image — Load Image

Upload your reference image through:

Image — الصورة

This image defines the character or object that will appear and move in the generated video.

To preserve identity, specify that the image controls:

- Facial identity.
- Hairstyle and clothing.
- Apparent age and body proportions.
- Important colors and visible details.

Also exclude unwanted elements such as the original background, pose, microphone, lighting, or framing.

━━━━━━━━━━━━━━━━━━━━

2️⃣ Write the Prompt — MiniMaxH3ImageToVideo

Enter the complete video description inside:

MiniMaxH3ImageToVideo

A strong prompt should contain:

- The exact role of the reference image.
- Identity and continuity locks.
- Location and atmosphere.
- Chronological actions.
- Character movement.
- Camera movement.
- Lighting and visual style.
- Dialogue and synchronized sound.
- Specific negative constraints.

Example reference definition:

[REFERENCE USE]

Image1 defines the subject’s exact facial identity, hairstyle, clothing, body proportions, and visible appearance only.

Do not inherit the original background, pose, framing, microphone, or lighting from Image1.

If you do not want speech or generated vocals, add:

No dialogue, no narration, no singing, no speech, and no generated vocals. Generate only synchronized environmental ambience and physical sound effects.

━━━━━━━━━━━━━━━━━━━━

3️⃣ Set the Duration

Set the video duration through:

Float — Duration

Current setting:

15 Seconds

The workflow automatically converts the selected duration into a compatible frame count at 24 FPS.

For a 15-second video:

15 × 24 = 360 Frames

The ComfyMathExpression node adjusts it automatically to:

362 Frames

This produces a frame count compatible with MiniMax H3’s temporal structure.

━━━━━━━━━━━━━━━━━━━━

4️⃣ Select the Resolution

Select the aspect ratio and resolution through:

ResolutionSelector

Default setting:

- Aspect Ratio: 16:9
- Megapixels: 0.3 MP
- Output Resolution: 736×416
- Size Multiple: 32

Start your tests at 0.2 or 0.3 MP to reduce rendering time and cost. Increase the resolution only after confirming that the prompt and motion work correctly.

━━━━━━━━━━━━━━━━━━━━

🧠 How the Workflow Works

1️⃣ Load Image receives the reference image.

2️⃣ MiniMaxH3ImageToVideo receives the image, prompt, duration, width, and height.

3️⃣ The workflow loads the video generation model:

minimax_h3_fl2va_bf16.safetensors

4️⃣ It uses the following vision-language model for prompt and image understanding:

qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors

5️⃣ It uses separate decoding models:

- MiniMax H3 Video VAE FP16
- MiniMax H3 Audio VAE FP32

6️⃣ MiniMaxH3MemoryEfficientSageAttentionPatch applies memory-efficient attention.

7️⃣ MiniMaxH3BlockCacheT8 reduces repeated computation and helps accelerate generation.

8️⃣ The sampling configuration is:

- Sampler: res_multistep
- Scheduler: simple
- Sampling Steps: 20
- Denoise: 1.0

9️⃣ VAEDecode decodes the video, while VAEDecodeAudio decodes the synchronized audio.

🔟 Video Combine combines the decoded frames and audio into the final video.

Export settings:

- Format: H.264 MP4
- Frame Rate: 24 FPS
- Pixel Format: YUV420P
- CRF: 19

━━━━━━━━━━━━━━━━━━━━

▶️ How to Run the Workflow

1️⃣ Upload your reference through Image — الصورة.
2️⃣ Enter the video description inside MiniMaxH3ImageToVideo.
3️⃣ Set the duration through Duration.
4️⃣ Select the aspect ratio and resolution through ResolutionSelector.
5️⃣ Start testing at 0.2 or 0.3 MP.
6️⃣ Press Queue Prompt.
7️⃣ Wait for the video and audio generation to finish.
8️⃣ Preview the result inside Video Combine.
9️⃣ Download the completed video as an MP4 file.

━━━━━━━━━━━━━━━━━━━━

📌 Important Notes

- This is an Image-to-Video workflow requiring one image.
- It does not require a Last Frame or an external audio file.
- The image defines appearance; the prompt defines motion, environment, events, camera, and sound.
- The seed is set to Randomize, so each run can produce a different result.
- Describe all events in a clear chronological order.
- Avoid including too many actions within 15 seconds.
- Give every dialogue line enough time to be delivered naturally.
- Clearly specify the required language and accent.
- This version uses 20 Sampling Steps and is not a Turbo 4-Step workflow.
- Avoid changing the group labeled «إعدادات ملناش دعوة بيها» unless you understand the nodes.
- Higher resolutions increase rendering time, memory consumption, and cost.

🎁 **زي ما وعدتكم، الورك فلو والتطبيق متاحين مجانًا للجميع في أول تعليق!**تقدروا تستخدموا التطبيق بسهولة من الموبايل أو م...
11/08/2026

🎁 **زي ما وعدتكم، الورك فلو والتطبيق متاحين مجانًا للجميع في أول تعليق!**

تقدروا تستخدموا التطبيق بسهولة من الموبايل أو من أي جهاز متصل بالإنترنت، من غير الحاجة إلى جهاز قوي أو إمكانيات تشغيل محلية، وبدون التعامل مع تعقيدات النودز داخل ComfyUI. 📱☁️

وده ضمن سلسلة **«كومفي للجميع | Comfy for Everyone»**، المخصصة لكل شخص حابب يستخدم ComfyUI ولكن إمكانيات جهازه لا تسمح بالتشغيل المحلي.

ودي أول خطوة لتسهيل الأمور على الناس الغالية علينا، اللي كان نفسهم يستخدموا تطبيقات وورك فلوهات ComfyUI بسهولة، وبواجهة بسيطة تدعم **اللغة العربية**. والقادم أفضل بإذن الله ❤️🔥

━━━━━━━━━━━━━━━━━━━━

🚀 **Fastest MiniMax H3 Text to Video — 15 Seconds**

ورك فلو سريع يحوّل الوصف النصي مباشرة إلى فيديو سينمائي مدته تصل إلى **15 ثانية**، مع توليد الحركة والحوار والمؤثرات الصوتية المتزامنة داخل MiniMax H3.

✅ لا يحتاج إلى صورة بداية أو نهاية.
✅ لا يحتاج إلى ملف صوتي خارجي.
✅ يدعم الحوار العربي والمؤثرات الصوتية.
✅ يستخدم **Turbo LoRA** للتوليد السريع.
✅ يولّد الفيديو والصوت خلال **4 خطوات للفيديو + 4 خطوات للصوت**.
✅ يدعم نسب عرض متعددة، منها: **16:9 و9:16 و3:4 و1:1**.
✅ يصدر الفيديو بصيغة **H.264 MP4** بمعدل **24 FPS**.

━━━━━━━━━━━━━━━━━━━━

📱 **طريقة تشغيل التطبيق — AI App**

1️⃣ افتح رابط التطبيق الموجود في أول تعليق.
2️⃣ سجّل الدخول إلى حسابك على **RunningHub**.
3️⃣ اكتب وصف الفيديو المطلوب داخل خانة **Prompt**.
4️⃣ حدد مدة الفيديو والدقة إذا كانت الخيارات متاحة.
5️⃣ اضغط **Run** لبدء التوليد.
6️⃣ انتظر حتى يكتمل إنشاء الفيديو والصوت.
7️⃣ شاهد النتيجة ثم حمّل الفيديو بصيغة **MP4**.

✅ يعمل من الهاتف والكمبيوتر.
✅ لا يحتاج إلى تثبيت ComfyUI.
✅ لا يحتاج إلى جهاز قوي أو كارت شاشة.
✅ مناسب للمبتدئين ويدعم البرومبت العربي.

━━━━━━━━━━━━━━━━━━━━

⚙️ **طريقة تشغيل الورك فلو — Workflow**

1️⃣ افتح رابط الورك فلو الموجود في أول تعليق.

2️⃣ اضغط **Run** أو **Clone** لفتحه داخل واجهة ComfyUI السحابية في RunningHub.

3️⃣ اكتب وصف الفيديو داخل نود:

**Text — البرومبت**

يُفضّل أن يحتوي الوصف على:

* الشخصيات والملابس.
* المكان والوقت.
* ترتيب الأحداث والحركات.
* حركة الكاميرا.
* الإضاءة والأجواء.
* الحوار وتوقيت كل جملة.
* المؤثرات والأصوات المحيطة.

4️⃣ حدد مدة الفيديو من:

**Float — Duration**

المدة الافتراضية والقصوى لهذا الإصدار هي:

**15 Seconds**

5️⃣ اختر نسبة العرض والدقة من:

**ResolutionSelector**

الإعداد الافتراضي:

**16:9 — 0.3 MP — 736×416**

6️⃣ يُفضّل بدء الاختبار بدقة **0.2 أو 0.3 MP** لتقليل وقت التوليد والتكلفة.

7️⃣ اضغط:

**Queue Prompt**

8️⃣ انتظر حتى يقوم MiniMax H3 بتوليد الفيديو والصوت وفك ترميزهما.

9️⃣ شاهد النتيجة داخل نود:

**Video Combine**

🔟 حمّل الفيديو النهائي بصيغة **MP4**.

━━━━━━━━━━━━━━━━━━━━

🔊 **التحكم في الصوت**

يمكن للنموذج توليد:

* الحوار المتزامن مع حركة الشفاه.
* أصوات البيئة.
* حركة السيارات والأجسام.
* أصوات الخطوات والرياح.
* المؤثرات السينمائية.
* الأجواء المحيطة بالمكان.

إذا كنت لا تريد حوارًا أو غناءً، أضف إلى البرومبت:

**No dialogue, no narration, no singing, no speech, and no generated vocals. Generate only synchronized environmental ambience and physical sound effects.**

━━━━━━━━━━━━━━━━━━━━

📌 **نصائح مهمة**

* اكتب الأحداث بترتيب زمني واضح.
* لا تضع عددًا مبالغًا فيه من الأحداث خلال 15 ثانية.
* وزّع الحوار على توقيت يسمح بنطقه بصورة طبيعية.
* اذكر أن الحوار باللهجة المصرية إذا كان مطلوبًا.
* ابدأ بدقة منخفضة قبل الرندر النهائي.
* جرّب أكثر من **Seed** لأن النتيجة تختلف في كل تشغيل.
* لا تعدّل منطقة **«إعدادات ملناش دعوة بيها»** إلا إذا كنت تعرف وظيفة النودز.
* رفع الدقة يزيد وقت الرندر واستهلاك الذاكرة والتكلفة.

━━━━━━━━━━━━━━━━━━━━

🎁 **As promised, the workflow and AI App are completely free for everyone!**

You can easily use the App from your mobile phone or any internet-connected device—without needing a powerful computer, local hardware, or any experience working with complex ComfyUI nodes. 📱☁️

This release is part of the **“Comfy for Everyone | كومفي للجميع”** series, created especially for people who want to explore and use ComfyUI but do not have the hardware required to run it locally.

This is our first step toward making everything easier for our valued community—especially those who have always wanted accessible ComfyUI Apps and workflows with a simple interface and **Arabic-language support**. More is coming soon! ❤️🔥

━━━━━━━━━━━━━━━━━━━━

🚀 **Fastest MiniMax H3 Text to Video — 15 Seconds**

This fast workflow transforms a text prompt directly into a cinematic video of up to **15 seconds**, generating motion, dialogue, environmental ambience, and synchronized sound effects with MiniMax H3.

✅ No First Frame or Last Frame is required.
✅ No external audio file is required.
✅ Supports Arabic dialogue and synchronized sound.
✅ Uses **Turbo LoRA** for faster generation.
✅ Generates video and audio using **4 video steps + 4 audio steps**.
✅ Supports multiple aspect ratios, including **16:9, 9:16, 3:4, and 1:1**.
✅ Exports the result as an **H.264 MP4** file at **24 FPS**.

━━━━━━━━━━━━━━━━━━━━

📱 **How to Run the AI App**

1️⃣ Open the AI App link in the first comment.
2️⃣ Sign in to your **RunningHub** account.
3️⃣ Enter your video description in the **Prompt** field.
4️⃣ Select the duration and resolution if these controls are available.
5️⃣ Press **Run** to begin generation.
6️⃣ Wait for the video and synchronized audio to be generated.
7️⃣ Preview the result and download it as an **MP4** file.

✅ Works on mobile phones and computers.
✅ No ComfyUI installation is required.
✅ No powerful computer or GPU is required.
✅ Beginner-friendly and supports Arabic prompts.

━━━━━━━━━━━━━━━━━━━━

⚙️ **How to Run the Workflow**

1️⃣ Open the Workflow link in the first comment.

2️⃣ Press **Run** or **Clone** to open it in RunningHub’s cloud ComfyUI interface.

3️⃣ Enter your video description inside:

**Text — Prompt**

Your prompt should describe:

* Characters and clothing.
* Location and time.
* Chronological actions.
* Camera movement.
* Lighting and atmosphere.
* Dialogue and approximate timing.
* Environmental and physical sound effects.

4️⃣ Set the video duration through:

**Float — Duration**

The default and maximum duration for this version is:

**15 Seconds**

5️⃣ Select the aspect ratio and resolution through:

**ResolutionSelector**

Default setting:

**16:9 — 0.3 MP — 736×416**

6️⃣ Begin testing at **0.2 or 0.3 MP** to reduce generation time and cost.

7️⃣ Press:

**Queue Prompt**

8️⃣ Wait for MiniMax H3 to generate and decode the video and audio.

9️⃣ Preview the final result inside:

**Video Combine**

🔟 Download the completed video as an **MP4** file.

━━━━━━━━━━━━━━━━━━━━

🔊 **Audio Control**

The model can generate:

* Lip-synchronized dialogue.
* Environmental ambience.
* Vehicle and object sounds.
* Footsteps and wind.
* Cinematic physical effects.
* Location-specific background sound.

If you do not want dialogue, speech, or singing, add:

**No dialogue, no narration, no singing, no speech, and no generated vocals. Generate only synchronized environmental ambience and physical sound effects.**

━━━━━━━━━━━━━━━━━━━━

📌 **Important Recommendations**

* Describe events in a clear chronological order.
* Avoid including too many actions within 15 seconds.
* Allow enough time for dialogue to be spoken naturally.
* Specify the required language and accent.
* Test at a low resolution before the final render.
* Try multiple seeds because each generation may differ.
* Avoid changing the group labeled **“إعدادات ملناش دعوة بيها”** unless you understand the nodes.
* Higher resolutions significantly increase rendering time, memory usage, and cost.

🎁 **زي ما وعدتكم، الورك فلو والتطبيق متاحين مجانًا للجميع في أول تعليق!**تقدروا تستخدموا التطبيق بسهولة من الموبايل أو م...
11/08/2026

🎁 **زي ما وعدتكم، الورك فلو والتطبيق متاحين مجانًا للجميع في أول تعليق!**

تقدروا تستخدموا التطبيق بسهولة من الموبايل أو من أي جهاز متصل بالإنترنت، من غير الحاجة إلى جهاز قوي أو إمكانيات تشغيل محلية، وبدون التعامل مع تعقيدات النودز داخل ComfyUI. 📱☁️

وده ضمن سلسلة **«كومفي للجميع | Comfy for Everyone»**، المخصصة لكل شخص حابب يستخدم ComfyUI ولكن إمكانيات جهازه لا تسمح بالتشغيل المحلي.

ودي أول خطوة لتسهيل الأمور على الناس الغالية علينا، اللي كان نفسهم يستخدموا تطبيقات وورك فلوهات ComfyUI بسهولة، وبواجهة بسيطة تدعم **اللغة العربية**. والقادم أفضل بإذن الله ❤️🔥

━━━━━━━━━━━━━━━━━━━━

📱 **طريقة تشغيل الـAI App:**

1️⃣ افتح رابط التطبيق الموجود في أول تعليق.
2️⃣ سجّل الدخول إلى حسابك على **RunningHub**.
3️⃣ ارفع الصورة التي تريد أن يبدأ بها الفيديو داخل **First Frame**.
4️⃣ ارفع الصورة التي تريد أن ينتهي عندها الفيديو داخل **Last Frame**.
5️⃣ اكتب وصف الحركة والانتقال المطلوب بين الصورتين داخل خانة **Prompt**.
6️⃣ اختر مدة الفيديو والدقة المناسبة إذا كانت الخيارات متاحة.
7️⃣ اضغط **Run** وانتظر حتى يكتمل التوليد.
8️⃣ شاهد الفيديو النهائي، ثم قم بتحميله على جهازك.

✅ يمكن تشغيل التطبيق من الهاتف أو الكمبيوتر.
✅ لا تحتاج إلى تثبيت ComfyUI.
✅ لا تحتاج إلى جهاز قوي أو كارت شاشة.
✅ الواجهة سهلة وتدعم اللغة العربية.

━━━━━━━━━━━━━━━━━━━━

⚙️ **طريقة تشغيل الـWorkflow:**

1️⃣ افتح رابط الورك فلو الموجود في أول تعليق.
2️⃣ سجّل الدخول إلى **RunningHub**.
3️⃣ اضغط **Run** أو **Clone** لفتح الورك فلو داخل واجهة ComfyUI السحابية.
4️⃣ ارفع صورة البداية داخل نود:

**اللقطة الأولى — First Frame**

5️⃣ ارفع صورة النهاية داخل نود:

**اللقطة الأخيرة — Last Frame**

6️⃣ عدّل البرومبت داخل نود:

**MiniMaxH3ImageToVideo**

واشرح بالترتيب الحركة والأحداث ومسار الكاميرا من الصورة الأولى إلى الصورة الأخيرة.

7️⃣ اختر المدة والدقة المناسبة. ويُفضّل إجراء الاختبار الأول بدقة منخفضة مثل **0.2 أو 0.3 MP** لتقليل وقت التوليد والتكلفة.

8️⃣ اضغط:

**Queue Prompt**

9️⃣ انتظر حتى ينتهي MiniMax H3 من توليد الفيديو والصوت المتزامن.

🔟 بعد اكتمال المعالجة، يمكنك مشاهدة الفيديو وتحميله بصيغة **MP4**.

⚠️ لا تعدّل نودز الموديل أو الـVAE أو الـSampler أو إعدادات التسريع إذا لم تكن تعرف وظيفتها. يكفي تغيير الصور والبرومبت والمدة والدقة فقط.

━━━━━━━━━━━━━━━━━━━━

🎁 **As promised, the workflow and AI App are completely free for everyone!**

You can easily use the App from your mobile phone or any internet-connected device—without needing a powerful computer, local hardware, or any experience working with complex ComfyUI nodes. 📱☁️

This release is part of the **“Comfy for Everyone | كومفي للجميع”** series, created especially for people who want to explore and use ComfyUI but do not have the hardware required to run it locally.

This is our first step toward making everything easier for our valued community—especially those who have always wanted accessible ComfyUI Apps and workflows with a simple interface and **Arabic-language support**. More is coming soon! ❤️🔥

━━━━━━━━━━━━━━━━━━━━

📱 **How to Run the AI App:**

1️⃣ Open the AI App link in the first comment.
2️⃣ Sign in to your **RunningHub** account.
3️⃣ Upload the image that will define the **First Frame**.
4️⃣ Upload the image that will define the **Last Frame**.
5️⃣ Enter a prompt describing the required motion, events, and transition between the two images.
6️⃣ Select the desired duration and resolution if these options are available.
7️⃣ Press **Run** and wait for the generation to finish.
8️⃣ Preview and download the final video.

✅ Works on mobile phones and computers.
✅ No ComfyUI installation is required.
✅ No powerful computer or GPU is required.
✅ Simple interface with Arabic-language support.

━━━━━━━━━━━━━━━━━━━━

⚙️ **How to Run the Workflow:**

1️⃣ Open the Workflow link in the first comment.
2️⃣ Sign in to **RunningHub**.
3️⃣ Press **Run** or **Clone** to open it in the cloud ComfyUI interface.
4️⃣ Upload the starting image through:

**اللقطة الأولى — First Frame**

5️⃣ Upload the ending image through:

**اللقطة الأخيرة — Last Frame**

6️⃣ Edit the prompt inside:

**MiniMaxH3ImageToVideo**

Describe the complete chronological movement, events, and camera journey from the first image to the final image.

7️⃣ Select the required duration and resolution. Begin testing at **0.2 or 0.3 MP** to reduce generation time and cost.

8️⃣ Press:

**Queue Prompt**

9️⃣ Wait for MiniMax H3 to generate the video and synchronized audio.

🔟 Preview and download the completed video as an **MP4** file.

⚠️ Avoid changing the model, VAE, sampler, cache, or acceleration nodes unless you understand their functions. You normally only need to change the images, prompt, duration, and resolution.

🇪🇬 شرح ورك فلو MiniMax H3 — تحويل أول وآخر فريم إلى فيديو مدته 15 ثانية

يحوّل هذا الورك فلو صورتين ثابتتين إلى فيديو سينمائي متصل باستخدام MiniMax H3:

🖼️ الصورة الأولى تحدد بداية الفيديو.
🖼️ الصورة الثانية تحدد نهاية الفيديو.
📝 البرومبت يصف الحركة والأحداث والانتقال المطلوب بينهما.

ويقوم النموذج بإنشاء كل الفريمات والحركات الموجودة بين الصورتين تلقائيًا، مع محاولة الوصول إلى تكوين الصورة الأخيرة في نهاية الفيديو.

الفكرة ببساطة:

First Frame + Last Frame + Prompt = Complete Video

━━━━━━━━━━━━━━━━━━━━

🎯 ما وظيفة الورك فلو؟

يمكن استخدام الورك فلو لإنشاء:

انتقالات سينمائية بين مكانين مختلفين.
Zoom من الفضاء إلى موقع على الأرض.
تحولات زمنية أو بيئية.
انتقال من الطبيعة إلى الدمار أو العكس.
انتقال بين لقطتين لشخصية واحدة.
تغيّر تعبيرات الوجه أو وضعية الجسم.
تحريك الكاميرا بين مشهدين.
مشاهد قصصية متصلة من البداية إلى النهاية.
فيديوهات إعلانية وسينمائية قصيرة.
فيديو يصل إلى نهاية محددة بدل ترك النهاية عشوائية.

يستخدم النموذج الصورتين كـ Boundary Frames، أي حدّين بصريين يحددان نقطة البداية ونقطة النهاية.

━━━━━━━━━━━━━━━━━━━━

🖼️ دور الصورة الأولى — First Frame

الصورة الأولى هي نقطة البداية الدقيقة للفيديو، وتحدد:

تكوين أول لقطة.
مكان الكاميرا واتجاهها.
الشخصيات الظاهرة في البداية.
الخلفية والموقع.
الإضاءة والألوان.
الملابس والهوية البصرية.
العناصر الموجودة داخل المشهد.
وضعية الجسم وتعبير الوجه.
نقطة انطلاق الحركة.

يجب أن يبدأ الفيديو قريبًا جدًا من تكوين الصورة الأولى، ثم يبدأ التحرك منها بصورة طبيعية.

ارفعها داخل النود المسماة:

اللقطة الأولى — First Frame

━━━━━━━━━━━━━━━━━━━━

🖼️ دور الصورة الأخيرة — Last Frame

الصورة الثانية تحدد الوجهة البصرية النهائية للفيديو، ومنها:

التكوين النهائي.
مكان الشخصية في نهاية المشهد.
ملامح الوجه والملابس.
وضعية الجسم واليدين.
موقع الكاميرا النهائي.
الخلفية النهائية.
الإضاءة والألوان.
الحالة العاطفية.
العناصر التي يجب أن تظهر في النهاية.

يجب أن يتطور الفيديو تدريجيًا حتى يصل إلى تكوين قريب من الصورة الثانية عند آخر فريم.

ارفعها داخل النود المسماة:

اللقطة الأخيرة — Last Frame

💡 الصورة الأخيرة لا تعني أن النموذج سيكررها حرفيًا بشكل مضمون، لكنها تعمل كمرساة قوية توجه الحركة والتكوين النهائي.

━━━━━━━━━━━━━━━━━━━━

📝 دور البرومبت

البرومبت يشرح للنموذج كيفية الانتقال من الصورة الأولى إلى الصورة الأخيرة.

يُفضل أن يحتوي على:

تعريف واضح لدور كل صورة.
ما يحدث بعد بداية الفيديو.
حركة الشخصية أو العناصر.
مسار الكاميرا.
الأحداث الوسطية.
طريقة ظهور عناصر الصورة الأخيرة.
التوقيت التقريبي لكل مرحلة.
المؤثرات الصوتية المطلوبة.
العناصر التي يجب الحفاظ عليها.
القيود التي تمنع تغير الهوية أو الملابس.
شكل اللقطة النهائية.

يحتوي الملف الحالي على برومبت تجريبي يبدأ من كوكب الأرض في الفضاء، ثم تتحرك الكاميرا إلى الطبيعة، وبعد ذلك تمر تدريجيًا بمناطق مدمرة، حتى تدخل منزلًا متضررًا وتصل إلى الطفلة الموجودة في الصورة الأخيرة.

يمكن استبدال هذا البرومبت بالكامل ليناسب صورك، مع الحفاظ على فكرة:

Image1 = Exact First-Frame Anchor
Image2 = Exact Final-Frame Anchor

━━━━━━━━━━━━━━━━━━━━

🔊 الصوت في الورك فلو

هذا الإصدار لا يحتاج إلى رفع ملف صوتي خارجي.

لا توجد نود:

Load Audio

يقوم MiniMax H3 بتوليد الصوت المتزامن مع الفيديو اعتمادًا على وصف المشهد، مثل:

أصوات البيئة.
الرياح.
الطيور.
حركة السيارات.
الانفجارات أو الدمار.
خطوات الشخصيات.
أصوات الأشياء المتحركة.
المؤثرات السينمائية.
أجواء المكان.

يتم توليد الصوت داخل المسار نفسه، ثم فك ترميزه باستخدام:

MiniMax H3 Audio VAE FP32

إذا كنت لا تريد كلامًا أو غناءً، اكتب ذلك بوضوح داخل البرومبت:

No dialogue, no singing, no speech, no whispering, and no generated vocals. Generate only synchronized environmental ambience and physical sound effects.

━━━━━━━━━━━━━━━━━━━━

🧠 كيف يعمل الورك فلو؟

1️⃣ يتم رفع صورة البداية داخل:

اللقطة الأولى — First Frame

2️⃣ يتم رفع صورة النهاية داخل:

اللقطة الأخيرة — Last Frame

3️⃣ يتم كتابة وصف الحركة والانتقال داخل نود:

MiniMaxH3ImageToVideo

4️⃣ يتم تحديد مدة الفيديو من:

المدة الزمنية — Duration

5️⃣ يتم اختيار نسبة العرض والدقة من:

حجم الفيديو — Aspect Ratio

6️⃣ يحسب الورك فلو عدد الفريمات اعتمادًا على المدة ومعدل:

24 FPS

7️⃣ يتم تعديل عدد الفريمات تلقائيًا ليتوافق مع البنية الزمنية التي يحتاجها MiniMax H3.

8️⃣ يتم تحليل البرومبت والصورتين باستخدام:

Qwen3‑VL 32B MiniMax H3

9️⃣ يستخدم الورك فلو نموذج التوليد:

MiniMax H3 FL2VA BF16

10️⃣ تتم معالجة الفيديو والصوت باستخدام:

MiniMax H3 Video VAE FP16
MiniMax H3 Audio VAE FP32

1️⃣1️⃣ يتم تطبيق تقنيات تقليل استهلاك الذاكرة وتسريع المعالجة.

1️⃣2️⃣ يبدأ الـSampling لإنشاء الحركة والفريمات الوسيطة والصوت المتزامن.

1️⃣3️⃣ يتم فك ترميز الفيديو والصوت بشكل منفصل.

1️⃣4️⃣ يتم دمجهما داخل ملف MP4 نهائي باستخدام:

Video Combine

━━━━━━━━━━━━━━━━━━━━

⚡ تقنيات التسريع الموجودة

يحتوي الورك فلو على:

MiniMax H3 Memory Efficient Sage Attention
Sage Attention Patch
MiniMax H3 Block Cache T8
معالجة الكاش على CPU
نموذج Qwen3‑VL NVFP4 AWQ
مسار MiniMax H3 FL2VA BF16

تساعد هذه الإعدادات على تقليل استهلاك الذاكرة وتسريع التوليد مقارنة بتشغيل المسار الكامل من دون تحسينات.

⚠️ لا يحتوي هذا الملف على:

Turbo LoRA.
EasyCache.
SpectrumApply.
RIFE Frame Interpolation.
ملف صوتي خارجي.

إذا ظهرت أخطاء توافق أو تغيّرت الحركة، يمكن تجربة تعطيل:

MiniMax H3 Block Cache
أو
Memory Efficient Sage Attention Patch

لكن لا تعدّل هذه المنطقة إلا إذا كنت تعرف تأثيرها على الأداء والذاكرة.

━━━━━━━━━━━━━━━━━━━━

⚙️ الإعدادات الحالية داخل الملف
Video Duration: 15 Seconds
Base Frame Rate: 24 FPS
Aspect Ratio: 16:9
Current Megapixels: 0.3 MP
Current Output Size: 736×416
Size Multiple: 32
Generated Frames for 15 Seconds: 374 Frames
Model: MiniMax H3 FL2VA BF16
Text/Vision Encoder: Qwen3‑VL 32B NVFP4 AWQ
Video VAE: MiniMax H3 Video VAE FP16
Audio VAE: MiniMax H3 Audio VAE FP32
Sampler: RES Multistep
Scheduler: Simple
Sampling Steps: 20
Denoise: 1.0
Seed Mode: Randomize
Output Frame Rate: 24 FPS
Output Format: H.264 MP4
Final Generated Audio: Enabled
Save Output: Enabled

━━━━━━━━━━━━━━━━━━━━

⏱️ حساب عدد الفريمات

يعتمد الورك فلو على المدة المحددة ومعدل 24 FPS.

عند اختيار 15 ثانية:

15 × 24 = 360 فريمًا

بعد ذلك يعدّل الورك فلو العدد تلقائيًا ليتوافق مع البنية الزمنية لنموذج MiniMax H3، فيصبح:

374 فريمًا

ويتم ذلك باستخدام معادلة داخل نود:

ComfyMathExpression

لذلك قد يكون طول الفيديو الناتج أكبر قليلًا من المدة الاسمية المحددة.

الملف النهائي يعمل عند 24 FPS، ولا تتم زيادة المعدل إلى 60 FPS لأن هذا الورك فلو لا يحتوي على RIFE.

━━━━━━━━━━━━━━━━━━━━

📐 الدقات المتاحة بنسبة 16:9
Megapixels الدقة
0.2 MP 608×352
0.3 MP 736×416
0.4 MP 864×480
0.5 MP 960×544
0.6 MP 1056×608
0.7 MP 1152×640
0.8 MP 1216×672
0.9 MP 1280×736
0.98 MP 1344×768
1.0 MP 1376×768
1.2 MP 1504×832
1.5 MP 1664×928
1.8 MP 1824×1024
2.0 MP 1920×1088

💡 ابدأ الاختبار على 0.2 أو 0.3 MP، وبعد التأكد من نجاح الحركة والوصول إلى الفريم الأخير استخدم 0.98 أو 1.0 MP للنتيجة النهائية.

رفع الدقة يزيد وقت التوليد واستهلاك الذاكرة بصورة كبيرة.

━━━━━━━━━━━━━━━━━━━━

☁️ طريقة تشغيل الورك فلو

1️⃣ افتح الورك فلو داخل ComfyUI أو RunningHub.

2️⃣ ارفع صورة البداية داخل:

اللقطة الأولى — First Frame

3️⃣ ارفع صورة النهاية داخل:

اللقطة الأخيرة — Last Frame

4️⃣ عدّل البرومبت الموجود داخل:

MiniMaxH3ImageToVideo

5️⃣ اشرح بالتسلسل ما يجب أن يحدث بين الصورتين.

6️⃣ حدد مدة الفيديو من:

المدة الزمنية — Duration

7️⃣ اختر نسبة العرض والدقة من:

حجم الفيديو — Aspect Ratio

8️⃣ ابدأ بدقة 0.2 أو 0.3 MP للاختبار.

9️⃣ اضغط:

Queue Prompt

🔟 بعد اكتمال التوليد، سيقوم الورك فلو بفك ترميز الفيديو والصوت ودمجهما تلقائيًا داخل ملف MP4.

━━━━━━━━━━━━━━━━━━━━

📌 نصائح للحصول على أفضل نتيجة
استخدم صورتين لهما نسبة عرض متقاربة.
اختر صورتين واضحتين وعاليتَي الجودة.
اجعل الانتقال بينهما منطقيًا وقابلًا للتنفيذ.
اشرح حركة الكاميرا بصورة زمنية ومتسلسلة.
لا تطلب عددًا كبيرًا من الأحداث خلال 15 ثانية.
لا تُظهر عناصر الفريم الأخير مبكرًا إذا كنت تريد مفاجأة بصرية.
حافظ على هوية الشخصية وملابسها داخل البرومبت.
وضّح متى تبدأ الشخصية في الظهور.
حدد شكل اللقطة الأخيرة بدقة.
استخدم حركة واحدة متصلة إذا كنت تريد انتقالًا سلسًا.
تجنب القطع المفاجئ إذا كان الهدف رحلة كاميرا مستمرة.
لا تطلب انتقالًا مستحيلًا من دون شرح بصري.
ابدأ بدقة منخفضة قبل التوليد النهائي.
جرّب أكثر من Seed لأن النتيجة قد تختلف بين كل تشغيل.

لمنع ظهور الفريم الأخير مبكرًا، استخدم:

Do not reveal Image2, its characters, location, or final composition before the camera physically reaches the final scene.

للحفاظ على الهوية:

Strict identity consistency—preserve the exact face, age, hairstyle, clothing, body proportions, colors, and accessories shown in Image2.

للحصول على نهاية قوية:

The final moments must settle naturally into the exact composition, camera angle, subject placement, lighting, and emotional state defined by Image2.

━━━━━━━━━━━━━━━━━━━━

⚠️ ملاحظات مهمة
الصورة الأولى تحدد بداية الفيديو فقط.
الصورة الثانية تحدد نهاية الفيديو المطلوبة.
البرومبت هو المسؤول عن بناء الرحلة بينهما.
لا يحتاج الورك فلو إلى ملف صوتي خارجي.
يقوم النموذج بتوليد المؤثرات الصوتية من وصف المشهد.
النتيجة النهائية تعمل عند 24 FPS.
لا توجد معالجة RIFE داخل هذا الإصدار.
إعداد 15 ثانية يولد 374 فريمًا داخل الملف الحالي.
الدقة الحالية هي 736×416.
الـSeed مضبوط على Randomize، لذلك تختلف النتيجة مع كل تشغيل.
الوصول إلى الفريم الأخير قد لا يكون مطابقًا 100% في كل محاولة.

لا تعدّل المنطقة المسماة:

إعدادات ملناش دعوة بيها

إلا إذا كنت تعرف تأثير النودز على النموذج والـVAE والـSampling والذاكرة.

🇬🇧 MiniMax H3 Workflow — First and Last Frame to a 15-Second Video

This workflow transforms two still images into one continuous cinematic video using MiniMax H3:

🖼️ The first image defines the beginning of the video.
🖼️ The second image defines the intended ending.
📝 The prompt describes the movement, events, camera path, and transition between them.

MiniMax H3 generates the motion and all intermediate frames while attempting to arrive naturally at the composition defined by the final image.

The concept is simple:

First Frame + Last Frame + Prompt = Complete Video

━━━━━━━━━━━━━━━━━━━━

🎯 Workflow Purpose

The workflow can generate:

Cinematic transitions between two locations.
Space-to-Earth zoom sequences.
Environmental or time-based transformations.
Transitions from natural beauty to destruction.
Character pose or expression changes.
Continuous camera journeys.
Short visual narratives.
Advertising and cinematic videos.
Videos with a visually guided ending rather than a random final shot.

The two images operate as boundary frames, defining the visual starting point and intended destination.

━━━━━━━━━━━━━━━━━━━━

🖼️ First Frame Role

The first image defines:

The opening composition.
Initial camera position and orientation.
Starting characters and objects.
Environment and background.
Lighting and colors.
Clothing and visual identity.
Body pose and facial expression.
The physical starting point of the action.

Upload it through:

اللقطة الأولى — First Frame

The video should begin close to this image before moving naturally away from it.

━━━━━━━━━━━━━━━━━━━━

🖼️ Last Frame Role

The second image defines the intended final state, including:

Final composition.
Character identity and appearance.
Final body and hand positions.
Clothing and accessories.
Camera position.
Background and location.
Lighting and colors.
Emotional state.
Objects visible at the end.

Upload it through:

اللقطة الأخيرة — Last Frame

The video should gradually evolve toward this image during its final moments.

The last frame is a strong visual anchor, although exact pixel-level reproduction is not guaranteed in every generation.

━━━━━━━━━━━━━━━━━━━━

📝 Prompt Role

The prompt explains how the video travels from the first frame to the last frame.

A strong prompt should define:

The role of each reference image.
Opening action.
Subject motion.
Camera movement.
Intermediate events.
Environmental progression.
When the final character or location appears.
Approximate timing.
Required sound effects.
Identity-preservation instructions.
Elements that must not change.
The exact intended ending.

The included example begins with Earth in deep space, descends toward a beautiful natural landscape, moves through progressively damaged areas, and eventually enters a destroyed home to reach the girl shown in the final image.

You can completely replace this prompt while preserving the structure:

Image1 = Exact First-Frame Anchor
Image2 = Exact Final-Frame Anchor

━━━━━━━━━━━━━━━━━━━━

🔊 Generated Audio

This workflow does not require an external audio file.

It does not contain a:

Load Audio

input node.

MiniMax H3 generates synchronized audio from the scene description, including:

Environmental ambience.
Wind and weather.
Birds and animals.
Vehicles.
Footsteps.
Moving objects.
Destruction and impact sounds.
Physical cinematic sound effects.
Location-specific atmosphere.

The generated audio is decoded through:

MiniMax H3 Audio VAE FP32

If dialogue, singing, or generated voices are not wanted, reinforce:

No dialogue, no singing, no speech, no whispering, and no generated vocals. Generate only synchronized environmental ambience and physical sound effects.

━━━━━━━━━━━━━━━━━━━━

🧠 How the Workflow Operates

1️⃣ Upload the starting image through:

اللقطة الأولى — First Frame

2️⃣ Upload the ending image through:

اللقطة الأخيرة — Last Frame

3️⃣ Enter the transition and action description inside:

MiniMaxH3ImageToVideo

4️⃣ Set the video duration through:

المدة الزمنية — Duration

5️⃣ Select the aspect ratio and resolution through:

حجم الفيديو — Aspect Ratio

6️⃣ The workflow calculates the initial frame count at 24 FPS.

7️⃣ It automatically adjusts that count to match MiniMax H3’s temporal structure.

8️⃣ The prompt and reference images are processed using:

Qwen3‑VL 32B MiniMax H3

9️⃣ Video and audio generation use:

MiniMax H3 FL2VA BF16

🔟 The visual and audio representations are processed through:

MiniMax H3 Video VAE FP16
MiniMax H3 Audio VAE FP32

1️⃣1️⃣ Memory-efficient attention and block caching are applied.

1️⃣2️⃣ Sampling generates the motion, intermediate frames, and synchronized audio.

1️⃣3️⃣ Video and audio are decoded separately.

1️⃣4️⃣ Both are combined and exported through:

Video Combine

━━━━━━━━━━━━━━━━━━━━

⚡ Acceleration Components

The workflow contains:

MiniMax H3 Memory Efficient Sage Attention
Sage Attention Patch
MiniMax H3 Block Cache T8
CPU-based block-cache processing.
Qwen3‑VL NVFP4 AWQ
MiniMax H3 FL2VA BF16

These components reduce memory usage and improve generation speed.

This file does not contain:

Turbo LoRA.
EasyCache.
SpectrumApply.
RIFE Frame Interpolation.
External audio input.

If compatibility errors or major motion changes appear, test the workflow with Block Cache or Sage Attention disabled.

━━━━━━━━━━━━━━━━━━━━

⚙️ Current File Settings
Video Duration: 15 Seconds
Base Frame Rate: 24 FPS
Aspect Ratio: 16:9
Current Megapixels: 0.3 MP
Current Output Size: 736×416
Size Multiple: 32
Generated Frames: 374
Model: MiniMax H3 FL2VA BF16
Text/Vision Encoder: Qwen3‑VL 32B NVFP4 AWQ
Video VAE: MiniMax H3 Video VAE FP16
Audio VAE: MiniMax H3 Audio VAE FP32
Sampler: RES Multistep
Scheduler: Simple
Sampling Steps: 20
Denoise: 1.0
Seed Mode: Randomize
Output Frame Rate: 24 FPS
Output Format: H.264 MP4
Generated Audio: Enabled
Save Output: Enabled

━━━━━━━━━━━━━━━━━━━━

⏱️ Frame Calculation

For a 15-second video at 24 FPS:

15 × 24 = 360 frames

The workflow then adjusts the sequence to satisfy MiniMax H3’s temporal structure:

Adjusted frame count = 374 frames

This calculation is performed automatically through:

ComfyMathExpression

The final output remains at 24 FPS. It is not interpolated to 60 FPS because this version does not contain RIFE.

━━━━━━━━━━━━━━━━━━━━

📐 Available 16:9 Resolutions
Megapixels Output Resolution
0.2 MP 608×352
0.3 MP 736×416
0.4 MP 864×480
0.5 MP 960×544
0.6 MP 1056×608
0.7 MP 1152×640
0.8 MP 1216×672
0.9 MP 1280×736
0.98 MP 1344×768
1.0 MP 1376×768
1.2 MP 1504×832
1.5 MP 1664×928
1.8 MP 1824×1024
2.0 MP 1920×1088

Begin testing at 0.2 or 0.3 MP. After confirming motion quality and successful arrival at the last frame, use 0.98 or 1.0 MP for the final generation.

━━━━━━━━━━━━━━━━━━━━

☁️ How to Run the Workflow

1️⃣ Open the workflow in ComfyUI or RunningHub.

2️⃣ Upload the starting image through:

اللقطة الأولى — First Frame

3️⃣ Upload the ending image through:

اللقطة الأخيرة — Last Frame

4️⃣ Edit the prompt inside:

MiniMaxH3ImageToVideo

5️⃣ Describe the complete chronological transition between the images.

6️⃣ Set the duration through:

المدة الزمنية — Duration

7️⃣ Select the aspect ratio and resolution through:

حجم الفيديو — Aspect Ratio

8️⃣ Begin at 0.2 or 0.3 MP for testing.

9️⃣ Press:

Queue Prompt

🔟 The workflow generates and decodes the video and audio, combines them, and saves the final result as an MP4 file.

━━━━━━━━━━━━━━━━━━━━

📌 Best-Result Recommendations
Use two clear, high-quality images.
Keep their aspect ratios reasonably similar.
Create a physically understandable transition.
Describe camera movement chronologically.
Avoid excessive events within 15 seconds.
Do not reveal the final frame’s elements too early.
Explicitly preserve character identity and clothing.
State exactly when the final subject should appear.
Describe the final composition clearly.
Use one continuous camera movement for seamless journeys.
Avoid abrupt cuts if continuity is required.
Test at low resolution before final generation.
Try multiple seeds because every run may produce a different result.

To prevent early appearance of the final scene:

Do not reveal Image2, its characters, location, or final composition before the camera physically reaches the final scene.

For identity consistency:

Strict identity consistency—preserve the exact face, age, hairstyle, clothing, body proportions, colors, and accessories shown in Image2.

For a stronger ending:

The final moments must settle naturally into the exact composition, camera angle, subject placement, lighting, and emotional state defined by Image2.

━━━━━━━━━━━━━━━━━━━━

⚠️ Important Notes
Image1 defines the video’s starting frame.
Image2 defines the intended final frame.
The prompt constructs the complete journey between them.
No external audio file is required.
MiniMax H3 generates synchronized environmental audio.
The final output runs at 24 FPS.
This version does not use RIFE interpolation.
A 15-second setting produces 374 frames.
The current resolution is 736×416.
The seed is randomized, so results vary between runs.
Exact last-frame reproduction is not guaranteed every time.

Avoid modifying the group labeled:

إعدادات ملناش دعوة بيها

unless you understand its effect on the model, VAE, sampling, cache, and memory usage.

Address

Cairo
11731

Alerts

Be the first to know and let us send you an email when Mohamed Nagy posts news and promotions. Your email address will not be used for any other purpose, and you can unsubscribe at any time.

Shortcuts

Share