11/08/2026
🎁 **زي ما وعدتكم، الورك فلو والتطبيق متاحين مجانًا للجميع في أول تعليق!**
تقدروا تستخدموا التطبيق بسهولة من الموبايل أو من أي جهاز متصل بالإنترنت، من غير الحاجة إلى جهاز قوي أو إمكانيات تشغيل محلية، وبدون التعامل مع تعقيدات النودز داخل ComfyUI. 📱☁️
وده ضمن سلسلة **«كومفي للجميع | Comfy for Everyone»**، المخصصة لكل شخص حابب يستخدم ComfyUI ولكن إمكانيات جهازه لا تسمح بالتشغيل المحلي.
ودي أول خطوة لتسهيل الأمور على الناس الغالية علينا، اللي كان نفسهم يستخدموا تطبيقات وورك فلوهات ComfyUI بسهولة، وبواجهة بسيطة تدعم **اللغة العربية**. والقادم أفضل بإذن الله ❤️🔥
━━━━━━━━━━━━━━━━━━━━
📱 **طريقة تشغيل الـAI App:**
1️⃣ افتح رابط التطبيق الموجود في أول تعليق.
2️⃣ سجّل الدخول إلى حسابك على **RunningHub**.
3️⃣ ارفع الصورة التي تريد أن يبدأ بها الفيديو داخل **First Frame**.
4️⃣ ارفع الصورة التي تريد أن ينتهي عندها الفيديو داخل **Last Frame**.
5️⃣ اكتب وصف الحركة والانتقال المطلوب بين الصورتين داخل خانة **Prompt**.
6️⃣ اختر مدة الفيديو والدقة المناسبة إذا كانت الخيارات متاحة.
7️⃣ اضغط **Run** وانتظر حتى يكتمل التوليد.
8️⃣ شاهد الفيديو النهائي، ثم قم بتحميله على جهازك.
✅ يمكن تشغيل التطبيق من الهاتف أو الكمبيوتر.
✅ لا تحتاج إلى تثبيت ComfyUI.
✅ لا تحتاج إلى جهاز قوي أو كارت شاشة.
✅ الواجهة سهلة وتدعم اللغة العربية.
━━━━━━━━━━━━━━━━━━━━
⚙️ **طريقة تشغيل الـWorkflow:**
1️⃣ افتح رابط الورك فلو الموجود في أول تعليق.
2️⃣ سجّل الدخول إلى **RunningHub**.
3️⃣ اضغط **Run** أو **Clone** لفتح الورك فلو داخل واجهة ComfyUI السحابية.
4️⃣ ارفع صورة البداية داخل نود:
**اللقطة الأولى — First Frame**
5️⃣ ارفع صورة النهاية داخل نود:
**اللقطة الأخيرة — Last Frame**
6️⃣ عدّل البرومبت داخل نود:
**MiniMaxH3ImageToVideo**
واشرح بالترتيب الحركة والأحداث ومسار الكاميرا من الصورة الأولى إلى الصورة الأخيرة.
7️⃣ اختر المدة والدقة المناسبة. ويُفضّل إجراء الاختبار الأول بدقة منخفضة مثل **0.2 أو 0.3 MP** لتقليل وقت التوليد والتكلفة.
8️⃣ اضغط:
**Queue Prompt**
9️⃣ انتظر حتى ينتهي MiniMax H3 من توليد الفيديو والصوت المتزامن.
🔟 بعد اكتمال المعالجة، يمكنك مشاهدة الفيديو وتحميله بصيغة **MP4**.
⚠️ لا تعدّل نودز الموديل أو الـVAE أو الـSampler أو إعدادات التسريع إذا لم تكن تعرف وظيفتها. يكفي تغيير الصور والبرومبت والمدة والدقة فقط.
━━━━━━━━━━━━━━━━━━━━
🎁 **As promised, the workflow and AI App are completely free for everyone!**
You can easily use the App from your mobile phone or any internet-connected device—without needing a powerful computer, local hardware, or any experience working with complex ComfyUI nodes. 📱☁️
This release is part of the **“Comfy for Everyone | كومفي للجميع”** series, created especially for people who want to explore and use ComfyUI but do not have the hardware required to run it locally.
This is our first step toward making everything easier for our valued community—especially those who have always wanted accessible ComfyUI Apps and workflows with a simple interface and **Arabic-language support**. More is coming soon! ❤️🔥
━━━━━━━━━━━━━━━━━━━━
📱 **How to Run the AI App:**
1️⃣ Open the AI App link in the first comment.
2️⃣ Sign in to your **RunningHub** account.
3️⃣ Upload the image that will define the **First Frame**.
4️⃣ Upload the image that will define the **Last Frame**.
5️⃣ Enter a prompt describing the required motion, events, and transition between the two images.
6️⃣ Select the desired duration and resolution if these options are available.
7️⃣ Press **Run** and wait for the generation to finish.
8️⃣ Preview and download the final video.
✅ Works on mobile phones and computers.
✅ No ComfyUI installation is required.
✅ No powerful computer or GPU is required.
✅ Simple interface with Arabic-language support.
━━━━━━━━━━━━━━━━━━━━
⚙️ **How to Run the Workflow:**
1️⃣ Open the Workflow link in the first comment.
2️⃣ Sign in to **RunningHub**.
3️⃣ Press **Run** or **Clone** to open it in the cloud ComfyUI interface.
4️⃣ Upload the starting image through:
**اللقطة الأولى — First Frame**
5️⃣ Upload the ending image through:
**اللقطة الأخيرة — Last Frame**
6️⃣ Edit the prompt inside:
**MiniMaxH3ImageToVideo**
Describe the complete chronological movement, events, and camera journey from the first image to the final image.
7️⃣ Select the required duration and resolution. Begin testing at **0.2 or 0.3 MP** to reduce generation time and cost.
8️⃣ Press:
**Queue Prompt**
9️⃣ Wait for MiniMax H3 to generate the video and synchronized audio.
🔟 Preview and download the completed video as an **MP4** file.
⚠️ Avoid changing the model, VAE, sampler, cache, or acceleration nodes unless you understand their functions. You normally only need to change the images, prompt, duration, and resolution.
🇪🇬 شرح ورك فلو MiniMax H3 — تحويل أول وآخر فريم إلى فيديو مدته 15 ثانية
يحوّل هذا الورك فلو صورتين ثابتتين إلى فيديو سينمائي متصل باستخدام MiniMax H3:
🖼️ الصورة الأولى تحدد بداية الفيديو.
🖼️ الصورة الثانية تحدد نهاية الفيديو.
📝 البرومبت يصف الحركة والأحداث والانتقال المطلوب بينهما.
ويقوم النموذج بإنشاء كل الفريمات والحركات الموجودة بين الصورتين تلقائيًا، مع محاولة الوصول إلى تكوين الصورة الأخيرة في نهاية الفيديو.
الفكرة ببساطة:
First Frame + Last Frame + Prompt = Complete Video
━━━━━━━━━━━━━━━━━━━━
🎯 ما وظيفة الورك فلو؟
يمكن استخدام الورك فلو لإنشاء:
انتقالات سينمائية بين مكانين مختلفين.
Zoom من الفضاء إلى موقع على الأرض.
تحولات زمنية أو بيئية.
انتقال من الطبيعة إلى الدمار أو العكس.
انتقال بين لقطتين لشخصية واحدة.
تغيّر تعبيرات الوجه أو وضعية الجسم.
تحريك الكاميرا بين مشهدين.
مشاهد قصصية متصلة من البداية إلى النهاية.
فيديوهات إعلانية وسينمائية قصيرة.
فيديو يصل إلى نهاية محددة بدل ترك النهاية عشوائية.
يستخدم النموذج الصورتين كـ Boundary Frames، أي حدّين بصريين يحددان نقطة البداية ونقطة النهاية.
━━━━━━━━━━━━━━━━━━━━
🖼️ دور الصورة الأولى — First Frame
الصورة الأولى هي نقطة البداية الدقيقة للفيديو، وتحدد:
تكوين أول لقطة.
مكان الكاميرا واتجاهها.
الشخصيات الظاهرة في البداية.
الخلفية والموقع.
الإضاءة والألوان.
الملابس والهوية البصرية.
العناصر الموجودة داخل المشهد.
وضعية الجسم وتعبير الوجه.
نقطة انطلاق الحركة.
يجب أن يبدأ الفيديو قريبًا جدًا من تكوين الصورة الأولى، ثم يبدأ التحرك منها بصورة طبيعية.
ارفعها داخل النود المسماة:
اللقطة الأولى — First Frame
━━━━━━━━━━━━━━━━━━━━
🖼️ دور الصورة الأخيرة — Last Frame
الصورة الثانية تحدد الوجهة البصرية النهائية للفيديو، ومنها:
التكوين النهائي.
مكان الشخصية في نهاية المشهد.
ملامح الوجه والملابس.
وضعية الجسم واليدين.
موقع الكاميرا النهائي.
الخلفية النهائية.
الإضاءة والألوان.
الحالة العاطفية.
العناصر التي يجب أن تظهر في النهاية.
يجب أن يتطور الفيديو تدريجيًا حتى يصل إلى تكوين قريب من الصورة الثانية عند آخر فريم.
ارفعها داخل النود المسماة:
اللقطة الأخيرة — Last Frame
💡 الصورة الأخيرة لا تعني أن النموذج سيكررها حرفيًا بشكل مضمون، لكنها تعمل كمرساة قوية توجه الحركة والتكوين النهائي.
━━━━━━━━━━━━━━━━━━━━
📝 دور البرومبت
البرومبت يشرح للنموذج كيفية الانتقال من الصورة الأولى إلى الصورة الأخيرة.
يُفضل أن يحتوي على:
تعريف واضح لدور كل صورة.
ما يحدث بعد بداية الفيديو.
حركة الشخصية أو العناصر.
مسار الكاميرا.
الأحداث الوسطية.
طريقة ظهور عناصر الصورة الأخيرة.
التوقيت التقريبي لكل مرحلة.
المؤثرات الصوتية المطلوبة.
العناصر التي يجب الحفاظ عليها.
القيود التي تمنع تغير الهوية أو الملابس.
شكل اللقطة النهائية.
يحتوي الملف الحالي على برومبت تجريبي يبدأ من كوكب الأرض في الفضاء، ثم تتحرك الكاميرا إلى الطبيعة، وبعد ذلك تمر تدريجيًا بمناطق مدمرة، حتى تدخل منزلًا متضررًا وتصل إلى الطفلة الموجودة في الصورة الأخيرة.
يمكن استبدال هذا البرومبت بالكامل ليناسب صورك، مع الحفاظ على فكرة:
Image1 = Exact First-Frame Anchor
Image2 = Exact Final-Frame Anchor
━━━━━━━━━━━━━━━━━━━━
🔊 الصوت في الورك فلو
هذا الإصدار لا يحتاج إلى رفع ملف صوتي خارجي.
لا توجد نود:
Load Audio
يقوم MiniMax H3 بتوليد الصوت المتزامن مع الفيديو اعتمادًا على وصف المشهد، مثل:
أصوات البيئة.
الرياح.
الطيور.
حركة السيارات.
الانفجارات أو الدمار.
خطوات الشخصيات.
أصوات الأشياء المتحركة.
المؤثرات السينمائية.
أجواء المكان.
يتم توليد الصوت داخل المسار نفسه، ثم فك ترميزه باستخدام:
MiniMax H3 Audio VAE FP32
إذا كنت لا تريد كلامًا أو غناءً، اكتب ذلك بوضوح داخل البرومبت:
No dialogue, no singing, no speech, no whispering, and no generated vocals. Generate only synchronized environmental ambience and physical sound effects.
━━━━━━━━━━━━━━━━━━━━
🧠 كيف يعمل الورك فلو؟
1️⃣ يتم رفع صورة البداية داخل:
اللقطة الأولى — First Frame
2️⃣ يتم رفع صورة النهاية داخل:
اللقطة الأخيرة — Last Frame
3️⃣ يتم كتابة وصف الحركة والانتقال داخل نود:
MiniMaxH3ImageToVideo
4️⃣ يتم تحديد مدة الفيديو من:
المدة الزمنية — Duration
5️⃣ يتم اختيار نسبة العرض والدقة من:
حجم الفيديو — Aspect Ratio
6️⃣ يحسب الورك فلو عدد الفريمات اعتمادًا على المدة ومعدل:
24 FPS
7️⃣ يتم تعديل عدد الفريمات تلقائيًا ليتوافق مع البنية الزمنية التي يحتاجها MiniMax H3.
8️⃣ يتم تحليل البرومبت والصورتين باستخدام:
Qwen3‑VL 32B MiniMax H3
9️⃣ يستخدم الورك فلو نموذج التوليد:
MiniMax H3 FL2VA BF16
10️⃣ تتم معالجة الفيديو والصوت باستخدام:
MiniMax H3 Video VAE FP16
MiniMax H3 Audio VAE FP32
1️⃣1️⃣ يتم تطبيق تقنيات تقليل استهلاك الذاكرة وتسريع المعالجة.
1️⃣2️⃣ يبدأ الـSampling لإنشاء الحركة والفريمات الوسيطة والصوت المتزامن.
1️⃣3️⃣ يتم فك ترميز الفيديو والصوت بشكل منفصل.
1️⃣4️⃣ يتم دمجهما داخل ملف MP4 نهائي باستخدام:
Video Combine
━━━━━━━━━━━━━━━━━━━━
⚡ تقنيات التسريع الموجودة
يحتوي الورك فلو على:
MiniMax H3 Memory Efficient Sage Attention
Sage Attention Patch
MiniMax H3 Block Cache T8
معالجة الكاش على CPU
نموذج Qwen3‑VL NVFP4 AWQ
مسار MiniMax H3 FL2VA BF16
تساعد هذه الإعدادات على تقليل استهلاك الذاكرة وتسريع التوليد مقارنة بتشغيل المسار الكامل من دون تحسينات.
⚠️ لا يحتوي هذا الملف على:
Turbo LoRA.
EasyCache.
SpectrumApply.
RIFE Frame Interpolation.
ملف صوتي خارجي.
إذا ظهرت أخطاء توافق أو تغيّرت الحركة، يمكن تجربة تعطيل:
MiniMax H3 Block Cache
أو
Memory Efficient Sage Attention Patch
لكن لا تعدّل هذه المنطقة إلا إذا كنت تعرف تأثيرها على الأداء والذاكرة.
━━━━━━━━━━━━━━━━━━━━
⚙️ الإعدادات الحالية داخل الملف
Video Duration: 15 Seconds
Base Frame Rate: 24 FPS
Aspect Ratio: 16:9
Current Megapixels: 0.3 MP
Current Output Size: 736×416
Size Multiple: 32
Generated Frames for 15 Seconds: 374 Frames
Model: MiniMax H3 FL2VA BF16
Text/Vision Encoder: Qwen3‑VL 32B NVFP4 AWQ
Video VAE: MiniMax H3 Video VAE FP16
Audio VAE: MiniMax H3 Audio VAE FP32
Sampler: RES Multistep
Scheduler: Simple
Sampling Steps: 20
Denoise: 1.0
Seed Mode: Randomize
Output Frame Rate: 24 FPS
Output Format: H.264 MP4
Final Generated Audio: Enabled
Save Output: Enabled
━━━━━━━━━━━━━━━━━━━━
⏱️ حساب عدد الفريمات
يعتمد الورك فلو على المدة المحددة ومعدل 24 FPS.
عند اختيار 15 ثانية:
15 × 24 = 360 فريمًا
بعد ذلك يعدّل الورك فلو العدد تلقائيًا ليتوافق مع البنية الزمنية لنموذج MiniMax H3، فيصبح:
374 فريمًا
ويتم ذلك باستخدام معادلة داخل نود:
ComfyMathExpression
لذلك قد يكون طول الفيديو الناتج أكبر قليلًا من المدة الاسمية المحددة.
الملف النهائي يعمل عند 24 FPS، ولا تتم زيادة المعدل إلى 60 FPS لأن هذا الورك فلو لا يحتوي على RIFE.
━━━━━━━━━━━━━━━━━━━━
📐 الدقات المتاحة بنسبة 16:9
Megapixels الدقة
0.2 MP 608×352
0.3 MP 736×416
0.4 MP 864×480
0.5 MP 960×544
0.6 MP 1056×608
0.7 MP 1152×640
0.8 MP 1216×672
0.9 MP 1280×736
0.98 MP 1344×768
1.0 MP 1376×768
1.2 MP 1504×832
1.5 MP 1664×928
1.8 MP 1824×1024
2.0 MP 1920×1088
💡 ابدأ الاختبار على 0.2 أو 0.3 MP، وبعد التأكد من نجاح الحركة والوصول إلى الفريم الأخير استخدم 0.98 أو 1.0 MP للنتيجة النهائية.
رفع الدقة يزيد وقت التوليد واستهلاك الذاكرة بصورة كبيرة.
━━━━━━━━━━━━━━━━━━━━
☁️ طريقة تشغيل الورك فلو
1️⃣ افتح الورك فلو داخل ComfyUI أو RunningHub.
2️⃣ ارفع صورة البداية داخل:
اللقطة الأولى — First Frame
3️⃣ ارفع صورة النهاية داخل:
اللقطة الأخيرة — Last Frame
4️⃣ عدّل البرومبت الموجود داخل:
MiniMaxH3ImageToVideo
5️⃣ اشرح بالتسلسل ما يجب أن يحدث بين الصورتين.
6️⃣ حدد مدة الفيديو من:
المدة الزمنية — Duration
7️⃣ اختر نسبة العرض والدقة من:
حجم الفيديو — Aspect Ratio
8️⃣ ابدأ بدقة 0.2 أو 0.3 MP للاختبار.
9️⃣ اضغط:
Queue Prompt
🔟 بعد اكتمال التوليد، سيقوم الورك فلو بفك ترميز الفيديو والصوت ودمجهما تلقائيًا داخل ملف MP4.
━━━━━━━━━━━━━━━━━━━━
📌 نصائح للحصول على أفضل نتيجة
استخدم صورتين لهما نسبة عرض متقاربة.
اختر صورتين واضحتين وعاليتَي الجودة.
اجعل الانتقال بينهما منطقيًا وقابلًا للتنفيذ.
اشرح حركة الكاميرا بصورة زمنية ومتسلسلة.
لا تطلب عددًا كبيرًا من الأحداث خلال 15 ثانية.
لا تُظهر عناصر الفريم الأخير مبكرًا إذا كنت تريد مفاجأة بصرية.
حافظ على هوية الشخصية وملابسها داخل البرومبت.
وضّح متى تبدأ الشخصية في الظهور.
حدد شكل اللقطة الأخيرة بدقة.
استخدم حركة واحدة متصلة إذا كنت تريد انتقالًا سلسًا.
تجنب القطع المفاجئ إذا كان الهدف رحلة كاميرا مستمرة.
لا تطلب انتقالًا مستحيلًا من دون شرح بصري.
ابدأ بدقة منخفضة قبل التوليد النهائي.
جرّب أكثر من Seed لأن النتيجة قد تختلف بين كل تشغيل.
لمنع ظهور الفريم الأخير مبكرًا، استخدم:
Do not reveal Image2, its characters, location, or final composition before the camera physically reaches the final scene.
للحفاظ على الهوية:
Strict identity consistency—preserve the exact face, age, hairstyle, clothing, body proportions, colors, and accessories shown in Image2.
للحصول على نهاية قوية:
The final moments must settle naturally into the exact composition, camera angle, subject placement, lighting, and emotional state defined by Image2.
━━━━━━━━━━━━━━━━━━━━
⚠️ ملاحظات مهمة
الصورة الأولى تحدد بداية الفيديو فقط.
الصورة الثانية تحدد نهاية الفيديو المطلوبة.
البرومبت هو المسؤول عن بناء الرحلة بينهما.
لا يحتاج الورك فلو إلى ملف صوتي خارجي.
يقوم النموذج بتوليد المؤثرات الصوتية من وصف المشهد.
النتيجة النهائية تعمل عند 24 FPS.
لا توجد معالجة RIFE داخل هذا الإصدار.
إعداد 15 ثانية يولد 374 فريمًا داخل الملف الحالي.
الدقة الحالية هي 736×416.
الـSeed مضبوط على Randomize، لذلك تختلف النتيجة مع كل تشغيل.
الوصول إلى الفريم الأخير قد لا يكون مطابقًا 100% في كل محاولة.
لا تعدّل المنطقة المسماة:
إعدادات ملناش دعوة بيها
إلا إذا كنت تعرف تأثير النودز على النموذج والـVAE والـSampling والذاكرة.
🇬🇧 MiniMax H3 Workflow — First and Last Frame to a 15-Second Video
This workflow transforms two still images into one continuous cinematic video using MiniMax H3:
🖼️ The first image defines the beginning of the video.
🖼️ The second image defines the intended ending.
📝 The prompt describes the movement, events, camera path, and transition between them.
MiniMax H3 generates the motion and all intermediate frames while attempting to arrive naturally at the composition defined by the final image.
The concept is simple:
First Frame + Last Frame + Prompt = Complete Video
━━━━━━━━━━━━━━━━━━━━
🎯 Workflow Purpose
The workflow can generate:
Cinematic transitions between two locations.
Space-to-Earth zoom sequences.
Environmental or time-based transformations.
Transitions from natural beauty to destruction.
Character pose or expression changes.
Continuous camera journeys.
Short visual narratives.
Advertising and cinematic videos.
Videos with a visually guided ending rather than a random final shot.
The two images operate as boundary frames, defining the visual starting point and intended destination.
━━━━━━━━━━━━━━━━━━━━
🖼️ First Frame Role
The first image defines:
The opening composition.
Initial camera position and orientation.
Starting characters and objects.
Environment and background.
Lighting and colors.
Clothing and visual identity.
Body pose and facial expression.
The physical starting point of the action.
Upload it through:
اللقطة الأولى — First Frame
The video should begin close to this image before moving naturally away from it.
━━━━━━━━━━━━━━━━━━━━
🖼️ Last Frame Role
The second image defines the intended final state, including:
Final composition.
Character identity and appearance.
Final body and hand positions.
Clothing and accessories.
Camera position.
Background and location.
Lighting and colors.
Emotional state.
Objects visible at the end.
Upload it through:
اللقطة الأخيرة — Last Frame
The video should gradually evolve toward this image during its final moments.
The last frame is a strong visual anchor, although exact pixel-level reproduction is not guaranteed in every generation.
━━━━━━━━━━━━━━━━━━━━
📝 Prompt Role
The prompt explains how the video travels from the first frame to the last frame.
A strong prompt should define:
The role of each reference image.
Opening action.
Subject motion.
Camera movement.
Intermediate events.
Environmental progression.
When the final character or location appears.
Approximate timing.
Required sound effects.
Identity-preservation instructions.
Elements that must not change.
The exact intended ending.
The included example begins with Earth in deep space, descends toward a beautiful natural landscape, moves through progressively damaged areas, and eventually enters a destroyed home to reach the girl shown in the final image.
You can completely replace this prompt while preserving the structure:
Image1 = Exact First-Frame Anchor
Image2 = Exact Final-Frame Anchor
━━━━━━━━━━━━━━━━━━━━
🔊 Generated Audio
This workflow does not require an external audio file.
It does not contain a:
Load Audio
input node.
MiniMax H3 generates synchronized audio from the scene description, including:
Environmental ambience.
Wind and weather.
Birds and animals.
Vehicles.
Footsteps.
Moving objects.
Destruction and impact sounds.
Physical cinematic sound effects.
Location-specific atmosphere.
The generated audio is decoded through:
MiniMax H3 Audio VAE FP32
If dialogue, singing, or generated voices are not wanted, reinforce:
No dialogue, no singing, no speech, no whispering, and no generated vocals. Generate only synchronized environmental ambience and physical sound effects.
━━━━━━━━━━━━━━━━━━━━
🧠 How the Workflow Operates
1️⃣ Upload the starting image through:
اللقطة الأولى — First Frame
2️⃣ Upload the ending image through:
اللقطة الأخيرة — Last Frame
3️⃣ Enter the transition and action description inside:
MiniMaxH3ImageToVideo
4️⃣ Set the video duration through:
المدة الزمنية — Duration
5️⃣ Select the aspect ratio and resolution through:
حجم الفيديو — Aspect Ratio
6️⃣ The workflow calculates the initial frame count at 24 FPS.
7️⃣ It automatically adjusts that count to match MiniMax H3’s temporal structure.
8️⃣ The prompt and reference images are processed using:
Qwen3‑VL 32B MiniMax H3
9️⃣ Video and audio generation use:
MiniMax H3 FL2VA BF16
🔟 The visual and audio representations are processed through:
MiniMax H3 Video VAE FP16
MiniMax H3 Audio VAE FP32
1️⃣1️⃣ Memory-efficient attention and block caching are applied.
1️⃣2️⃣ Sampling generates the motion, intermediate frames, and synchronized audio.
1️⃣3️⃣ Video and audio are decoded separately.
1️⃣4️⃣ Both are combined and exported through:
Video Combine
━━━━━━━━━━━━━━━━━━━━
⚡ Acceleration Components
The workflow contains:
MiniMax H3 Memory Efficient Sage Attention
Sage Attention Patch
MiniMax H3 Block Cache T8
CPU-based block-cache processing.
Qwen3‑VL NVFP4 AWQ
MiniMax H3 FL2VA BF16
These components reduce memory usage and improve generation speed.
This file does not contain:
Turbo LoRA.
EasyCache.
SpectrumApply.
RIFE Frame Interpolation.
External audio input.
If compatibility errors or major motion changes appear, test the workflow with Block Cache or Sage Attention disabled.
━━━━━━━━━━━━━━━━━━━━
⚙️ Current File Settings
Video Duration: 15 Seconds
Base Frame Rate: 24 FPS
Aspect Ratio: 16:9
Current Megapixels: 0.3 MP
Current Output Size: 736×416
Size Multiple: 32
Generated Frames: 374
Model: MiniMax H3 FL2VA BF16
Text/Vision Encoder: Qwen3‑VL 32B NVFP4 AWQ
Video VAE: MiniMax H3 Video VAE FP16
Audio VAE: MiniMax H3 Audio VAE FP32
Sampler: RES Multistep
Scheduler: Simple
Sampling Steps: 20
Denoise: 1.0
Seed Mode: Randomize
Output Frame Rate: 24 FPS
Output Format: H.264 MP4
Generated Audio: Enabled
Save Output: Enabled
━━━━━━━━━━━━━━━━━━━━
⏱️ Frame Calculation
For a 15-second video at 24 FPS:
15 × 24 = 360 frames
The workflow then adjusts the sequence to satisfy MiniMax H3’s temporal structure:
Adjusted frame count = 374 frames
This calculation is performed automatically through:
ComfyMathExpression
The final output remains at 24 FPS. It is not interpolated to 60 FPS because this version does not contain RIFE.
━━━━━━━━━━━━━━━━━━━━
📐 Available 16:9 Resolutions
Megapixels Output Resolution
0.2 MP 608×352
0.3 MP 736×416
0.4 MP 864×480
0.5 MP 960×544
0.6 MP 1056×608
0.7 MP 1152×640
0.8 MP 1216×672
0.9 MP 1280×736
0.98 MP 1344×768
1.0 MP 1376×768
1.2 MP 1504×832
1.5 MP 1664×928
1.8 MP 1824×1024
2.0 MP 1920×1088
Begin testing at 0.2 or 0.3 MP. After confirming motion quality and successful arrival at the last frame, use 0.98 or 1.0 MP for the final generation.
━━━━━━━━━━━━━━━━━━━━
☁️ How to Run the Workflow
1️⃣ Open the workflow in ComfyUI or RunningHub.
2️⃣ Upload the starting image through:
اللقطة الأولى — First Frame
3️⃣ Upload the ending image through:
اللقطة الأخيرة — Last Frame
4️⃣ Edit the prompt inside:
MiniMaxH3ImageToVideo
5️⃣ Describe the complete chronological transition between the images.
6️⃣ Set the duration through:
المدة الزمنية — Duration
7️⃣ Select the aspect ratio and resolution through:
حجم الفيديو — Aspect Ratio
8️⃣ Begin at 0.2 or 0.3 MP for testing.
9️⃣ Press:
Queue Prompt
🔟 The workflow generates and decodes the video and audio, combines them, and saves the final result as an MP4 file.
━━━━━━━━━━━━━━━━━━━━
📌 Best-Result Recommendations
Use two clear, high-quality images.
Keep their aspect ratios reasonably similar.
Create a physically understandable transition.
Describe camera movement chronologically.
Avoid excessive events within 15 seconds.
Do not reveal the final frame’s elements too early.
Explicitly preserve character identity and clothing.
State exactly when the final subject should appear.
Describe the final composition clearly.
Use one continuous camera movement for seamless journeys.
Avoid abrupt cuts if continuity is required.
Test at low resolution before final generation.
Try multiple seeds because every run may produce a different result.
To prevent early appearance of the final scene:
Do not reveal Image2, its characters, location, or final composition before the camera physically reaches the final scene.
For identity consistency:
Strict identity consistency—preserve the exact face, age, hairstyle, clothing, body proportions, colors, and accessories shown in Image2.
For a stronger ending:
The final moments must settle naturally into the exact composition, camera angle, subject placement, lighting, and emotional state defined by Image2.
━━━━━━━━━━━━━━━━━━━━
⚠️ Important Notes
Image1 defines the video’s starting frame.
Image2 defines the intended final frame.
The prompt constructs the complete journey between them.
No external audio file is required.
MiniMax H3 generates synchronized environmental audio.
The final output runs at 24 FPS.
This version does not use RIFE interpolation.
A 15-second setting produces 374 frames.
The current resolution is 736×416.
The seed is randomized, so results vary between runs.
Exact last-frame reproduction is not guaranteed every time.
Avoid modifying the group labeled:
إعدادات ملناش دعوة بيها
unless you understand its effect on the model, VAE, sampling, cache, and memory usage.