النص أو الصورة
ابدأ بلقطة مختصرة أو استخدم رابطًا مرئيًا معتمدًا للموضوع والتكوين.
Native 30-second multimodal production
Generate continuous single-shot video up to 30 seconds in one pass, with up to 50 multimodal references, native 4K, and region-level editing — no stitching, no sequence breaks.
مساحة عمل المسار الدقيق
الضوابط الحالية والأرصدة المشحونة تأتي من المولد المباشر. قم بتأكيدها قبل التقديم.
مدخلات PixMind الحالية
ابدأ بلقطة مختصرة أو استخدم رابطًا مرئيًا معتمدًا للموضوع والتكوين.
حدد كلا نقطتي النهاية، ثم قم بوصف الحركة الجسدية ومسار الكاميرا بينهما فقط.
قم بتعيين الهوية والإعداد والحركة والصوت لفصل مراجع الصور والفيديو والصوت.
Long-form multimodal route
Seedance 2.5 extends the 2.0 multimodal architecture: longer native duration, the highest reference limit among commercial models, and region-level editing that preserves overall consistency.
01
Direct a continuous 30-second take in one generation instead of stitching multiple 5–15s clips.

02
Combine up to 50 image, video, text, and audio inputs to control characters, scenes, style, and sound together.

03
Change a product, prop, or background area while the rest of the clip keeps its motion, lighting, and identity.

04
Output crisp 4K natively for large displays, product films, and high-DPI social.

05
Keep identity, eyelines, and lip-sync stable while directing a 30-second performance with linked room sound.

06
Test opening beats as controlled variants while preserving the same product, palette, and final packshot.

مقارنة الطريق
يفصل الجدول لقطة مسار PixMind الحالية عن المطالبات على مستوى عائلة ByteDance. يجب إعادة فحص الاعتمادات وعناصر التحكم المكشوفة في المولد.
| الإصدار | الأفضل ل | المدخلات | المدة | القرار | الصوت | لقطة الائتمان |
|---|---|---|---|---|---|---|
| Seedance 2.0 Pro | مراجع معقدة متعددة الوسائط ولقطات نهائية مختارة | المراجع النصية والصورة والفيديو والصوت | 5 أو 10 أو 15 ثانية | 480p، 720p، 1080p؛ يتطلب 4K التحقق المباشر | طريق قادر على الصوت | اقرأ من المولد المباشر |
| Seedance 2.5الحالي | Native 30-second single-shot video with up to 50 multimodal references | Text, image, video, and audio references (up to 50) | Up to 30 seconds in a single shot | 480p, 720p, 1080p, native 4K | طريق قادر على الصوت | اقرأ من المولد المباشر |
| Seedance 2.0 Fast | إيقاعات القصة المصورة، واختبارات أ/ب للكاميرا، والخطافات الاجتماعية | المراجع النصية والصورة والفيديو والصوت | 5 أو 10 أو 15 ثانية | 480p أو 720p | طريق قادر على الصوت | اقرأ من المولد المباشر |
| Seedance 2.0 Mini | اختبار الملاءمة المرجعية والتكوين والتعارضات السريعة | المراجع النصية والصورة والفيديو والصوت | 5 أو 10 أو 15 ثانية | 480p أو 720p | طريق قادر على الصوت | اقرأ من المولد المباشر |
| Seedance 1.5 Pro | المطالبات الحالية التي تقودها الصور واختبارات الترحيل الخاضعة للرقابة | نص أو ما يصل إلى صورتين في الاستجابة الحالية | المولد المباشر نهائي؛ تعارض حقول واجهة برمجة التطبيقات | 480p أو 720p أو 1080p | طريق قادر على الصوت | لقطة واجهة برمجة التطبيقات (API) من 40 نقطة؛ تأكيد العيش |
تشير الاستجابة 1.5 Pro إلى 40 نقطة، لكن حقول المدة الخاصة بها تتعارض. لا تعرض Pro وFast وMini حقل نقاط ثابتة في لقطة نقطة النهاية التي تم التحقق منها.
ما يجب مشاهدته بعد الاختبار
استخدم موضوعًا كاملاً وحواف نظيفة وعمقًا قابلاً للقراءة ومساحة كافية للحركة المطلوبة.
افصل حركة الهدف عن حركة الكاميرا وتجنب تجميع العديد من التغييرات الكبيرة في لقطة قصيرة واحدة.
انتقل إلى 2.0 Pro عندما تحتاج اللقطة إلى مراجع متعددة الوسائط، أو الإطارات الأولى/الأخيرة، أو اتجاه صوتي أكثر وضوحًا.
كيفية استخدام هذا الإصدار
Give every image, video, text, or audio input a single explicit role (identity, shape, motion, palette, or rhythm).
Fix identity, wardrobe, props, screen direction, and the ending composition before adding complex motion.
Edit one region per pass and review identity, geometry, lighting, and timing in the full clip each time.
Read the live generator for the current credits at the chosen duration and resolution before submitting.
المطالبات الخاصة بالإصدار

موجه 01
Use the full native duration for one continuous reveal.
One continuous 30-second take. Use @Image 1 for exact product shape and @Image 2 for the studio palette. Slow controlled orbit around the product, one subtle environmental interaction at the midpoint, finish on the approved hero composition. Preserve shape, material, ports, and color. No generated text.

موجه 02
Coordinate identity, motion, and sound across references.
Keep the character from @Image 1 unchanged. Use @Video 1 only for the body motion and @Audio 1 for the ambient rhythm. The character walks through the station, pauses, and looks toward the exit. Preserve facial features, wardrobe, and left-to-right screen direction.

موجه 03
Change one area without re-generating the shot.
Keep the full clip motion, lighting, and identity fixed. Only replace the background product on the right shelf with the object in @Image 2, matching perspective, scale, and reflections. Do not alter the subject or camera path.
Pro
Quality-first multimodal video with reference and audio control.
عرض الإصدارPro
Native 30-second single-shot video with up to 50 multimodal references, 4K output, and region-level editing.
Fast
مسار Seedance 2.0 الأسرع لسير العمل الاجتماعي والحملات والتصور المسبق كثيف التكرار.
عرض الإصدارMini
خيار Seedance 2.0 أخف لاختبار الأفكار متعددة الوسائط قبل الالتزام بتقديم الجودة أولاً.
عرض الإصدارPro
خيار تحويل صورة إلى فيديو أثبت كفاءته مع نسب عرض إلى ارتفاع واسعة وتحكم في الدقة للحصول على لقطات واضحة ومرجعية.
عرض الإصدارNative single-shot video up to 30 seconds, up to 50 multimodal references (image, video, text, audio), region-level editing that preserves consistency, and native 4K output.
2.0 generates 5–15s clips with up to 9 references; 2.5 extends single-shot duration to 30 seconds, raises the reference limit to 50, and adds region-level editing.
Pass up to 50 inputs across image, video, text, and audio in one request. Give each a single explicit role and remove assets that compete for the same property.
Yes. Region-level editing lets you change a specific area while the rest of the clip keeps its motion, lighting, and identity. Edit one region per pass and review the full clip each time.
Single-shot clips up to 30 seconds and resolutions up to native 4K. Confirm the exact options in the live generator before submitting.
Pricing is not yet published. Read the live generator for the current credits at your chosen duration and resolution before submitting.
Yes. The unified audio-video architecture generates synchronized audio in the same task when the selected mode supports it. Review speech, lip movement, source matching, distortion, and timing before publishing.
Review identity, faces, hands, product geometry, text, logos, continuity, dialogue, and audio timing frame by frame across the full clip. Confirm rights for every uploaded person, brand, voice, image, video, and audio asset.
خطوة الإنتاج التالية
Return to the family hub for route selection, current parameter notes, and migration guidance.