Texto o imagen
Comience con un resumen de la toma o utilice un anclaje visual aprobado para el tema y la composición.
Native 30-second multimodal production
Generate continuous single-shot video up to 30 seconds in one pass, with up to 50 multimodal references, native 4K, and region-level editing — no stitching, no sequence breaks.
Espacio de trabajo de ruta exacta
Los controles actuales y los créditos cargados provienen del generador en vivo. Confírmalos antes de enviarlos.
Entradas actuales de PixMind
Comience con un resumen de la toma o utilice un anclaje visual aprobado para el tema y la composición.
Defina ambos puntos finales, luego describa solo el movimiento físico y la ruta de la cámara entre ellos.
Asigne identidad, entorno, movimiento y sonido para separar referencias de imagen, vídeo y audio.
Long-form multimodal route
Seedance 2.5 extends the 2.0 multimodal architecture: longer native duration, the highest reference limit among commercial models, and region-level editing that preserves overall consistency.
01
Direct a continuous 30-second take in one generation instead of stitching multiple 5–15s clips.

02
Combine up to 50 image, video, text, and audio inputs to control characters, scenes, style, and sound together.

03
Change a product, prop, or background area while the rest of the clip keeps its motion, lighting, and identity.

04
Output crisp 4K natively for large displays, product films, and high-DPI social.

05
Keep identity, eyelines, and lip-sync stable while directing a 30-second performance with linked room sound.

06
Test opening beats as controlled variants while preserving the same product, palette, and final packshot.

Comparación de rutas
La tabla separa la instantánea de la ruta actual de PixMind de los reclamos a nivel de familia de ByteDance. Se deben volver a verificar los créditos y controles expuestos en el generador.
| Versión | Lo mejor para | Entradas | Duración | Resolución | Audio | Instantánea de crédito |
|---|---|---|---|---|---|---|
| Seedance 2.0 Pro | Referencias multimodales complejas y tomas finales seleccionadas. | Referencias de texto, imagen, video y audio. | 5, 10 o 15 segundos | 480p, 720p, 1080p; 4K requiere verificación en vivo | Ruta con capacidad de audio | Leer del generador en vivo |
| Seedance 2.5Actual | Native 30-second single-shot video with up to 50 multimodal references | Text, image, video, and audio references (up to 50) | Up to 30 seconds in a single shot | 480p, 720p, 1080p, native 4K | Ruta con capacidad de audio | Leer del generador en vivo |
| Seedance 2.0 Fast | Ritmos de guiones gráficos, pruebas de cámara A/B y ganchos sociales | Referencias de texto, imagen, video y audio. | 5, 10 o 15 segundos | 480p o 720p | Ruta con capacidad de audio | Leer del generador en vivo |
| Seedance 2.0 Mini | Prueba de ajuste de referencia, composición y conflictos de indicaciones | Referencias de texto, imagen, video y audio. | 5, 10 o 15 segundos | 480p o 720p | Ruta con capacidad de audio | Leer del generador en vivo |
| Seedance 1.5 Pro | Avisos basados en imágenes existentes y pruebas de migración controlada | Texto o hasta dos imágenes en la respuesta actual | El generador en vivo es definitivo; Conflicto de campos API | 480p, 720p o 1080p | Ruta con capacidad de audio | Instantánea de API de 40 puntos; confirmar en vivo |
La respuesta 1.5 Pro informa 40 puntos, pero sus campos de duración entran en conflicto. Pro, Fast y Mini no exponen un campo de puntos estables en la instantánea del punto final verificado.
Qué mirar después de la prueba
Utilice un tema completo, bordes limpios, profundidad legible y suficiente espacio para el movimiento solicitado.
Separe la acción del sujeto del movimiento de la cámara y evite acumular varios cambios grandes en una toma corta.
Pase a 2.0 Pro cuando la toma necesite referencias multimodales, primeros/últimos fotogramas o una dirección de audio más explícita.
Cómo utilizar esta versión
Give every image, video, text, or audio input a single explicit role (identity, shape, motion, palette, or rhythm).
Fix identity, wardrobe, props, screen direction, and the ending composition before adding complex motion.
Edit one region per pass and review identity, geometry, lighting, and timing in the full clip each time.
Read the live generator for the current credits at the chosen duration and resolution before submitting.
Avisos específicos de la versión

rápido 01
Use the full native duration for one continuous reveal.
One continuous 30-second take. Use @Image 1 for exact product shape and @Image 2 for the studio palette. Slow controlled orbit around the product, one subtle environmental interaction at the midpoint, finish on the approved hero composition. Preserve shape, material, ports, and color. No generated text.

rápido 02
Coordinate identity, motion, and sound across references.
Keep the character from @Image 1 unchanged. Use @Video 1 only for the body motion and @Audio 1 for the ambient rhythm. The character walks through the station, pauses, and looks toward the exit. Preserve facial features, wardrobe, and left-to-right screen direction.

rápido 03
Change one area without re-generating the shot.
Keep the full clip motion, lighting, and identity fixed. Only replace the background product on the right shelf with the object in @Image 2, matching perspective, scale, and reflections. Do not alter the subject or camera path.
Pro
Quality-first multimodal video with reference and audio control.
Ver versiónPro
Native 30-second single-shot video with up to 50 multimodal references, 4K output, and region-level editing.
Fast
La ruta más rápida de Seedance 2.0 para flujos de trabajo de previsualización, campañas y redes sociales con muchas iteraciones.
Ver versiónMini
Una opción más ligera de Seedance 2.0 para probar ideas multimodales antes de comprometerse con un renderizado de calidad.
Ver versiónPro
Una opción comprobada de imagen a vídeo con amplias relaciones de aspecto y control de resolución para tomas sencillas basadas en referencias.
Ver versiónNative single-shot video up to 30 seconds, up to 50 multimodal references (image, video, text, audio), region-level editing that preserves consistency, and native 4K output.
2.0 generates 5–15s clips with up to 9 references; 2.5 extends single-shot duration to 30 seconds, raises the reference limit to 50, and adds region-level editing.
Pass up to 50 inputs across image, video, text, and audio in one request. Give each a single explicit role and remove assets that compete for the same property.
Yes. Region-level editing lets you change a specific area while the rest of the clip keeps its motion, lighting, and identity. Edit one region per pass and review the full clip each time.
Single-shot clips up to 30 seconds and resolutions up to native 4K. Confirm the exact options in the live generator before submitting.
Pricing is not yet published. Read the live generator for the current credits at your chosen duration and resolution before submitting.
Yes. The unified audio-video architecture generates synchronized audio in the same task when the selected mode supports it. Review speech, lip movement, source matching, distortion, and timing before publishing.
Review identity, faces, hands, product geometry, text, logos, continuity, dialogue, and audio timing frame by frame across the full clip. Confirm rights for every uploaded person, brand, voice, image, video, and audio asset.
Siguiente paso de producción
Return to the family hub for route selection, current parameter notes, and migration guidance.