Texto ou imagem
Comece com um resumo da cena ou use uma âncora visual aprovada para o assunto e a composição.
Native 30-second multimodal production
Generate continuous single-shot video up to 30 seconds in one pass, with up to 50 multimodal references, native 4K, and region-level editing — no stitching, no sequence breaks.
Espaço de trabalho de rota exata
Os controles atuais e os créditos cobrados vêm do gerador ao vivo. Confirme-os antes de enviar.
Entradas atuais do PixMind
Comece com um resumo da cena ou use uma âncora visual aprovada para o assunto e a composição.
Defina ambos os pontos finais e descreva apenas o movimento físico e o caminho da câmera entre eles.
Atribua identidade, cenário, movimento e som a referências separadas de imagem, vídeo e áudio.
Long-form multimodal route
Seedance 2.5 extends the 2.0 multimodal architecture: longer native duration, the highest reference limit among commercial models, and region-level editing that preserves overall consistency.
01
Direct a continuous 30-second take in one generation instead of stitching multiple 5–15s clips.

02
Combine up to 50 image, video, text, and audio inputs to control characters, scenes, style, and sound together.

03
Change a product, prop, or background area while the rest of the clip keeps its motion, lighting, and identity.

04
Output crisp 4K natively for large displays, product films, and high-DPI social.

05
Keep identity, eyelines, and lip-sync stable while directing a 30-second performance with linked room sound.

06
Test opening beats as controlled variants while preserving the same product, palette, and final packshot.

Comparação de rotas
A tabela separa o instantâneo da rota PixMind atual das reivindicações de nível de família ByteDance. Os créditos e controles expostos deverão ser verificados novamente no gerador.
| Versão | Melhor para | Entradas | Duração | Resolução | Áudio | Instantâneo de crédito |
|---|---|---|---|---|---|---|
| Seedance 2.0 Pro | Referências multimodais complexas e tomadas finais selecionadas | Referências de texto, imagem, vídeo e áudio | 5, 10 ou 15 segundos | 480p, 720p, 1080p; 4K requer verificação ao vivo | Rota com capacidade de áudio | Leia do gerador ao vivo |
| Seedance 2.5Atual | Native 30-second single-shot video with up to 50 multimodal references | Text, image, video, and audio references (up to 50) | Up to 30 seconds in a single shot | 480p, 720p, 1080p, native 4K | Rota com capacidade de áudio | Leia do gerador ao vivo |
| Seedance 2.0 Fast | Batidas de storyboard, testes A/B de câmera e ganchos sociais | Referências de texto, imagem, vídeo e áudio | 5, 10 ou 15 segundos | 480p ou 720p | Rota com capacidade de áudio | Leia do gerador ao vivo |
| Seedance 2.0 Mini | Testando ajuste de referência, composição e conflitos imediatos | Referências de texto, imagem, vídeo e áudio | 5, 10 ou 15 segundos | 480p ou 720p | Rota com capacidade de áudio | Leia do gerador ao vivo |
| Seedance 1.5 Pro | Prompts existentes baseados em imagens e testes de migração controlados | Texto ou até duas imagens na resposta atual | O gerador ao vivo é definitivo; Conflito de campos da API | 480p, 720p ou 1080p | Rota com capacidade de áudio | Instantâneo da API de 40 pontos; confirmar ao vivo |
A resposta do 1.5 Pro reporta 40 pontos, mas seus campos de duração são conflitantes. Pro, Fast e Mini não expõem um campo de pontos estáveis no snapshot do endpoint verificado.
O que assistir após o teste
Use um assunto completo, bordas limpas, profundidade legível e espaço suficiente para o movimento solicitado.
Separe a ação do assunto do movimento da câmera e evite empilhar várias mudanças grandes em uma única foto.
Mude para o 2.0 Pro quando a cena precisar de referências multimodais, primeiro/último quadro ou direção de áudio mais explícita.
Como usar esta versão
Give every image, video, text, or audio input a single explicit role (identity, shape, motion, palette, or rhythm).
Fix identity, wardrobe, props, screen direction, and the ending composition before adding complex motion.
Edit one region per pass and review identity, geometry, lighting, and timing in the full clip each time.
Read the live generator for the current credits at the chosen duration and resolution before submitting.
Prompts específicos da versão

Alerta 01
Use the full native duration for one continuous reveal.
One continuous 30-second take. Use @Image 1 for exact product shape and @Image 2 for the studio palette. Slow controlled orbit around the product, one subtle environmental interaction at the midpoint, finish on the approved hero composition. Preserve shape, material, ports, and color. No generated text.

Alerta 02
Coordinate identity, motion, and sound across references.
Keep the character from @Image 1 unchanged. Use @Video 1 only for the body motion and @Audio 1 for the ambient rhythm. The character walks through the station, pauses, and looks toward the exit. Preserve facial features, wardrobe, and left-to-right screen direction.

Alerta 03
Change one area without re-generating the shot.
Keep the full clip motion, lighting, and identity fixed. Only replace the background product on the right shelf with the object in @Image 2, matching perspective, scale, and reflections. Do not alter the subject or camera path.
Pro
Quality-first multimodal video with reference and audio control.
Ver versãoPro
Native 30-second single-shot video with up to 50 multimodal references, 4K output, and region-level editing.
Fast
A rota mais rápida do Seedance 2.0 para fluxos de trabalho sociais, de campanha e de pré-visualização com muitas iterações.
Ver versãoMini
Uma opção mais leve do Seedance 2.0 para testar ideias multimodais antes de se comprometer com uma renderização de qualidade.
Ver versãoPro
Uma opção comprovada de imagem para vídeo com amplas proporções e controle de resolução para fotos diretas com referência.
Ver versãoNative single-shot video up to 30 seconds, up to 50 multimodal references (image, video, text, audio), region-level editing that preserves consistency, and native 4K output.
2.0 generates 5–15s clips with up to 9 references; 2.5 extends single-shot duration to 30 seconds, raises the reference limit to 50, and adds region-level editing.
Pass up to 50 inputs across image, video, text, and audio in one request. Give each a single explicit role and remove assets that compete for the same property.
Yes. Region-level editing lets you change a specific area while the rest of the clip keeps its motion, lighting, and identity. Edit one region per pass and review the full clip each time.
Single-shot clips up to 30 seconds and resolutions up to native 4K. Confirm the exact options in the live generator before submitting.
Pricing is not yet published. Read the live generator for the current credits at your chosen duration and resolution before submitting.
Yes. The unified audio-video architecture generates synchronized audio in the same task when the selected mode supports it. Review speech, lip movement, source matching, distortion, and timing before publishing.
Review identity, faces, hands, product geometry, text, logos, continuity, dialogue, and audio timing frame by frame across the full clip. Confirm rights for every uploaded person, brand, voice, image, video, and audio asset.
Próxima etapa de produção
Return to the family hub for route selection, current parameter notes, and migration guidance.