Geração de vídeo de nível profissional com áudio nativo, referenciamento multimodal profundo e edição contínua — destilado em 18 exemplos de prompts anotados.
O Dreamina Seedance 2.0 é um modelo de geração de vídeo de nível profissional que suporta nativamente a saída conjunta de áudio e vídeo, com excelente compreensão semântica e interação multimodal. Este guia percorre as técnicas principais — fórmulas de texto, controle de referência, renderização de texto, referências de imagem/vídeo e edição não destrutiva — com 18 exemplos oficiais.
Todos os vídeos e imagens abaixo são gerados autonomamente pelos modelos visuais Seedance/Seedream. Reutilizados com permissão da BytePlus ModelArk.
01
Princípios gerais
1.1Fórmula básica para instruções de texto
O Seedance 2.0 se destaca em seguir a lógica da linguagem natural. Combine estes elementos de forma flexível para corresponder à sua intenção criativa:
Sujeito + ação — a base lógica. Defina claramente quem está realizando qual ação.
Atmosfera — defina o tom geral descrevendo o fundo espacial, detalhes de iluminação ou um estilo visual específico.
Design de som — instruções avançadas podem incluir efeitos sonoros de cena ou ambiente para uma saída audiovisual imersiva e sincronizada.
1.2Controle de referência para entradas multimodais
Além das descrições de texto, bloqueie o estado ideal do quadro com materiais de referência. O Seedance 2.0 suporta referenciamento profundo de imagens, áudio e vídeo.
No seu prompt, especifique claramente o objeto de referência — ex: “Use the composition of Image 1” ou “Match the motion of Video 2”.
O modelo extrai características principais da referência e as mescla com seu texto, mantendo alta fidelidade e previsibilidade, permitindo ainda variação criativa.
02
Renderização de texto
O Seedance 2.0 gera texto legível em T2V (Text-to-Video), I2V (Image-to-Video), R2V (Reference-to-Video) e V2V (Video-to-Video). Ele adapta o estilo e a cor da fonte à sua cena automaticamente e oferece controle granular sobre estilo, tempo e layout.
2.1Slogans
O Seedance 2.0 detecta automaticamente o contexto da cena para corresponder à estética de fonte mais apropriada. Para consistência de marca rigorosa, combine com uma referência de logo (veja 3.2).
Display subtitles at the bottom-center with the text. The subtitles must be perfectly synchronized with the audio rhythm and pacing.
2.3Balões de fala
Prompt template
[Character] says, "[Dialogue]." Speech bubbles appear around the character containing the spoken text.
Melhores práticas
Vocabulário comum. Palavras padrão e amplamente reconhecidas são renderizadas com mais precisão.
Evite palavras obscuras. Termos muito técnicos ou raros podem produzir glifos inconsistentes.
Minimize símbolos especiais. Pontuação complexa ou não padrão prejudica a fidelidade da fonte.
2.1 · Example 1
Resultado
Imagem de referência
Image 1
Prompt
Hand-drawn comic style: Three people are sitting around a table enjoying the fried chicken shown in Image 1, with a friendly and joyful atmosphere. The frame then gradually blurs, and the text "Bite", "Laugh", "Seedance" in order appears in the center of the screen.
2.2 · Example 2Locução
Resultado
Imagem de referência
Image 1
Prompt
I2V: A time-lapse of a mountain landscape transitioning from a vast, starry night to a vibrant dawn. Voiceover: A deep, serene male voice says: 'In the vast silence of the cosmos, our world is but a fleeting moment. Yet, within it, life defiantly thrives.' > Text Integration: Render the narration as subtitles at the bottom-center. Subtitles must be perfectly synchronized with audio timing.
2.2 · Example 3Dublagem
Resultado
Imagem de referência
Image 1
Prompt
R2V: A shot of these two people in Image 1 chatting in a modern office. The woman speaks first with a playful tone: "You always arrive right on time, don't you just love that perfect timing?" followed by the man's smiling reply: "I have my own rhythm." > Text Integration: Render the dialogue as subtitles at the bottom-center of the screen. Subtitles should appear sequentially as each character speaks.
2.3 · Example 4Cena de parquinho
Resultado
Imagem de referência
Image 1
Prompt
The two characters from Image 1, both dressed in sportswear, are running on the school playground. The girl looks at the boy, smiling confidently as she says: "We can definitely do it!". Cut to a close-up of the boy. He hesitates and replies: "Are you sure?". Cut back to a medium close-up of the girl. She speaks in a light, upbeat tone: "Yes!" Her demeanor is bright and resolute. Speech bubbles containing the corresponding lines appear around the speaking character.
2.3 · Example 5Campo de maçãs
Resultado
Imagem de referência
Image 1
Prompt
Refer to the character design of the girl in Image 1 and Image 2. The scene is set in an apple field: the girl picks one apple, takes a bite, smiles and says "This is the real deal!". A speech bubble pops up beside the girl, with this line written inside.
03
Referência de imagem
O Seedance 2.0 suporta referências de múltiplas perspectivas para sujeitos e referências multi-imagem para layouts de cena e sequências. Se o seu fluxo de trabalho exigir uma ordem específica, carregue as imagens em sequência e referencie-as no seu prompt como Image 1, Image 2, … Image N.
3.1Referência de sujeito multi-perspectiva
Prompt template
Refer to / Extract / Combine / Use the [Subject] from [Image N] to generate [Scene Description], maintaining consistent [Subject] features.
3.2Referência multi-imagem
Prompt template
Refer to / Extract / Combine / Follow the [Description of referenced elements] from [Image N] to generate [Scene Description], while maintaining the consistency of [Referenced Elements].
3.1 · Example 1Eletrônicos de consumo
Resultado
Imagem de referência
Image 1
Prompt
Use the cameras featured in Image 1, Image 2 and Image 3. Replace the original background with a white one, and place the cameras on a white table. The shooting lens first focuses on the cameras in close-up, then slowly rotates 360° with the cameras as the main subject, clearly displaying the front, sides and back of each camera.
3.1 · Example 2Casa e estilo de vida
Resultado
Imagem de referência
Image 1
Prompt
In a warm-toned home setting, present the thermos shown in the reference image in a medium shot. Then smoothly push the camera into a close-up of the thermos. Next, a hand naturally enters the frame off-screen, gently grips the thermos body and picks it up. The camera follows the slight rotating motion of the hand to showcase the thermos.
3.1 · Example 3Personagens
Resultado
Imagem de referência
Image 1
Prompt
Refer to the image of the woman in Image 1, Image 2 and Image 3, and generate a scene of her eating a cake in a coffee shop.
3.2 · Example 4Referência de logo
Resultado
Imagem de referência
Image 1
Prompt
The scene is set on an aerial corridor in a neon-drenched futuristic metropolis, where flying vehicles and holographic ads intertwine. Featuring the girl from Reference Image 2, the sequence opens with a medium shot of her releasing a silver floating lantern embedded with a holographic projection. The camera then pulls back to reveal floating lanterns flooding the sky, which gradually converge at the center of the frame to form the logo from Reference Image 1. The entire piece adopts a 3D cyberpunk sci-fi animation style.
3.2 · Example 5Referência de múltiplos sujeitos
Resultado
Imagem de referência
Image 1
Prompt
Using the cat and dog from the reference Image 1 and Image 2 as prototypes, the scene unfolds in a cozy apartment. The dog is lying on the ground eating dog food when the cat approaches, extending a paw to nudge the dog. The dog pauses its meal upon noticing the cat, and the cat snuggles up next to the dog. The entire scene features a warm colored tone.
3.2 · Example 6Referência de múltiplos elementos
Resultado
Imagem de referência
Image 1
Prompt
The scene is set in the restaurant from Image 4 with people coming and going. The girl from Image 1, wearing the clothes from Image 2, is organizing the items on the counter. The boy, a customer, from Image 3 approaches her to ask for her contact information. The logo from Image 5 remains in the bottom right corner throughout.
3.2 · Example 7Referência de sequência multi-painel
Resultado
Imagem de referência
Image 1
Prompt
Refer to the sequence in Image 1 to create an intense high-energy fight sequence. All frame compositions from Image 1 shall be presented in strict predefined order, after which the two characters engage in fierce, fast-paced combat.
3.2 · Example 8Referência de sequência
Resultado
Imagem de referência
Image 1
Prompt
Refer to the composition in Image 3. A girl (her character design refers to Image 1) is waiting for her father to finish cooking, and she says: "아빠, 배고파요! 밥 다 됐어요?" Then the camera pans right and cuts to the frame and composition shown in Image 4. The father (his character design refers to Image 2) replies to her: "거의 다 됐어, 조금만 기다려!" Next, the camera cuts back to a close-up shot of the daughter's slightly disappointed facial expression, and she says: "아직 멀었어요? 맛있는 냄새 나는데..." Then the shot switches to a close-up of the father's face, and he says: "이제 진짜 금방이야. '빨리빨리' 하지 말고 손부터 씻고 와!"
04
Referência de vídeo
O Seedance 2.0 suporta referenciamento baseado em vídeo para movimento, movimento de câmera e efeitos visuais. Carregue os vídeos em sequência e referencie-os como Video 1, Video 2, … Video N.
4.1Referência de movimento
Prompt template
Refer to the [Motion Description] from [Video N] to generate [Scene Description], keeping the motion details consistent.
4.2Referência de movimento de câmera
Prompt template
Refer to the [Camera Movement Description] from [Video N] to generate [Scene Description], keeping the scene consistent.
4.3Referência de efeitos visuais (VFX)
Prompt template
Refer to the [VFX Effects Description] from [Video N] to generate [Scene Description], keeping the special effects consistent.
4.1 · Example 1Artístico
Resultado
References
Image 1
Video 1
Prompt
Refer to the character movements and shot language in Video 1 to create a fight scene with the character from Image 2 on the left and the character from Image 1 on the right. Include intense background music.
4.1 · Example 2Marketing
Resultado
Vídeo de referência
Video 1
Prompt
Referencing the running shape of the horse in the video, generate a scene: a golden steed runs on the grassland, then freezes its magnificent running posture and turns into a horse-shaped gold pendant.
4.2 · Example 3
Resultado
References
Image 1
Video 1
Prompt
Referring to the camera movement in Video 1, create a concept video for a science and technology park, with the tall building in Image 1 as the visual center, also using a first-person diving perspective, to reflect the sense of technology in the park from Image 1.
4.3 · Example 4Produção de vídeo
Resultado
References
Image 1
Video 1
Prompt
Refer to the golden particle effects in Video 1, so that when the character in Image 1 plays the flute, the same particle effects surround their body.
4.3 · Example 5FX criativo
Resultado
References
Image 1
Video 1
Prompt
Refer to the special effects shown in Video 1 to generate identical wings for the girl in Image 1, ensuring the wing formation trajectory follows the exact same motion path and sequence depicted in the video.
05
Edição de vídeo
O Seedance 2.0 suporta edição de vídeo não destrutiva — adicionando, removendo ou modificando elementos; estendendo para frente ou para trás; e completando trilhas em vários clipes. Os segmentos originais são preservados para uma continuidade perfeita.
5.1Adicionar, remover ou modificar elementos
Prompt template
Adding: At [Timestamp] and [Spatial Location] of [Video N], add [intended element].
Removing: Remove [Element] from [Video N], keeping the rest unchanged.
Modifying: Replace [original element] in [Video N] with [intended element].
5.2Estender vídeos
Prompt template
Extend [Video N] forward/backward + [Description of extended content]
Generate content before/after [Video N] + [Description of extended content]
5.3Completar trilhas
Prompt template
[Video 1] + [Transition Description] + followed by [Video 2] + [Transition Description] + followed by [Video 3]
Note: A conclusão de trilha suporta até 3 clipes de vídeo com uma duração combinada de 15 segundos. O modelo corta automaticamente os segmentos de conexão para uma síntese contínua.
5.1 · Example 1Adicionar elementos
Resultado
Vídeo de referência
Video 1
Prompt
Add snacks such as fried chicken and pizza to the countertop in Video 1.
5.1 · Example 2Remover elementos
Resultado
Vídeo de referência
Video 1
Prompt
Remove everything that isn't office stuff from the table in Video 1, keeping the rest of the video content unchanged.
5.1 · Example 3Modificar elementos
Resultado
References
Image 1
Video 1
Prompt
Replace the perfume featured in Video 1 with the face cream from Image 1, with all original motions and camera work preserved.
5.2 · Example 4Estender para frente
Resultado
Vídeo de referência
Video 1
Prompt
Generate the content after Video 1: the two men who are late run towards them, the five people finally meet and have a friendly chat.
5.2 · Example 5Estender para trás
Resultado
Vídeo de referência
Video 1
Prompt
Extend the opening segment of Video 1: Set up an over-the-shoulder shot of the man in a hoodie, and the man says: "It's not that bad. You're just stressed. Everyone goes through this, you just need to keep going."
5.3 · Example 6
Resultado
Vídeos de referência
Video 1
Video 2
Prompt
Video 1. The moment a leaf falls to the ground, it sets off a special effect of golden particles. A gust of wind blows by, leading into Video 2.
Pronto para criar com o Seedance 2.0?
Comece a gerar vídeos com áudio integrado, controle de referência e edição não destrutiva — diretamente no Doitong.