Generación de video de nivel profesional con audio nativo, referencia multimodal profunda y edición fluida, resumida en 18 ejemplos de prompts anotados.
Dreamina Seedance 2.0 es un modelo de generación de video de nivel profesional que admite de forma nativa la salida conjunta de audio y video, con una comprensión semántica y una interacción multimodal excepcionales. Esta guía recorre las técnicas principales (fórmulas de texto, control de referencia, renderizado de texto, referencias de imagen/video y edición no destructiva) con 18 ejemplos oficiales.
Todos los videos e imágenes a continuación han sido generados de forma autónoma por los modelos visuales Seedance/Seedream. Reutilizados con permiso de BytePlus ModelArk.
01
Principios generales
1.1Fórmula básica para instrucciones de texto
Seedance 2.0 destaca por seguir la lógica del lenguaje natural. Combina estos elementos con flexibilidad para ajustarlos a tu intención creativa:
Sujeto + acción — la base lógica. Define claramente quién realiza qué acción.
Atmósfera — establece el tono general describiendo el fondo espacial, detalles de iluminación o un estilo visual específico.
Diseño de sonido — las instrucciones avanzadas pueden incluir efectos de sonido ambientales o de escena para una salida audiovisual inmersiva y sincronizada.
1.2Control de referencia para entradas multimodales
Más allá de las descripciones de texto, fija el estado ideal del fotograma con materiales de referencia. Seedance 2.0 admite referencias profundas de imágenes, audio y video.
En tu prompt, especifica claramente el objeto de referencia; por ejemplo, “Use the composition of Image 1” o “Match the motion of Video 2”.
El modelo extrae las características principales de la referencia y las fusiona con tu texto, manteniendo una alta fidelidad y previsibilidad, permitiendo al mismo tiempo variaciones creativas.
02
Renderizado de texto
Seedance 2.0 genera texto legible en T2V (Text-to-Video), I2V (Image-to-Video), R2V (Reference-to-Video) y V2V (Video-to-Video). Adapta automáticamente el estilo y el color de la fuente a tu escena, y te ofrece un control granular sobre el estilo, el tiempo y el diseño.
2.1Eslóganes
Seedance 2.0 detecta automáticamente el contexto de la escena para aplicar la estética de fuente más apropiada. Para una consistencia de marca estricta, combínalo con una referencia de logotipo (ver 3.2).
Display subtitles at the bottom-center with the text. The subtitles must be perfectly synchronized with the audio rhythm and pacing.
2.3Bocadillos de texto
Prompt template
[Character] says, "[Dialogue]." Speech bubbles appear around the character containing the spoken text.
Mejores prácticas
Vocabulario común. Las palabras estándar y ampliamente reconocidas se renderizan con mayor precisión.
Evitar palabras poco comunes. Los términos demasiado técnicos o de diccionario pueden producir glifos inconsistentes.
Minimizar símbolos especiales. La puntuación compleja o no estándar perjudica la fidelidad de la fuente.
2.1 · Example 1
Resultado
Imagen de referencia
Image 1
Prompt
Hand-drawn comic style: Three people are sitting around a table enjoying the fried chicken shown in Image 1, with a friendly and joyful atmosphere. The frame then gradually blurs, and the text "Bite", "Laugh", "Seedance" in order appears in the center of the screen.
2.2 · Example 2Locución
Resultado
Imagen de referencia
Image 1
Prompt
I2V: A time-lapse of a mountain landscape transitioning from a vast, starry night to a vibrant dawn. Voiceover: A deep, serene male voice says: 'In the vast silence of the cosmos, our world is but a fleeting moment. Yet, within it, life defiantly thrives.' > Text Integration: Render the narration as subtitles at the bottom-center. Subtitles must be perfectly synchronized with audio timing.
2.2 · Example 3Doblaje
Resultado
Imagen de referencia
Image 1
Prompt
R2V: A shot of these two people in Image 1 chatting in a modern office. The woman speaks first with a playful tone: "You always arrive right on time, don't you just love that perfect timing?" followed by the man's smiling reply: "I have my own rhythm." > Text Integration: Render the dialogue as subtitles at the bottom-center of the screen. Subtitles should appear sequentially as each character speaks.
2.3 · Example 4Escena de patio de recreo
Resultado
Imagen de referencia
Image 1
Prompt
The two characters from Image 1, both dressed in sportswear, are running on the school playground. The girl looks at the boy, smiling confidently as she says: "We can definitely do it!". Cut to a close-up of the boy. He hesitates and replies: "Are you sure?". Cut back to a medium close-up of the girl. She speaks in a light, upbeat tone: "Yes!" Her demeanor is bright and resolute. Speech bubbles containing the corresponding lines appear around the speaking character.
2.3 · Example 5Campo de manzanos
Resultado
Imagen de referencia
Image 1
Prompt
Refer to the character design of the girl in Image 1 and Image 2. The scene is set in an apple field: the girl picks one apple, takes a bite, smiles and says "This is the real deal!". A speech bubble pops up beside the girl, with this line written inside.
03
Referencia de imagen
Seedance 2.0 admite referencias de múltiples perspectivas para sujetos y referencias de múltiples imágenes para diseños de escena y secuencias. Si tu flujo de trabajo requiere un orden específico, sube las imágenes en secuencia y nómbralas en tu prompt como Image 1, Image 2, … Image N.
3.1Referencia de sujeto con múltiples perspectivas
Prompt template
Refer to / Extract / Combine / Use the [Subject] from [Image N] to generate [Scene Description], maintaining consistent [Subject] features.
3.2Referencia de múltiples imágenes
Prompt template
Refer to / Extract / Combine / Follow the [Description of referenced elements] from [Image N] to generate [Scene Description], while maintaining the consistency of [Referenced Elements].
3.1 · Example 1Electrónica de consumo
Resultado
Imagen de referencia
Image 1
Prompt
Use the cameras featured in Image 1, Image 2 and Image 3. Replace the original background with a white one, and place the cameras on a white table. The shooting lens first focuses on the cameras in close-up, then slowly rotates 360° with the cameras as the main subject, clearly displaying the front, sides and back of each camera.
3.1 · Example 2Hogar y estilo de vida
Resultado
Imagen de referencia
Image 1
Prompt
In a warm-toned home setting, present the thermos shown in the reference image in a medium shot. Then smoothly push the camera into a close-up of the thermos. Next, a hand naturally enters the frame off-screen, gently grips the thermos body and picks it up. The camera follows the slight rotating motion of the hand to showcase the thermos.
3.1 · Example 3Personajes
Resultado
Imagen de referencia
Image 1
Prompt
Refer to the image of the woman in Image 1, Image 2 and Image 3, and generate a scene of her eating a cake in a coffee shop.
3.2 · Example 4Referencia de logotipo
Resultado
Imagen de referencia
Image 1
Prompt
The scene is set on an aerial corridor in a neon-drenched futuristic metropolis, where flying vehicles and holographic ads intertwine. Featuring the girl from Reference Image 2, the sequence opens with a medium shot of her releasing a silver floating lantern embedded with a holographic projection. The camera then pulls back to reveal floating lanterns flooding the sky, which gradually converge at the center of the frame to form the logo from Reference Image 1. The entire piece adopts a 3D cyberpunk sci-fi animation style.
3.2 · Example 5Referencia de múltiples sujetos
Resultado
Imagen de referencia
Image 1
Prompt
Using the cat and dog from the reference Image 1 and Image 2 as prototypes, the scene unfolds in a cozy apartment. The dog is lying on the ground eating dog food when the cat approaches, extending a paw to nudge the dog. The dog pauses its meal upon noticing the cat, and the cat snuggles up next to the dog. The entire scene features a warm colored tone.
3.2 · Example 6Referencia de múltiples elementos
Resultado
Imagen de referencia
Image 1
Prompt
The scene is set in the restaurant from Image 4 with people coming and going. The girl from Image 1, wearing the clothes from Image 2, is organizing the items on the counter. The boy, a customer, from Image 3 approaches her to ask for her contact information. The logo from Image 5 remains in the bottom right corner throughout.
3.2 · Example 7Referencia de secuencia de múltiples paneles
Resultado
Imagen de referencia
Image 1
Prompt
Refer to the sequence in Image 1 to create an intense high-energy fight sequence. All frame compositions from Image 1 shall be presented in strict predefined order, after which the two characters engage in fierce, fast-paced combat.
3.2 · Example 8Referencia de secuencia
Resultado
Imagen de referencia
Image 1
Prompt
Refer to the composition in Image 3. A girl (her character design refers to Image 1) is waiting for her father to finish cooking, and she says: "아빠, 배고파요! 밥 다 됐어요?" Then the camera pans right and cuts to the frame and composition shown in Image 4. The father (his character design refers to Image 2) replies to her: "거의 다 됐어, 조금만 기다려!" Next, the camera cuts back to a close-up shot of the daughter's slightly disappointed facial expression, and she says: "아직 멀었어요? 맛있는 냄새 나는데..." Then the shot switches to a close-up of the father's face, and he says: "이제 진짜 금방이야. '빨리빨리' 하지 말고 손부터 씻고 와!"
04
Referencia de video
Seedance 2.0 admite referencias basadas en video para el movimiento, el desplazamiento de cámara y los efectos visuales. Sube los videos en secuencia y nómbralos como Video 1, Video 2, … Video N.
4.1Referencia de movimiento
Prompt template
Refer to the [Motion Description] from [Video N] to generate [Scene Description], keeping the motion details consistent.
4.2Referencia de movimiento de cámara
Prompt template
Refer to the [Camera Movement Description] from [Video N] to generate [Scene Description], keeping the scene consistent.
4.3Referencia de efectos visuales (VFX)
Prompt template
Refer to the [VFX Effects Description] from [Video N] to generate [Scene Description], keeping the special effects consistent.
4.1 · Example 1Artístico
Resultado
References
Image 1
Video 1
Prompt
Refer to the character movements and shot language in Video 1 to create a fight scene with the character from Image 2 on the left and the character from Image 1 on the right. Include intense background music.
4.1 · Example 2Marketing
Resultado
Video de referencia
Video 1
Prompt
Referencing the running shape of the horse in the video, generate a scene: a golden steed runs on the grassland, then freezes its magnificent running posture and turns into a horse-shaped gold pendant.
4.2 · Example 3
Resultado
References
Image 1
Video 1
Prompt
Referring to the camera movement in Video 1, create a concept video for a science and technology park, with the tall building in Image 1 as the visual center, also using a first-person diving perspective, to reflect the sense of technology in the park from Image 1.
4.3 · Example 4Producción de video
Resultado
References
Image 1
Video 1
Prompt
Refer to the golden particle effects in Video 1, so that when the character in Image 1 plays the flute, the same particle effects surround their body.
4.3 · Example 5Efectos creativos
Resultado
References
Image 1
Video 1
Prompt
Refer to the special effects shown in Video 1 to generate identical wings for the girl in Image 1, ensuring the wing formation trajectory follows the exact same motion path and sequence depicted in the video.
05
Edición de video
Seedance 2.0 admite la edición de video no destructiva: añadir, eliminar o modificar elementos; extender hacia adelante o hacia atrás; y completar pistas a través de múltiples clips. Los segmentos originales se preservan para una continuidad perfecta.
5.1Añadir, eliminar o modificar elementos
Prompt template
Adding: At [Timestamp] and [Spatial Location] of [Video N], add [intended element].
Removing: Remove [Element] from [Video N], keeping the rest unchanged.
Modifying: Replace [original element] in [Video N] with [intended element].
5.2Extender videos
Prompt template
Extend [Video N] forward/backward + [Description of extended content]
Generate content before/after [Video N] + [Description of extended content]
5.3Completar pistas
Prompt template
[Video 1] + [Transition Description] + followed by [Video 2] + [Transition Description] + followed by [Video 3]
Note: La función de completar pistas admite hasta 3 clips de video con una duración combinada de 15 segundos. El modelo recorta automáticamente los segmentos de conexión para una síntesis fluida.
5.1 · Example 1Añadir elementos
Resultado
Video de referencia
Video 1
Prompt
Add snacks such as fried chicken and pizza to the countertop in Video 1.
5.1 · Example 2Eliminar elementos
Resultado
Video de referencia
Video 1
Prompt
Remove everything that isn't office stuff from the table in Video 1, keeping the rest of the video content unchanged.
5.1 · Example 3Modificar elementos
Resultado
References
Image 1
Video 1
Prompt
Replace the perfume featured in Video 1 with the face cream from Image 1, with all original motions and camera work preserved.
5.2 · Example 4Extender hacia adelante
Resultado
Video de referencia
Video 1
Prompt
Generate the content after Video 1: the two men who are late run towards them, the five people finally meet and have a friendly chat.
5.2 · Example 5Extender hacia atrás
Resultado
Video de referencia
Video 1
Prompt
Extend the opening segment of Video 1: Set up an over-the-shoulder shot of the man in a hoodie, and the man says: "It's not that bad. You're just stressed. Everyone goes through this, you just need to keep going."
5.3 · Example 6
Resultado
Videos de referencia
Video 1
Video 2
Prompt
Video 1. The moment a leaf falls to the ground, it sets off a special effect of golden particles. A gust of wind blows by, leading into Video 2.
¿Listo para crear con Seedance 2.0?
Empieza a generar video con audio conjunto, control de referencia y edición no destructiva, directamente en Doitong.