In February 2026, ByteDance dropped Seedance 2.0. Within 24 hours it was everywhere, because a filmmaker generated a Hollywood-quality fight scene with a two-line prompt. The Deadpool screenwriter saw it and publicly posted that it was likely over for screenwriters. Elon Musk replied. The Motion Picture Association issued a formal statement the same day. ByteDance pulled the ability to generate real celebrity likenesses, tightened the guardrails, and the internet held its breath wondering if the model had been gutted in the process.
It wasn't. The underlying capability, the action, the physics, the motion, the emotional performance, all survived intact. And honestly, those conversations need to happen. Generative AI is moving faster than any legal framework can keep up with, and how the industry handles that matters. But that is a different video. What we can tell you is that Seedance 2.0 is the most capable video model on the market right now, and it is live on OpenArt.
Ermin Monzon, our Head of Content Media & Education, broke down why the model matters technically. Our in-house creator Bob stress-tested every reference type he could think of. And our collaborator Mia Meow built a 10-tip workflow guide from weeks of testing. This Handbook pulls the strongest material from all three to give you one complete picture.
Lo que Seedance 2.0 puede hacer → Ermin's overview
Before we get into workflow tips, it is worth understanding why this model is generating the reaction it is. Ermin walked through a series of demos that show where Seedance 2.0 pulls away from everything else on the market.
The fight scene that broke the internet is the obvious headline, but the demo that surprised Ermin more was a dancer with neon light trails sweeping around them in orbital arcs. It looks like a Nike commercial or a music video with a serious production budget. The reason this is technically significant is that the light trails are not composited over the dancer. The model generated both simultaneously. The motion and the effect interact with each other: the trails bend around the dancer's body, and the lighting from the trails actually reflects on the floor beneath them. That is called coherent VFX generation. The model understands that light behaves physically. It does not just paint glowing lines on top of video. For anyone doing music video content, UGC ads, or anything with a visual effects component, this is where Seedance 2.0 stands apart.
Another demo worth watching: a skier launching off a mountain with an alpine range filling the background. The physics of the snow spray, the body rotation, the landing. All generated, all physically plausible. The model handles action and motion at a level that no other publicly available tool is matching right now.
Text-to-video: the single-take prompt structure → 0:50 in Mia's video
Everything you hear in Mia's text-to-video demos, the audio, the ambient sound, was generated at the same time as the video. One prompt, one take. Other models struggle to hold this many details in a single generation. Seedance 2.0 actually follows what you give it.
Mia's prompt structure for single-take shots is straightforward. Start with "continuous single take" and your camera actions. Then describe what the camera sees as it moves. End with control and style keywords: no cuts, seamless transition, cinematic, high definition. That formula alone will get you strong results.
Para escenas más complejas, como un fight sequence at 2:13, Mia recommends starting with a clear beginning and end state. Describe the fight scene, describe how it ends (everyone on the ground, for example), and let the model fill the middle. She also found that describing specific camera angles creates a more cinematic feel, and adding a film genre reference to the end of the prompt helps set the visual tone. One thing she discovered through iteration: if you describe your characters physically in the prompt rather than leaving it vague, the model does a better job designing the scene and keeping characters visually distinct.
Mia's key insight on prompting: the model is well-trained on film language. Terms like "tracking shot," "crane shot," "whip pan," and genre cues like "gritty war film" or "neo-noir thriller" go a long way.

El tutorial completo de Mia recoge sus 10 trucos de flujo de trabajo, incluidos los que hemos destilado a lo largo de este manual. Míralo mientras avanzas o vuelve a él más tarde.
The reference system: tagging images, video, and audio into your prompt → 4:27 en el vídeo de Mia · → Análisis a fondo de las referencias de Bob
This is the core mechanic that powers everything else in Seedance 2.0 on OpenArt, and once you understand it, every section that follows will make more sense.
La idea es sencilla: puedes subir imágenes, vídeos y archivos de audio como referencias y luego etiquetarlos directamente en tu prompt de texto con el símbolo @. Escribe @ y elige qué archivo subido quieres referenciar. A partir de ahí, le dices al modelo exactamente qué hacer con cada uno. Usa la cara de la imagen uno. Usa el movimiento de cámara del vídeo uno. Sincroniza los labios con el audio uno. El modelo lee tu prompt, mira los archivos etiquetados y combina todo en una sola generación.
Mia's tutorial multimodal en el minuto 4:27 is the cleanest explainer of how this works. She uploaded a dance video, two character images, and a song, then told the model to use the video as a camera motion reference, image one as the left dancer, image two as the right dancer, and the audio as the background music. All tagged with @, all in one prompt.
Bob pushed this further. In one of his demos, he referenced four different files in a single prompt: the face from image one, the overalls from image three, the shirt from image four, and a location from a reference video. The prompt was: the man in image one is walking down the street from the video wearing the overalls from image three and carrying the shirt from image four, talking to the camera. The model held consistency across every element. That is a remarkable amount of compositing happening from a single text prompt.

A few things both creators learned about how references behave:
Si tu vídeo de referencia tiene audio y además etiquetas un archivo de audio aparte, the model tends to favor the video's existing audio over your uploaded track. Mia's workaround: upload the video without audio if you want your own music or voiceover to take priority.
Las referencias no se limitan a elementos concretos. You can also point to a video and say "I just want it to look like this" without copying anything else. Bob used this to transfer a visual style (fisheye lens, flickering light) from one video to an entirely new scene.
You do not have to use all reference types at once. A single character image in a text prompt is a perfectly valid use of the system. The power scales with how many elements you layer in, but it works fine at every level of complexity.

Lip sync, voice, and performance → 7:57 in Mia's video · → Análisis a fondo del audio de Bob
El vídeo completo de Bob cubre en profundidad las referencias en imágenes, audio y vídeo. Las secciones de lip sync y clonación de voz son donde realmente brilla.
This is the section with the most tips from the most creators, and for good reason. Seedance 2.0's audio capabilities are the feature that separates it from everything else on the market right now.
Bob's transcript trick
This is the single most useful tip across all three videos. When you are doing lip sync with an audio reference, include the actual transcript of the words in your prompt alongside the audio file. Here is why.
Bob tested this with a clip of himself talking to camera. In his first attempt, he tagged the audio file and described the scene but did not include the exact words being spoken. The model listened to the audio and tried to replicate it, but got details wrong ("CGI 2.0" instead of "Seedance 2.0"). When he modified the prompt to explicitly spell out the words being said, the audio file combined with the written transcript produced dramatically better results. The lip sync was tighter, the words were accurate, and the overall performance was more convincing.

La clave: siempre que hagas lip sync, proporciona la transcripción real junto con el archivo de audio. Prácticamente tienes garantizado un mejor resultado.
Ajusta la duración de tu audio a la de tu vídeo
Mia lo aprendió por las malas. Si tu referencia de audio dura 11 segundos y ajustas el vídeo a 15 segundos, el modelo se ve obligado a estirar o comprimir el audio para que encaje, y es casi seguro que perderás la sincronización. Ajusta siempre la duración a la longitud real de tu archivo de audio.
Voice cloning
Bob demonstrated voice cloning with a 15-second clip of Ermin's voice (Uncle Monz around the OpenArt office). The setup: record about 15 seconds of a voice in a way that captures the qualities you want duplicated. Tag that audio as a voice reference rather than a lip sync source. Then write the actual dialogue in the prompt, and the model generates new speech in that voice.
El primer intento fue demasiado rápido para la cadencia natural de Ermin. Bob le dio al modelo 14 segundos completos en lugar de 9 y el resultado sonaba muchísimo más parecido a cómo hablaría Ermin de verdad. La lección: el ritmo importa tanto como la muestra de voz. Dale al modelo suficiente tiempo para pronunciar las palabras a la velocidad adecuada para esa voz.
Puedes reutilizar la misma referencia de voz en varios vídeos para mantener la consistencia vocal de un personaje.
Mia's performance directing tips → 12:20 in Mia's video
Mia found that you can direct the emotional performance and energy of a character through the prompt itself. The words you write influence how the character delivers the dialogue, not just what they say. If you write "he says it excitedly," the character's body language and vocal energy shift. If you write "she whispers nervously," the whole performance changes.
Esto también se aplica a varios personajes en la misma escena. En 11:10 in Mia's video, she demonstrates giving two characters different vocal personalities and emotional states in one prompt, and the model keeps them distinct.
Mia's negative guardrail tip
Cuando quieres transiciones limpias de una escena a otra, sin morphing entre ellas (por ejemplo, narrando a través de cuatro imágenes distintas), Mia aprendió a añadir restricciones negativas al prompt: «sin morphing, sin ghosting, sin cortes de cámara». Sin esto, el modelo a veces mezcla las escenas de formas que parecen fallos visuales. Con estas restricciones, obtienes transiciones nítidas de una escena a la siguiente.

Advanced references: motion, storyboards, and style → 13:22 in Mia's video · → Referencias de vídeo de Bob
Una vez que entiendas cómo funciona el sistema de referencias para caras y voces, aquí tienes qué más puedes usar como referencia.
Motion and camera transfer
If you have a shot where you like the way the camera moves, you can transfer that camera motion to an entirely new scene. Bob demonstrated this with a video of a dog jumping up and down as the camera dollied around it. He brought in that video as a reference and said "use that video as a reference for the camera and character movement to guide a scarecrow jumping on a trampoline." The result transferred both the camera dolly and the jumping action to the new character and scene.
He also showed motion transfer with real footage. Using a clip of Emily running at 19:15, dijo «reemplaza a la mujer del vídeo uno por el hombre de la imagen uno y que esté nevando». El modelo cambió el personaje, añadió clima invernal y mantuvo intacto el movimiento de carrera. Algunos detalles del fondo cambiaron (la atmósfera general pasó a ser gris e invernal), pero el movimiento principal se mantuvo.
Character Swap Split Screen LRB.mp4
Del storyboard al vídeo → 13:22 in Mia's video
Puedes subir una cuadrícula estilo viñetas de cómic o un storyboard y el modelo intentará animar viñeta a viñeta. Mia lo probó con una cuadrícula dibujada a mano y comprobó que el modelo seguía las viñetas en orden, pero la lógica direccional entre viñetas no siempre le resultaba evidente al modelo solo con las imágenes.
Two tips that made this work better. First, describe the narrative logic between panels in your prompt. The model needs to understand the story progression, not just the visual sequence. Second, if your storyboard has text or captions on it, add "text overlay no captions on screen" to the prompt to prevent text from bleeding into the generated video.
Mia también señala que puedes usar páginas reales de cómic para esto, pero ten cuidado con cualquier cosa que implique propiedad intelectual existente.
Referencias de estilo y replicación de plantillas → 14:45 en el vídeo de Mia
Mia's template replication tip is a practical one for anyone producing content at volume. If you have a video with a visual style you like, you can reference it and tell the model to replicate just the style, not the content. She tested this by pointing to a video and saying she only wanted the fisheye lens effect and the flickering double-exposure look, then combined it with character images and outfit references to create a fashion sequence. The model picked up the abstract visual qualities without copying the specific content of the reference video.
Esto es útil para mantener una identidad visual coherente en una serie de vídeos. Crea un vídeo que te guste y luego usa su estilo como referencia para todas las generaciones siguientes.
Post-generation: extend and edit → 16:06 in Mia's video
Un vídeo no está terminado una vez generado. Seedance 2.0 te permite ampliarlo, editarlo y reescribirlo después.

Extend. Sube un vídeo que quieras continuar, elige extender hacia delante (o hacia atrás), indica cuántos segundos, ajusta la duración para que coincida y describe qué quieres que pase en la parte nueva. Mia extendió un vídeo 10 segundos hacia delante, luego lo extendió 10 segundos hacia atrás y los combinó en una escena fluida más larga que el límite de generación de 15 segundos. Eso es un desbloqueo real. El límite de 15 segundos solía ser un techo infranqueable. Ahora puedes seguir construyendo escenas fluidas más largas sin cortes.
Editar. Si quieres cambiar algo que ya ocurrió en un vídeo generado, puedes hablar con el modelo sobre el vídeo existente y decirle qué modificar. Funciona de forma parecida a la edición de vídeo a vídeo de Wan 2.7, pero con la mayor comprensión de Seedance 2.0 sobre lo que hay en la escena.
The bottom line
Seedance 2.0 is the most capable video model available on OpenArt right now, and the reference system is the reason to pay attention. The ability to tag images, video, and audio into a single prompt and have the model hold consistency across all of them is not something any other model is doing at this level. The lip sync is strong. The voice cloning works. The motion transfer is clean. The coherent VFX generation, where lighting and effects actually interact with the scene physically, is new territory entirely.
Tres creadores, tres vídeos, una misma conclusión: este modelo escucha. Sigue prompts complejos y con múltiples elementos mejor que cualquier otro. Lo mejor que puedes hacer es lanzarte y empezar a pedirle lo que quieres, porque tiene una imaginación increíble. Palabras de Bob, y tiene toda la razón.