Write each Seedance 2.0 Mini prompt as one compact shot brief: name one concrete subject, one visible action, a simple camera intention, a setting or visual treatment, and-only if it helps-one small proof detail. Keep that core brief fixed when revising, then change one meaningful element at a time.
Seedance 2.0 Mini is described in Dreamina as an AI video model for generating video from text prompts or images. The most useful prompt is therefore not a complete video concept or a brand manifesto. It is a clear description of what the viewer should see in one moment. Dreamina's creator guide describes the text-to-video and image-to-video workflow.
Start with One Visible Shot
Before writing, reduce the idea to a single observable event. "Make a premium coffee campaign" is a creative direction, not yet a shot. A usable shot might be: a barista pours espresso into a white cup at a café counter.
Ask three questions:
- 1
- Who or what is on screen? Name a concrete subject. 2
- What happens once? Choose one dominant action the viewer can observe. 3
- What should the camera emphasize? Decide what deserves attention: the person, product, gesture, or result.
If the answer includes a sequence of events, multiple locations, or several competing actions, split it into separate shots. A simpler brief gives each generation one job.
Use the Compact Shot Formula
A practical writing pattern is:
Concrete subject + one visible action and object + framing and one movement + setting or visual treatment + optional proof detail or mood
This is not required model syntax. It is a reliable order for turning a creative intention into compact natural language. The underlying structure-subject, action, camera, setting, and tone-is supported by Seedance 2.0 Mini prompt guidance.
Here is a complete example:
A barista in a dark-green apron pours espresso into a white cup, medium close-up with a subtle push-in, warm café light, crema swirling on the surface.
Each phrase has a purpose:
- 1
- Subject: "A barista in a dark-green apron" identifies the performer without overloading the description. 2
- Action: "Pours espresso into a white cup" gives the shot a visible event and object. 3
- Camera cue: "Medium close-up with a subtle push-in" suggests what to prioritize. 4
- Setting and look: "Warm café light" establishes context and visual treatment. 5
- Proof detail: "Crema swirling on the surface" gives the action a small, visible result.
Leave out decorative details that do not change the shot. More words are not automatically more direction.
Direct the Camera and Look Without Competing Instructions
Camera language works best as a restrained cue, not a long list of cinematic commands. Choose one framing and, if needed, one movement.
For example:
- 1
- Product detail: "Medium close-up, tripod-stable" 2
- Reveal: "Wide shot, gentle push-in" 3
- Character gesture: "Handheld close-up, slight move forward"
Then add one short visual cue-such as warm morning light, cool dawn tones, or an upbeat mood-after the core shot. Treat framing, movement, pacing, and mood as descriptive intentions rather than exact controls.
A proof detail is optional. Use one when it makes the action easier to understand: steam rising from a cup, condensation on a bottle, or fabric texture visible on a hand. Do not stack several tiny requirements into the same shot.
Fix Weak Results by Simplifying the Brief
When a generation feels vague or drifts from the intended shot, revise the clearest source of ambiguity instead of rewriting everything.
For instance, this prompt is broad:
Make a premium coffee advertisement.
This version gives the shot a clearer job:
A barista twists the grinder dial, then pours espresso into a white cup, medium close-up, tripod-stable, warm café light.
The goal is not to make the prompt more elaborate. It is to make the intended screen action easier to describe and evaluate.
Iterate One Variable at a Time
Treat the first generation as a baseline, not a final verdict. Keep the subject, action, and setting unchanged, then adjust one element for the next version-perhaps the framing, camera movement, proof detail, or mood cue. This creates a cleaner comparison, even though a single prompt change does not prove it caused every difference in the output.
Review each candidate against the original shot brief:
- 1
- Is the subject recognizable and appropriately framed? 2
- Is the primary action readable? 3
- Does the clip show visual artifacts, unnatural motion, or deviations from the prompt? 4
- Is the footage usable for the role it needs to play in the edit?
When a recognizable person, product, or visual anchor matters, a reference input may be useful rather than adding layers of identity description. For image-led work, keep the requested motion restrained and assess the resulting clip rather than assuming every detail will carry through.
Choose one simple shot now, write it with the formula, and generate a baseline. Revise only one element at a time, then bring the strongest usable clips into CapCut to sequence, pace, caption, sound-design, and shape them into a finished video.